Best Large Language Models (LLMs) Software

How Many Large Language Models (LLMs) Software Products Does G2 Track?

Total Products under this Category: 33

Category Stats (Sep 2026)

  • Average Rating: 4.36/5 (↓0.02 vs Aug 2026) The average rating of products in this category, based on all submitted ratings
  • Top Trending Product: Gemini (+0.34%) - Among all products in this category, Gemini recorded the largest rating increase compared to last month

Last updated: September 01, 2026

How Does G2 Rank Large Language Models (LLMs) Software Products?

Why You Can Trust G2's Software Rankings:

  • 30 Analysts and Data Experts
  • 4,300+ Authentic Reviews
  • 33+ Products
  • Unbiased Rankings

G2's software rankings are built on verified user reviews, rigorous moderation, and a consistent research methodology maintained by a team of analysts and data experts. Each product is measured using the same transparent criteria, with no paid placement or vendor influence. While reviews reflect real user experiences, which can be subjective, they offer valuable insight into how software performs in the hands of professionals. Together, these inputs power the G2 Score, a standardized way to compare tools within every category.

G2 Grid® for Large Language Models (LLMs) Software

G2 Grid® for Large Language Models (LLMs) Software plotting products by satisfaction and market presence

Highlighted products: ChatGPT, Claude, Gemini, Mistral AI, Deepseek, Grok, and Llama.

Underlying data: [Grid® JSON](https://www.g2.com/categories/large-language-models-llms/grids.json?focus%5B%5D=chatgpt&focus%5B%5D=claude-2025-12-11&focus%5B%5D=google-gemini&focus%5B%5D=mistral-ai&focus%5B%5D=deepseek&focus%5B%5D=xai-grok&focus%5B%5D=llama)

ChatGPT

ChatGPT is an advanced AI language model developed by OpenAI, designed to assist users in generating human-like text based on the input it receives. It serves as a versatile tool for a wide range of applications, including drafting emails, writing code, creating content, and providing detailed explanations on various topics. ChatGPT is continually evolving to enhance user experience and meet diverse needs. Key Features and Functionality: - Natural Language Understanding: ChatGPT can comprehend and generate text that closely resembles human conversation, making interactions intuitive and engaging. - Versatile Applications: It supports tasks such as content creation, coding assistance, learning new concepts, and more, catering to both personal and professional use cases. - Continuous Improvement: OpenAI regularly updates ChatGPT to improve its performance, accuracy, and safety, ensuring it remains a reliable tool for users. Primary Value and User Solutions: ChatGPT addresses the need for efficient and accessible assistance in various domains. By leveraging its advanced language processing capabilities, it helps users save time, enhance productivity, and access information seamlessly. Whether it's drafting documents, learning new subjects, or automating routine tasks, ChatGPT provides a valuable resource that adapts to individual requirements, making it an indispensable tool in today's digital landscape.

Average Rating: 4.6/5.0

Total Reviews: 2,938

How Do G2 Users Rate ChatGPT?

  • Quality of Support: 8.4/10 (Category avg: 7.9/10)
  • Content Moderation: 8.3/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.5/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.7/10 (Category avg: 8.0/10)

Who Is the Company Behind ChatGPT?

  • Seller: OpenAI
  • Year Founded: 2015
  • HQ Location: San Francisco, CA
  • Twitter: @OpenAI
    4,941,980 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    10,438 employees on LinkedIn®

Who Uses This Product?

  • Who Uses This: Student, Software Engineer
  • Top Industries: Information Technology and Services, Computer Software
  • Company Size: 54% Small, 28% Medium

What Do G2 Reviewers Say About ChatGPT?

AI-generated summary from verified user reviews

Pros
  • Users find ChatGPT’s ease of use invaluable, as it simplifies tasks and enhances daily productivity effortlessly.
  • Users appreciate the quick and tailored responses of ChatGPT, enjoying its versatility for various tasks and inquiries.
  • Users value the quick and helpful responses of ChatGPT, enhancing understanding and resolving queries efficiently.
  • Users value the time-saving capabilities of ChatGPT, enabling faster brainstorming and efficient design processes.
  • Users love ChatGPT for its time-saving capabilities, streamlining tasks like content creation and scheduling efficiently.
Cons
  • Users note that ChatGPT has accuracy limitations, often providing incorrect answers and misunderstanding questions with overconfidence.
  • Users find context understanding in ChatGPT lacking, leading to confusion and repetitive responses in conversations.
  • Users find the limited context window of ChatGPT frustrating, affecting its ability to process extensive information effectively.
  • Users note the inaccuracy of ChatGPT, often requiring validation of its responses for correctness and detail.
  • Users experience inaccurate responses from ChatGPT, leading to confusion and the need for clarification to achieve satisfactory answers.

What Are Recent G2 Reviews of ChatGPT?

What Are G2 Users Discussing About ChatGPT?

Claude

Claude is a state-of-the-art large language model (LLM) developed by Anthropic, designed to serve as a helpful, honest, and harmless AI assistant. With its advanced reasoning capabilities and conversational tone, Claude excels in tasks ranging from complex coding to in-depth financial analysis, making it a versatile tool for developers, enterprises, and financial professionals. Key Features and Functionality: - Advanced Coding Capabilities: Claude Opus 4 leads in coding performance, achieving top scores on benchmarks like SWE-bench and Terminal-bench. It supports sustained, long-running tasks, enabling continuous work for several hours, which is ideal for complex software development projects. - Financial Analysis Tools: Claude integrates seamlessly with financial data platforms such as Databricks and Snowflake, providing a unified interface for market analysis, research, and investment decision-making. It offers direct hyperlinks to source materials for instant verification, enhancing the efficiency of financial workflows. - Extended Context Windows: With an enhanced 500k context window available in Claude Sonnet 4, users can upload extensive documents, including hundreds of sales transcripts or large codebases, facilitating comprehensive analysis and collaboration. - Tool Use and Integration: Claude's extended thinking capabilities allow it to utilize tools like web search during reasoning processes, improving response accuracy. It also supports background tasks via GitHub Actions and integrates natively with development environments like VS Code and JetBrains for seamless pair programming. - Enterprise-Grade Security: The Claude Enterprise plan offers advanced security features, including Single Sign-On (SSO), Just-in-Time Provisioning (JIT), role-based permissions, audit logs, and custom data retention controls, ensuring data safety and compliance for organizations. Primary Value and User Solutions: Claude addresses the need for a reliable and intelligent AI assistant capable of handling complex tasks across various domains. For developers, it enhances productivity through advanced coding support and integration with development tools. Financial professionals benefit from its ability to unify and analyze diverse data sources, streamlining research and decision-making processes. Enterprises gain from its scalable solutions and robust security features, enabling efficient and secure deployment of AI capabilities within their operations. Overall, Claude empowers users to achieve higher efficiency, accuracy, and innovation in their respective fields.

Average Rating: 4.6/5.0

Total Reviews: 460

How Do G2 Users Rate Claude?

  • Quality of Support: 8.2/10 (Category avg: 7.9/10)
  • Content Moderation: 6.9/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.7/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.3/10 (Category avg: 8.0/10)

Who Is the Company Behind Claude?

  • Seller: Anthropic
  • HQ Location: San Francisco, California
  • Twitter: @AnthropicAI
    1,440,248 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    5,886 employees on LinkedIn®

Who Uses This Product?

  • Who Uses This: Software Engineer, Data Analyst
  • Top Industries: Computer Software, Information Technology and Services
  • Company Size: 53% Small, 33% Medium

What Do G2 Reviewers Say About Claude?

AI-generated summary from verified user reviews

Pros
  • Users value Claude's ease of use, allowing for clear, structured, and organized content creation effortlessly.
  • Users value Claude for its deep discussion capabilities and effective context management for intellectual pursuits.
  • Users find Claude to be exceptionally helpful for deep discussions, aiding in intellectual work and creative projects.
  • Users commend the accuracy of Claude, noting its well-articulated and thoroughly researched answers.
  • Users value Claude's clarity and structured communication, which greatly enhances health education and patient interactions.
Cons
  • Users face usage limitations with Claude, including access restrictions, performance inconsistency, and file handling issues.
  • Users note significant limitations with Claude, including focus on text, lack of visual support, and slow responsiveness.
  • Users find Claude's limited functionality restricts visual content creation and rapid research, impacting efficiency and usability.
  • Users find that Claude can be overly cautious and long-winded, hindering quick and effective responses.
  • Users find the resource limitations frustrating, impacting productivity and increasing costs for their workload management.

What Are Recent G2 Reviews of Claude?

Gemini

Gemini is a family of multimodal, generative AI models. These models were developed by Google DeepMind and Google Research. They are designed to understand, operate across, and combine different types of information. This includes text, images, audio, video, and code. Gemini serves as a versatile, everyday AI assistant and powers a conversational chatbot. Key Product Features & Capabilities Multimodal Understanding: Gemini understands and combines text, images, audio, video, and code. It can analyze complex documents, code repositories, and long videos. Conversational AI: Gemini allows for natural conversations. It functions as an intelligent assistant that can brainstorm, plan, and discuss topics. Deep Research & Analysis: Gemini can analyze websites and user files to generate reports. It can also create audio overviews of the information. Agentic Capabilities: Users can create custom "Gems" (specialized AI experts). The models can act as agents to take actions in tools like Chrome. Integrated Productivity: Gemini is integrated into Gmail, Google Docs, Drive, and Meet. This helps summarize, write, edit, and organize information. Creative Tools: Features include image generation and video creation, enabling the generation of 8-second videos with sound. Long Context Window: High-end models feature up to a 1 million-token context window. This is capable of analyzing large amounts of data.

Average Rating: 4.4/5.0

Total Reviews: 422

How Do G2 Users Rate Gemini?

  • Quality of Support: 8.6/10 (Category avg: 7.9/10)
  • Content Moderation: 8.3/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.3/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.9/10 (Category avg: 8.0/10)

Who Is the Company Behind Gemini?

  • Seller: Google
  • Year Founded: 1998
  • HQ Location: Mountain View, CA
  • Twitter: @google
    31,899,995 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    301,144 employees on LinkedIn®
  • Ownership: NASDAQ:GOOG

Who Uses This Product?

  • Who Uses This: Student, Software Engineer
  • Top Industries: Information Technology and Services, Computer Software
  • Company Size: 50% Small, 29% Medium

What Do G2 Reviewers Say About Gemini?

AI-generated summary from verified user reviews

Pros
  • Users find Gemini's interface incredibly easy to use, enhancing productivity with its clear and structured messaging.
  • Users find Gemini to be a reliable tool for problem-solving and summarizing information, particularly in technical scenarios.
  • Users value the helpful suggestions from Gemini, enhancing productivity and simplifying task management effectively.
  • Users appreciate the efficient content creation capabilities of Gemini, finding it quick and intuitive for drafting tasks.
  • Users love the speed of Gemini, delivering creative suggestions and solutions in mere seconds.
Cons
  • Users note the limited customization options and less accurate answers of Gemini compared to competitors like GPT and Claude.
  • Users find inaccuracy in responses concerning, as it often lacks detail and can provide incorrect technical information.
  • Users note usage limitations with Gemini, including inconsistent responses and regional feature restrictions affecting reliability.
  • Users experience technical issues with Gemini, including bugs in video analysis and limited conversational abilities.
  • Users note that the context understanding in Gemini could be enhanced, particularly with video file analysis.

What Are Recent G2 Reviews of Gemini?

Mistral AI

Mistral AI is a French artificial intelligence company specializing in developing open-source large language models (LLMs) and AI solutions tailored for diverse applications. Founded in 2023, Mistral AI focuses on creating efficient, high-performance models that empower developers and enterprises to build intelligent applications across various domains. Key Features and Functionality: - Diverse Model Offerings: Mistral AI provides a range of models, including: - Mistral Large 2: A top-tier reasoning model designed for complex tasks, supporting multiple languages and a large context window of 128K tokens. - Codestral: A specialized model optimized for coding tasks, trained on over 80 programming languages, and featuring a 32K token context window. - Pixtral Large: A multimodal model capable of analyzing and understanding both text and images. - Developer Platform (La Plateforme): Offers APIs for accessing and customizing Mistral's models, enabling deployment in various environments such as on-premises or cloud. - Le Chat: A multilingual AI assistant available on mobile platforms, known for its speed and functionalities like web search, document understanding, and code assistance. Primary Value and Solutions: Mistral AI addresses the growing demand for customizable and efficient AI models by providing open-source solutions that offer greater flexibility and control to users. Their models are designed to be deployed across various platforms, ensuring privacy and adaptability to specific enterprise needs. By focusing on open and efficient AI models, Mistral AI empowers developers and businesses to integrate advanced AI capabilities into their applications, enhancing productivity and innovation.

Average Rating: 4.2/5.0

Total Reviews: 59

How Do G2 Users Rate Mistral AI?

  • Quality of Support: 7.8/10 (Category avg: 7.9/10)
  • Content Moderation: 9.1/10 (Category avg: 8.4/10)
  • Contextual Understanding: 9.1/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.9/10 (Category avg: 8.0/10)

Who Is the Company Behind Mistral AI?

  • Seller: Mistral
  • Year Founded: 2023
  • HQ Location: Paris, Île-de-France, France
  • Twitter: @MistralAI
    195,825 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    1,540 employees on LinkedIn®

Who Uses This Product?

  • Top Industries: Information Technology and Services, Computer Software
  • Company Size: 61% Small, 34% Medium

What Do G2 Reviewers Say About Mistral AI?

AI-generated summary from verified user reviews

Pros
  • Users value the free API services of Mistral AI, enabling practical testing and comparisons with other models.
  • Users appreciate the knowledge access of Mistral AI, benefiting from its extensive capabilities and useful testing options.
Cons
  • Users find that Mistral AI has a lack of creativity, often falling short on specific tasks compared to other models.
  • Users find Mistral AI's limited capabilities often inadequate for specific tasks, leading to reliance on other models.

What Are Recent G2 Reviews of Mistral AI?

FAQs About Large Language Models (LLMs) Software

Generated using AI

Last updated: June 3, 2026

Best Large Language Models avoiding vendor lock-in concerns about data portability and long-term independence

Based on G2 reviews, these products are commonly mentioned for flexible workflows and broad day-to-day adoption.

  • ChatGPT — broad drafting, coding, and research workflows.
  • Gemini — document analysis and workspace-based productivity.
  • Claude — long documents, coding, and structured writing.
  • Deepseek — lower-cost reasoning and coding support.

Large Language Models with flexible pricing that doesn't explode as token usage increases over time unexpectedly

According to verified users, pricing concerns usually show up as usage caps, paid-plan limits, or pressure to upgrade during heavier workloads. In recent G2 reviews, buyers most often describe value in terms of time saved on drafting, research, coding, summarization, and documentation rather than raw token economics. Reviewers mention that lower-cost or free-access options can be useful for everyday tasks, but they also note tradeoffs like weaker integrations, inconsistent depth, or the need to verify outputs. For teams expecting sustained usage, the most grounded takeaway from G2 reviews is to compare plan limits, workflow fit, and how often users hit caps during normal work rather than assuming one pricing model will stay efficient at scale.

What are the best Large Language Models for small marketing teams automating customer messaging workflows

Based on G2 reviews, these products appear often in messaging, content, and workflow-related use cases.

  • ChatGPT — email drafting, messaging, and content ideas.
  • Claude — polished writing and customer communication support.
  • Gemini — Gmail-connected drafting and daily communication tasks.
  • Deepseek — email drafting and content curation.

What features matter most in llm software

According to verified users, the most valued features in llm software are fast response times, clear explanations, strong context handling, easy setup, and versatility across writing, research, coding, summarization, and analysis. Recent G2 reviews also point to workflow features such as file handling, chat history, memory, document summarization, image support, and the ability to refine outputs through follow-up questions. For workplace use, buyers repeatedly mention integrations with tools like email, documents, spreadsheets, project tools, and internal workflows as important. At the same time, reviewers consistently flag limits around accuracy, outdated information, context drift in long conversations, and plan or usage caps, so reliability and usability matter as much as feature breadth.

How do teams use Large Language Models for documentation

G2 reviewers mention that teams use Large Language Models to speed up documentation work across reports, SOPs, technical documents, summaries, presentations, emails, and customer-facing materials. In recent reviews, users describe turning rough notes into structured drafts, summarizing long files, refining tone, and preparing repeatable documentation faster than manual workflows. Technical teams also mention using these tools for code explanations, report preparation, ticket refinement, and knowledge-base style outputs. The buyer takeaway is that documentation value comes from reducing first-draft time and organizing complex information quickly, but reviewers still recommend human review for specialized, business-critical, or rapidly changing content because answers can sometimes be too generic, inaccurate, or inconsistent.

Deepseek

DeepSeek LLM is a series of high-performance, open-source large language models from China-based DeepSeek AI.

Average Rating: 4.4/5.0

Total Reviews: 18

How Do G2 Users Rate Deepseek?

  • Quality of Support: 7.5/10 (Category avg: 7.9/10)
  • Content Moderation: 8.6/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.6/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.8/10 (Category avg: 8.0/10)

Who Is the Company Behind Deepseek?

  • Seller: DeepSeek
  • Year Founded: 2023
  • HQ Location: Hangzhou
  • LinkedIn® Page: www.linkedin.com
    200 employees on LinkedIn®

Who Uses This Product?

  • Top Industries: Computer Software
  • Company Size: 72% Small, 22% Medium

What Do G2 Reviewers Say About Deepseek?

AI-generated summary from verified user reviews

Pros
  • Users praise Deepseek for its exceptional performance and quick results, making it ideal for various tasks and industries.
  • Users find Deepseek to be extremely easy to use, facilitating quick research and content creation effortlessly.
  • Users appreciate the accuracy of Deepseek, enjoying its ability to provide reliable and prompt responses for various tasks.
  • Users enjoy the effective content creation capabilities of DeepSeek, finding it fast and easy for their needs.
  • Users appreciate the creativity enhancement of DeepSeek, generating fresh and innovative ideas effectively for various needs.
Cons
  • Users find that Deepseek struggles with context understanding, occasionally failing to respond accurately to complex prompts.
  • Users report concerns about low accuracy, with Deepseek often failing to meet expectations for detailed and reliable responses.
  • Users have faced technical issues with Deepseek, particularly regarding real-time data and server errors affecting usability.
  • Users express concerns over bias and censorship in Deepseek, questioning its reliability for unbiased information generation.
  • Users express significant concerns about data security, particularly regarding privacy risks and potential biases in information accuracy.

What Are Recent G2 Reviews of Deepseek?

Grok

Grok is your truth-seeking AI companion for unfiltered answers with advanced capabilities in reasoning, coding, and visual processing.

Average Rating: 4.1/5.0

Total Reviews: 53

How Do G2 Users Rate Grok?

  • Quality of Support: 7.7/10 (Category avg: 7.9/10)
  • Content Moderation: 8.5/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.8/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.4/10 (Category avg: 8.0/10)

Who Is the Company Behind Grok?

  • Seller: xAI
  • Year Founded: 2022
  • HQ Location: Asnières-sur-Seine, FR
  • LinkedIn® Page: www.linkedin.com
    3 employees on LinkedIn®

Who Uses This Product?

  • Top Industries: Information Technology and Services, Computer Software
  • Company Size: 70% Small, 17% Large

What Do G2 Reviewers Say About Grok?

AI-generated summary from verified user reviews

Pros
  • Users find Grok to be incredibly user-friendly, making it easy to create engaging content effortlessly.
  • Users appreciate the speed and clarity of Grok's research, making content preparation and information gathering efficient and straightforward.
  • Users praise Grok for its fast research capabilities, significantly enhancing efficiency and clarity in complex health topics.
  • Users value Grok for its rapid response times, enhancing their ability to conduct efficient research and support teaching.
  • Users appreciate Grok's versatility in research and content creation, seamlessly supporting various professional needs.
Cons
  • Users report low accuracy with Grok, leading to wasted time and frustration due to repeated errors and misinformation.
  • Users often confront technical issues with Grok, leading to frustration and significant time wasted on unproductive queries.
  • Users find Grok's limited context understanding inadequate for deep research, impacting accuracy and usability in complex tasks.
  • Users often face inaccurate responses from Grok, leading to misunderstandings and frustration during their tasks.
  • Users experience hallucinations with Grok, as it struggles with misinformation and unverified data from social media.

What Are Recent G2 Reviews of Grok?

Llama

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model developed by Meta, designed to handle both text and image inputs while generating multilingual text and code outputs across 12 languages. Built on a mixture-of-experts (MoE) architecture with 128 experts, it activates 17 billion parameters per forward pass out of a total of 400 billion, ensuring efficient processing. Optimized for vision-language tasks, Maverick is instruction-tuned to exhibit assistant-like behavior, perform image reasoning, and facilitate general-purpose multimodal interactions. It features early fusion for native multimodality and supports a context window of up to 1 million tokens. Trained on approximately 22 trillion tokens from a curated mix of public, licensed, and Meta-platform data, with a knowledge cutoff in August 2024, Maverick was released on April 5, 2025, under the Llama 4 Community License. It is well-suited for research and commercial applications requiring advanced multimodal understanding and high model throughput. Key Features and Functionality: - Multimodal Input Support: Processes both text and image inputs, enabling comprehensive understanding and generation capabilities. - Multilingual Output: Generates text and code outputs in 12 languages, including Arabic, English, French, German, Hindi, Indonesian, Italian, Portuguese, Spanish, Tagalog, Thai, and Vietnamese. - Mixture-of-Experts Architecture: Utilizes 128 experts with 17 billion active parameters per forward pass, optimizing computational efficiency and performance. - Instruction-Tuned: Fine-tuned for assistant-like behavior, image reasoning, and general-purpose multimodal interactions, enhancing its applicability across various tasks. - Extended Context Window: Supports a context length of up to 1 million tokens, facilitating the processing of extensive and complex inputs. Primary Value and User Solutions: Llama 4 Maverick 17B Instruct addresses the growing demand for advanced AI models capable of understanding and generating content across multiple modalities and languages. Its multimodal and multilingual capabilities make it an invaluable tool for developers and researchers working on applications that require nuanced language understanding, image processing, and code generation. The model's instruction-tuned nature ensures it can perform a wide range of tasks with high accuracy, from serving as an intelligent assistant to executing complex reasoning tasks. Its efficient architecture and extended context window allow for the handling of large-scale data inputs, making it suitable for both research and commercial applications that demand high throughput and advanced multimodal understanding.

Average Rating: 4.3/5.0

Total Reviews: 151

How Do G2 Users Rate Llama?

  • Quality of Support: 7.1/10 (Category avg: 7.9/10)
  • Content Moderation: 7.6/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.5/10 (Category avg: 8.6/10)
  • Bias Mitigation: 7.8/10 (Category avg: 8.0/10)

Who Is the Company Behind Llama?

Who Uses This Product?

  • Who Uses This: Software Engineer
  • Top Industries: Computer Software, Information Technology and Services
  • Company Size: 58% Small, 24% Medium

What Do G2 Reviewers Say About Llama?

AI-generated summary from verified user reviews

Pros
  • Users highlight the exceptional accuracy of Llama 3, enhancing interactions with its advanced contextual understanding.
  • Users find Llama 3 to be incredibly user-friendly, excelling in complex tasks and delivering quick, accurate results.
  • Users value the fast response times of Llama 3, enjoying enhanced accuracy and efficiency in their tasks.
  • Users love the open-source nature of Llama, enabling cost-effective solutions and fostering development of new AI tools.
  • Users find Llama to be very helpful with excellent instruction-following and quick, accurate responses for various applications.
Cons
  • Users face limitations in response capabilities, struggling with tasks like text recognition and indexing support.
  • Users note the slow performance of Llama, especially when compared to competitors, impacting overall efficiency and usability.
  • Users find Llama's responses to be poorly quality, often repetitive and lacking in depth, making it frustrating to use.
  • Users report inaccuracies in Llama's responses, necessitating careful verification before relying on generated content.
  • Users note a limited understanding in Llama, struggling with complex topics and context retention.

What Are Recent G2 Reviews of Llama?

bloom

The BLOOM model has been proposed with its various versions through the BigScience Workshop. BigScience is inspired by other open science initiatives where researchers have pooled their time and resources to collectively achieve a higher impact. The architecture of BLOOM is essentially similar to GPT3 (auto-regressive model for next token prediction), but has been trained on 46 different languages and 13 programming languages. Several smaller versions of the models have been trained on the same dataset. BLOOM is available in the following versions:

Average Rating: 4.3/5.0

Total Reviews: 7

How Do G2 Users Rate bloom?

  • Quality of Support: 8.0/10 (Category avg: 7.9/10)
  • Content Moderation: 10.0/10 (Category avg: 8.4/10)
  • Contextual Understanding: 8.7/10 (Category avg: 8.6/10)
  • Bias Mitigation: 10.0/10 (Category avg: 8.0/10)

Who Is the Company Behind bloom?

  • Seller: Hugging Face
  • Year Founded: 2016
  • HQ Location: United States
  • Twitter: @huggingface
    708,886 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    984 employees on LinkedIn®

Who Uses This Product?

  • Company Size: 71% Small, 29% Medium

What Are Recent G2 Reviews of bloom?

Kimi

1. Kimi Reviews & Product Details Kimi K3 is Moonshot AI's most capable open-source model, with 2.8T parameters and a 1M context built for real enterprise workloads — from long-horizon reasoning to complex problem-solving. Empowered by a versatile ecosystem of plugins, it slots directly into your team's workflows and takes complex tasks end to end: coding, deep research grounded in reliable data sources, presentation decks and files — all delivered with professionalism and agency-level aesthetics. 2. Key Features and Functionality: - Enterprise-Ready by Default: Built with strict data isolation to safeguard your corporate assets and intellectual property. Coupled with streamlined procurement—one contract, one invoice, seamlessly empowering every team. - Professional Data Plugins and Skills: With professional data sources and multiple expert skills covering equity research, investment banking & PE, and corporate finance, deliver low-hallucination analysis. - Kimi Work — Your Digital Employees: Goal mode takes a target and drives it to delivery; Agent Swarm coordinates up to 300 sub-agents on a single objective. Outputs arrive as finished work: Word documents, Excel models, slide decks, live dashboards, and web apps. - Long-Horizon Agentic Execution: K3 sustains deep reasoning over extended workflows, autonomously planning, researching, and iterating to drive complex projects from start to finish. - Design-Grade Frontend Output: K3 brings agency-level aesthetics to everyday business deliverables - from landing pages to dashboards - all come out visually polished and on-brand. Teams get production-ready interfaces from one prompt. 3. Primary Value and User Solutions: Kimi gives enterprise teams a single AI platform that takes complex work from request to delivery. Long-horizon reasoning handles multi-hour coding, deep research, and financial analysis, while Kimi Work turns that capability into finished work products, grounded by professional data plugins. Kimi covering engineering, research, finance, and content replaces a patchwork of vertical tools — with no training on customer data, unified procurement and invoicing, and flagship-level performance at a fraction of the typical cost.

Average Rating: 4.0/5.0

Total Reviews: 2

How Do G2 Users Rate Kimi?

  • Quality of Support: 8.3/10 (Category avg: 7.9/10)

Who Is the Company Behind Kimi?

Who Uses This Product?

  • Company Size: 100% Small

What Are Recent G2 Reviews of Kimi?

Phi

Phi-4 is a state-of-the-art language model developed by Microsoft Research, designed to deliver advanced reasoning capabilities within a compact architecture. With 14 billion parameters, this dense decoder-only Transformer model is optimized for text-based inputs, particularly excelling in chat-based prompts. Trained on a diverse dataset comprising 9.8 trillion tokens—including synthetic datasets, filtered public domain content, academic literature, and Q&A datasets—Phi-4 emphasizes high-quality data to enhance its reasoning abilities. The model underwent rigorous enhancement and alignment processes, incorporating both supervised fine-tuning and direct preference optimization to ensure precise instruction adherence and robust safety measures. Released on December 12, 2024, under the MIT license, Phi-4 is tailored for applications requiring efficient performance in memory or compute-constrained environments, latency-sensitive scenarios, and tasks demanding advanced reasoning and logic. Key Features and Functionality: - Advanced Reasoning: Phi-4 is engineered to perform complex reasoning tasks, making it suitable for applications that require logical processing and decision-making. - Efficient Architecture: With 14 billion parameters, the model offers a balance between performance and resource utilization, catering to environments with memory and compute constraints. - Extensive Training Data: The model is trained on a vast dataset of 9.8 trillion tokens, including high-quality synthetic data, filtered public domain content, academic books, and Q&A datasets, ensuring a comprehensive understanding of diverse topics. - Optimized for Chat Prompts: Phi-4 excels in generating coherent and contextually relevant responses to chat-based inputs, enhancing user interaction experiences. - Safety and Alignment: The model incorporates supervised fine-tuning and direct preference optimization to adhere to instructions accurately and maintain robust safety measures. Primary Value and User Solutions: Phi-4 addresses the need for a powerful yet efficient language model capable of advanced reasoning in resource-constrained environments. Its optimized architecture and extensive training enable developers to integrate sophisticated AI capabilities into applications without compromising performance. By focusing on high-quality data and safety measures, Phi-4 ensures reliable and contextually appropriate responses, making it a valuable tool for enhancing user engagement and decision-making processes in various applications.

Average Rating: 4.0/5.0

Total Reviews: 1

How Do G2 Users Rate Phi?

  • Quality of Support: 8.3/10 (Category avg: 7.9/10)
  • Contextual Understanding: 8.3/10 (Category avg: 8.6/10)

Who Is the Company Behind Phi?

  • Seller: Microsoft
  • Year Founded: 1975
  • HQ Location: Redmond, Washington
  • Twitter: @microsoft
    13,091,739 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    232,750 employees on LinkedIn®
  • Ownership: MSFT

Who Uses This Product?

  • Company Size: 100% Large

What Do G2 Reviewers Say About Phi?

AI-generated summary from verified user reviews

Pros
  • Users value the easy integrations of Phi, especially its seamless compatibility with Microsoft Azure tools.
  • Users highlight the efficiency of Phi, noting it outperforms many similar-sized models while being cost-effective.
Cons
  • Users find that Phi may struggle with complex tasks compared to larger models like GPT-4.

What Are Recent G2 Reviews of Phi?

Amazon Nova

Amazon Nova is a suite of advanced foundation models developed by Amazon, designed to deliver state-of-the-art intelligence and industry-leading price performance. Integrated within Amazon Bedrock, these models support a wide range of tasks across multiple modalities, including text, image, and video processing. Amazon Nova aims to simplify the development of generative AI applications by offering versatile and cost-effective solutions for businesses and developers.

Who Is the Company Behind Amazon Nova?

  • Seller: Amazon Web Services (AWS)
  • Year Founded: 2006
  • HQ Location: Seattle, WA
  • Twitter: @awscloud
    2,232,483 Twitter followers
  • LinkedIn® Page: www.linkedin.com
    147,094 employees on LinkedIn®
  • Ownership: NASDAQ: AMZN

Athene 70B

Athene-70B is an advanced open-weight language model developed by Nexusflow, built upon Meta's Llama-3-70B-Instruct architecture. Utilizing Reinforcement Learning from Human Feedback , Athene-70B achieves a 77.8% score on the Arena-Hard-Auto benchmark, positioning it competitively against proprietary models like Claude-3.5-Sonnet and GPT-4o. This model excels in tasks requiring precise instruction following, complex reasoning, comprehensive coding assistance, creative writing, and multilingual understanding. Its open-weight nature allows for broad accessibility, enabling developers and researchers to integrate and adapt the model for various applications. Key Features and Functionality: - High Performance: Achieves a 77.8% score on the Arena-Hard-Auto benchmark, closely matching leading proprietary models. - Advanced Training: Fine-tuned using RLHF to enhance desired behaviors and performance. - Versatile Capabilities: Excels in instruction following, complex reasoning, coding assistance, creative writing, and multilingual tasks. - Open-Weight Accessibility: Provides transparency and adaptability for developers and researchers. Primary Value and User Solutions: Athene-70B offers a high-performing, open-weight alternative to proprietary language models, enabling users to develop sophisticated AI applications without the constraints of closed-source systems. Its advanced capabilities in understanding and generating human-like text make it suitable for a wide range of applications, including conversational agents, content creation, and complex problem-solving tasks. By providing an accessible and adaptable model, Athene-70B empowers users to innovate and tailor AI solutions to their specific needs.

Who Is the Company Behind Athene 70B?

Command

Command A is Cohere's most advanced large language model, specifically engineered to meet the complex demands of enterprise applications. With 111 billion parameters and a context length of 256,000 tokens, it excels in tasks such as tool use, retrieval-augmented generation , agent-based workflows, and multilingual processing across 23 languages. Designed for efficient deployment, Command A operates effectively on just two GPUs, making it a cost-effective solution for businesses seeking high-performance AI capabilities. Key Features and Functionality: - High Performance: Delivers top-tier results in enterprise tasks, including tool integration, RAG, and agentic operations. - Extended Context Length: Supports up to 256,000 tokens, enabling the processing of extensive documents and complex datasets. - Multilingual Support: Proficient in 23 languages, facilitating global business applications. - Efficient Deployment: Operates on minimal hardware—specifically, two A100 or H100 GPUs—reducing infrastructure costs. - Data Security: Designed for on-premise or Virtual Private Cloud deployment, ensuring sensitive data remains within the organization's control. Primary Value and User Solutions: Command A addresses the critical need for enterprises to integrate advanced AI into their operations without compromising on performance, scalability, or data security. By automating complex workflows, enhancing content generation, and supporting multilingual communication, it empowers organizations to boost productivity and maintain a competitive edge in the global market. Its efficient deployment requirements make it accessible to businesses seeking powerful AI solutions without significant hardware investments.

Who Is the Company Behind Command?

  • Seller: Cohere
  • Year Founded: 2019
  • HQ Location: Toronto, Ontario, Canada
  • LinkedIn® Page: www.linkedin.com
    818 employees on LinkedIn®
Bijou Barry
BB
Researched and written by Bijou Barry
Updated April 9, 2026