Thought ArticlesSeptember 9, 202619 min read

Top 11 Generative AI Companies in 2026: Products, Pricing, Funding, and Competitive Analysis

Key Highlights

Generative AI is accelerating as leading companies release new frontier models, expand multimodal and agentic capabilities, and compete on performance, pricing, and enterprise adoption. This article compares 11 generative AI companies, including OpenAI, Anthropic, Google DeepMind, Microsoft, DeepSeek, Mistral AI, Perplexity AI, xAI, Meta AI, Sarvam AI, and Kimi AI.

The generative AI industry continues to reshape content creation, software development, enterprise automation, search, and customer engagement at a breakneck pace. Guided by the rapid evolution of large language models (LLMs), multimodal AI, and autonomous AI agents, the last few weeks alone (late August into early September 2026) produced four separate frontier-model launches within a 72-hour window, a pace that underlines just how compressed the competitive cycle has become.

This article analyzes the top 11 generative AI companies in the world, OpenAI, Anthropic, Google DeepMind, Microsoft, DeepSeek, Mistral AI, Perplexity AI, xAI (SpaceXAI), Meta AI, Sarvam AI, and Kimi AI (Moonshot AI), and examines their product portfolios, pricing, funding history, and the newest models each has shipped.

1. OpenAI

Company Overview

Parameter

Details

Headquarters

San Francisco, California, USA

CEO

Sam Altman

Founded

2015

Website

https://openai.com/

Core Focus

Frontier flagship models, multimodal AI, agentic tools, enterprise AI

OpenAI remains the largest commercial generative AI company globally. Its core business spans ChatGPT, the GPT model series, enterprise APIs, agentic AI products, and cloud partnerships.

Subscription Plans Pricing

Plan

Price

Free

$0

ChatGPT Go

$8/month (varies by region)

ChatGPT Plus

$20/month

ChatGPT Pro

$200/month

Enterprise

Custom pricing

Selected Open API Pricing for Flagship Models

Model

Processing Model

Input (per 1M tokens)

Cached Input (per 1M tokens)

Output (per 1M tokens)

Application

GPT 6 Astra

Standard

$10.00

$1.00

$50.00

Complex reasoning, coding, computer use, research, and document creation

GPT-5.5

Standard

$5.00

$0.50

$30.00

Premium multimodal model for coding and professional work

Batch-50%

$2.50

$0.25

$15.00

Data Residency +10%

$5.50

$0.55

$33.00

GPT-5.4

Standard

$2.50

$0.25

$15.00

Affordable for coding & professional work

Batch-50%

$1.25

$0.125

$7.50

Data Residency +10%

$2.75

$0.275

$16.50

GPT-5.4 mini

Standard

$0.75

$0.075

$4.50

Strongest mini model for coding & subagents

Batch-50%

$0.375

$0.0375

$2.25

Data Residency +10%

$0.825

$0.0825

$4.95

Latest Model: GPT-6 Astra (launched September 3, 2026)

The biggest OpenAI news since this report was last published is the release of GPT-6 Astra, which OpenAI calls its most powerful and capable model yet. Astra is notable for several reasons:

  • It is OpenAI's first model to cross the Critical cybersecurity capability threshold under the company's Preparedness Framework, meaning it can independently discover and develop exploits for previously unknown vulnerabilities across well-protected systems.

  • OpenAI is rolling it out in phases: first to organizations in its Daybreak cybersecurity program, then to ChatGPT Plus, Pro, Business, and Enterprise users, and via the OpenAI API and Amazon Web Services in the coming days after the September 3 announcement.

  • Astra is pitched as a new frontier on computer and browser use, with specs including a 1.05-million-token context window, 128K max output, text-and-image input, and a knowledge cutoff of April 30, 2026. It uses a recurrent depth (looped transformer) reasoning technique that improves efficiency but has drawn scrutiny because it obscures parts of the model's chain of thought.

  • OpenAI president Greg Brockman described the launch as a generational leap, with some commentary framing it as an early signal of the AGI era, a characterization that has proven controversial given the model's dual-use cyber capability.

Other Recent Developments

  • OpenAI's frontier models and Codex became broadly available through Amazon Bedrock in June 2026, easing enterprise adoption within existing AWS security and compliance workflows.

  • OpenAI rolled out Active Sessions, a security feature for reviewing account sessions, in June 2026. Further, in May 2026, OpenAI updated GPT-5.5 Instant on ChatGPT and in the API to enhance response mannerisms and quality for smoother flow and more natural interaction.

  • In March 2026, OpenAI reported raising $122 billion at an $852 billion post-money valuation to fund frontier-model development, compute infrastructure, and deeper integration across ChatGPT, Codex, browsing, and agentic workflows.

  • The GPT-5.6 family (Sol, Terra, Luna), ChatGPT Work, and GPT-Live were introduced in July 2026 and remain in active use as OpenAI's mid-tier lineup beneath Astra.

2. Anthropic

Company Overview

Parameter

Details

Headquarters

San Francisco, California, USA

CEO

Dario Amodei

Founded

2021

Website

https://www.anthropic.com/

Core Focus

Safety-centered AI systems, enterprise reasoning, coding models

Anthropic develops the Claude family of models, known for an emphasis on safety-oriented development, long-context reasoning, strong coding performance, and enterprise-grade deployment. Claude is widely used in software engineering, knowledge work, and agentic workflows.

Subscription Plans Pricing

Plan

Price

Claude (Free)

$0

Claude Pro

$17/month

Claude Max

5x: $100/person/month · 20x: $200/person/month

Team

Standard seat: $20–25/month · Premium seat: $100–125/month

Enterprise

Custom pricing

Selected Claude API Pricing

Model

Processing Model

Input (per 1M tokens)

Output (per 1M tokens)

Prompt caching (per 1M tokens)

Application

Write

Read

Opus 4.8

Standard

$5

$25

$6.25

$0.50

Most intelligent model for agents and coding

Batch-50%

$2.50

$12.50

$3.13

$0.25

Sonnet 4.6

Standard

$3

$15

$3.75

$0.30

Optimal balance of intelligence, cost, and speed

Batch-50%

$1.50

$7.50

$1.88

$0.15

Haiku 4.5

Standard

$1

$5

$1.25

$0.10

Strongest mini model for coding & subagents

Batch-50%

$0.50

$2.50

$0.63

$0.05

Products Offering

Product

Description

Claude

Main chat interface for conversation, writing, research, and analysis

Claude Code

Agentic coding system that reads codebases, edits files, runs tests, and ships committed code.

Claude Cowork

Desktop automation for non-coders, orchestrating files and tasks across a computer

Claude Security

AI-powered security tooling for threat discovery, investigation, and remediation

Latest Models: Claude Fable 5.1 and Claude Mythos 5.1 (launched September 1, 2026)

Anthropic's most recent releases update the “Mythos-class” model line first introduced in June 2026:

  • Claude Fable 5.1 is the generally available version, built for demanding reasoning, long-running agents, coding, multistep research, and document-heavy professional work. It offers a 1-million-token context window, up to 128K output tokens, and “adaptive thinking” that stays active by default.

  • Claude Mythos 5.1 shares the same underlying model but with safety mitigations lifted in select areas (cybersecurity, biology/chemistry, and model distillation), and remains available only to a limited set of vetted U.S. cyberdefense and life-science organizations through Project Glasswing, with international expansion planned in coordination with the U.S. government.

  • Both models can only be replayed against thinking blocks from equal-or-newer Claude models, a new safeguard that prevents older models from reading their extended reasoning.

  • Anthropic also introduced Enterprise Frontier Safeguards, a phased system giving enterprise customers control over activity storage and automated misuse monitoring.

Other Recent Developments

  • Anthropic's recent flagship model was Claude Opus 4.8, which was launched in May of 2026. It provides more coherent reasoning, faster answers, and better coding. It also greatly improves Claude's performance on complex, multi-step tasks to make it more reliable for enterprise and developer use cases.

  • In May 2026, Anthropic launched its Series H funding round, with billions raised to boost research and infrastructure for frontier AI. The funds are focused on speeding up model development, increasing compute resources, and deepening partnerships with companies adopting Claude.

  • In May 2026, Anthropic announced an expansion of partnerships with SpaceX to extend compute limits for Claude models. They work together to allow Claude greater processing of high-capacity workloads and advanced outputs, thereby serving key parts of the economy that rely on AI operations in much larger data volumes.

  • SAP and Anthropic partnered to bring Claude into SAP Business AI in May 2026. It aspires to equip SAP with cutting-edge natural language capabilities over its enterprise platform, empowering organizations in areas like workflow automation, data analysis, and AI-driven decision-making.

3. Google DeepMind

Company Overview

Parameter

Details

Headquarters

London, United Kingdom

CEO

Demis Hassabis

Founded

2010

Website

https://deepmind.google

Core Focus

Foundation models, multimodal AI, enterprise AI, AI agents

Google DeepMind is basically Google’s flagship AI research and commercialization segment, and it’s also the team behind the Gemini model family. Through tighter integration across Search, Android, Workspace, Cloud, YouTube, and the developer ecosystem, Gemini has ended up being one of the most widely spread generative AI platforms around the world. Lately, they’ve been putting more emphasis on multimodal AI, agentic workflows, video generation, and enterprise AI infrastructure.

Subscription Plans Pricing

Plan

Price

Google AI (Free)

$0

Google AI Plus

$7.99/month

Google AI Pro

$19.99/month

Google AI Ultra

$99.99/month

Enterprise

Custom pricing

Selected Gemini API Pricing

Model

Processing Model

Input (per 1M tokens)

Context caching price (per 1M tokens)

Output (per 1M tokens)

Application

Gemini 3.5 Flash

Standard

$1.50

$0.50

$9.00

It is an intelligent model built for speed, though, mixing frontier intelligence with sharper search and grounding.

Batch

$0.75

$0.075

$4.50

Flex

$0.75

$0.08

$4.50

Priority

$2.70

$0.27

$16.20

Gemini 3.1 Flash-Lite

Standard

$0.25 (text/image/video)

$0.50 (audio)

$0.025 (text/image/video)
$0.05 (audio)

$1.50

 Best model family globally for multimodal understanding, agentic capabilities, and coding.

Batch

$0.125 (text/image/video)

$0.25 (audio)

$0.0125 (text/image/video)

$0.025 (audio)

$0.75

Flex

$0.125 (text/image/video)

$0.25 (audio)

$0.0125 (text/image/video)

$0.025 (audio)

$0.75

Priority

$0.45 (text/image/video)
$0.90 (audio)

$0.045 (text/image/video)
$0.09 (audio)

$2.70

Gemini 3.1 Pro Preview

Standard

$2.00, prompts <= 200k tokens
$4.00, prompts > 200k tokens

$12.00, prompts <= 200k tokens
$18.00, prompts > 200k

$0.20, prompts <= 200k tokens
$0.40, prompts > 200k

A high-performance reasoning model designed for complex problem-solving, coding, multimodal understanding, and advanced agentic applications.

Batch

$1.00, prompts <= 200k tokens
$2.00, prompts > 200k tokens

$0.20, prompts <= 200k tokens
$0.40, prompts > 200k

$6.00, prompts <= 200k tokens
$9.00, prompts > 200k

Flex

$1.00, prompts <= 200k tokens
$2.00, prompts > 200k tokens

$0.20, prompts <= 200k tokens
$0.40, prompts > 200k

$6.00, prompts <= 200k tokens
$9.00, prompts > 200k

Priority

               $3.60, prompts <= 200k tokens

$7.20, prompts > 200k tokens

$0.36, prompts <= 200k tokens
$0.72, prompts > 200k

$21.60, prompts <= 200k tokens

$32.40, prompts > 200k

Products Offering

Product

Description

Veo

Cinematic-quality video generation with audio

Imagen

Text-to-image generation and editing

Lyria

High-fidelity music and audio generation

Latest Model: Gemini 3.8 Flash (launched September 2, 2026)

Google shipped Gemini 3.8 Flash alongside a defenders-only “Cyber” variant on September 2, 2026, continuing a fast Flash-tier cadence. Notably, as of this update, Google still has not shipped a refreshed flagship Gemini Pro model; Gemini 3.1 Pro (released February 2026) remains the top-of-line reasoning model even as Google's Flash tier iterates rapidly on coding, efficiency, and cybersecurity use cases. Industry commentary has flagged the delayed Pro refresh as notable given how aggressively OpenAI and Anthropic have been shipping frontier-tier models in the same window.

Other Recent Developments

  • Gemini 3.7 Flash launched in August 2026 as Google's most intelligent workhorse model yet for coding and agents, alongside Gemini 3.5 Transcribe (speech-to-text) reaching general availability.

  • The Gemini app crossed 1 billion monthly users in August 2026.

  • Google also launched the Pixel 11 series with deep Gemini integration, expanded Gemini in Chrome on Android, and added hands-free voice tools across Workspace and Gemini Live.

  • Computer-use capability was added to Gemini 3.5 Flash in June 2026, along with Gemini 3.5 Live Translate for low-latency voice translation.

  • In May 2026, Google DeepMind opened a new AI research lab in Singapore focused on responsible frontier AI development for the Asia-Pacific region, and Gemini Spark launched as a proactive task-completing AI agent.

4. Microsoft

Company Overview

Parameter

Details

Headquarters

Redmond, Washington, USA

CEO

Satya Nadella

Founded

1975

Website

https://copilot.microsoft.com/

Focus

Enterprise AI, productivity AI, coding AI

Microsoft created one of the biggest enterprise AI ecosystems worldwide through Copilot, Azure AI, GitHub Copilot, and Microsoft 365 integrations. Additionally, it uses both OpenAI technologies as well as internally developed MAI models to reinforce its whole AI platform strategy.

Subscription Plans Pricing

Plan

Price

Microsoft 365 (Free)

$0

Microsoft 365 Basic

$1.99/month

Microsoft 365 Personal

$9.99/month

Microsoft 365 Family

$12.99/month

Microsoft 365 Premium

$19.99/month

Selected Azure OpenAI API Pricing

Model

Processing Model

Input (per 1M tokens)

Cached Input (per 1M tokens)

Output (per 1M tokens)

Application

GPT 5.5 Series

Standard

$5

$0.50

$30

Provides users with advanced reasoning, instruction following, and agentic capabilities for production AI workloads.

Priority Processing

$12.50

$1.25

$75

GPT 5.4 Series

Standard

$2.50

$0.25

$15

Offer Stronger reasoning, dependable execution, and agentic workflows at scale.

Priority Processing

$5

$0.50

$30

Pricing with Batch API

$1.25

$0.13

$7.50

GPT-5.3 Series

Standard

$1.75

$0.18

$14

It unifies the frontier coding performance of GPT-5.2-Codex with the reasoning and professional knowledge abilities of GPT-5.2.

Priority Processing

$3.50

$0.35

 $28

GPT-5.2

Standard

$1.75

$0.18

$14

Offers the deep reasoning and expanded context handling needed for building more sophisticated AI agents

Priority Processing

$3.50

$0.35

 $28

GPT-5.1

Standard

$1.25

 $0.13

$10

Provide faster responses to users in diverse situation faster to users with adaptive reasoning.

Priority Processing

$2.50

$0.25

$20

Products Offering

Product

Description

Microsoft 365 Copilot

AI assistant integrated across Microsoft 365 apps

Microsoft Copilot Studio

Platform for building and deploying custom AI agents

Microsoft Azure

Enterprise-grade platform for building and governing AI applications

Recent Developments

  • At Microsoft Build (June 2, 2026), Microsoft detailed seven internally developed MAI models: MAI-Thinking-1, MAI-Image-2.5 (plus a Flash variant), MAI Transcribe 1.5, MAI-Voice-2 (plus a Flash variant), and MAI-Code-1, covering reasoning, image generation/editing, transcription, voice, and coding.

  • On July 2, 2026, Microsoft announced Microsoft Frontier Company, backed by a $2.5 billion investment and roughly 6,000 specialists, to help enterprises design and deploy AI systems.

  • Microsoft 365 Business with Copilot launched in May 2026 as a small-business-focused subscription combining productivity apps, Copilot features, and built-in security.

  • Microsoft's frontier model access spans GPT-6 Astra, GPT-5.6, and prior GPT-5.x series via Azure OpenAI, alongside Claude availability through Microsoft Foundry.

  • In April 2026, Microsoft announced a $10 billion investment in Japan to strengthen AI infrastructure, cybersecurity, and workforce development.

5. DeepSeek

Company Overview

Parameter

Details

Headquarters

Hangzhou, China

CEO

Liang Wenfeng

Website

https://www.deepseek.com/

Founded

2023

Focus

Open-weight frontier reasoning LLMs

DeepSeek is one of the more disruptive AI companies globally, after it released reasoning models that seemed really capable, with inference costs that were way lower than a lot of Western competitors. Further, the open-weight strategy kind of shifted the market pricing and adoption paths, with more users adopting it.

The company has various models tailored for diverse applications like deepseek-v4-flash, deepseek-v4-pro, deepseek-chat, and deepseek-reasoner, with plans to deprecate older versions, i.e., deepseek-chat and deepseek-reasoner, after 24th July 2026. DeepSeek offers free consumer chat access, while API usage is paid according to the input and output token rates shown below.

Selected API Pricing

Model

Input (cache hit)

Input (cache miss)

Output

Concurrency

DeepSeek-V4-Flash

$0.0028/1M

$0.14/1M

$0.28/1M

2,500

DeepSeek-V4-Pro

$0.003625/1M

$0.435/1M

$0.87/1M

500

DeepSeek-V3.1 API

$0.07/1M

$0.56/1M

$1.68/1M

Product Offerings

Product

Description

DeepSeek-V3.1

It is a hybrid model with stronger agent skills, tool use, and 128K context.

DeepSeek-R1 0528

It is an upgraded AI model with front-end capabilities and fewer hallucinations, JSON output, and function calling support.

DeepSeek V3-0324

It is a V3 update with major reasoning, front-end, and tool-use improvements; released under the MIT License.

Recent Developments

  • DeepSeek shipped a DeepSeek V4 Flash Vision (experimental) update on August 21, 2026, extending the V4 line toward multimodal vision tasks.

  • The V4 family (introduced April 2026) treats a ~1M-token context length as standard across services, with V4-Pro (1.6T total / 49B active parameters) tuned for reasoning and coding benchmarks, and V4-Flash (284B total / 13B active parameters) tuned for speed and cost efficiency. Both support dual Thinking/Non-Thinking modes.

  • Older models, deepseek-chat and deepseek-reasoner, are scheduled for deprecation after July 24, 2026, as DeepSeek consolidates around the V4 line.

  • DeepSeek continues to compete closely with rising Chinese open-weight rivals such as Alibaba's Qwen3.6/3.8 line and Zhipu AI's GLM-5.x series, both of which have shipped multiple releases (e.g., GLM-5.3 in August 2026, Qwen3.8 Flash in late August 2026) in the same window.

6. Mistral AI

Company Overview

Parameter

Details

Headquarters

Paris, France

CEO

Arthur Mensch

Website

https://mistral.ai/

Founded

2023

Core Focus

Frontier open-weight LLMs, assistants, agents

Mistral AI is Europe’s leading frontier AI startup, and it’s a strong supporter of open-weight AI development. The company targets both developers and enterprises, and it keeps promoting European AI sovereignty through offerings such as enterprise-grade and self-hostable foundation models.

Subscription Plans Pricing

Plan

Price

Mistral AI (free)

$0

Mistral Pro

$14.99/month

Mistral Team

$24.99/month

Mistral Enterprise

Custom pricing

Mistral Education Plan

$5.99/month

Selected API Pricing

Model

Input

Output

Detail

Mistral Medium 3.5

$1.50

$7.50

Instruction-following, reasoning, and coding in a 128B dense model

Mistral Large 3

$0.50

$1.50

Open-weight flagship multimodal, multilingual model

Devstral 2

$0.40

$2.00

Open-weight agentic coding model

Mistral Small 4

$0.15

$0.60

SOTA multimodal, multilingual, Apache 2.0

Magistral Medium

$2.00

$5.00

Transparent, multilingual reasoning model

Product Offerings

Product

Description

Codestral

It is the first-ever code model and is an open-weight generative AI model for code generation work.

Mistral Large

It is a flagship and cutting-edge text generation model utilized for complex multilingual reasoning tasks.

Recent Developments

Mistral has kept an unusually fast release cadence through the back half of 2026:

  • OCR 4 launched June 23, 2026, supporting 170 languages, up to 2,000 pages/minute on a single GPU, and a self-hostable container option. It was quickly followed by OCR 4.1 (by mid-August 2026), which became the default mistral-ocr-latest API alias.

  • Leanstral 1.5, a formal-proof/verification-focused model, launched July 2, 2026.

  • Robostral Navigate, aimed at robotics/navigation applications, launched July 8, 2026.

  • In May 2026, Mistral introduced Vibe Remote Agents (a framework for deploying agents across distributed infrastructure) alongside the Mistral Medium 3.5 upgrade, and launched Physics AI, an initiative combining AI with physics research and simulation.

  • Mistral also announced an agreement to acquire Emmi AI, a physics-simulation company, extending Mistral beyond conversational AI into manufacturing, engineering, and industrial simulation.

7. Perplexity AI

Company Overview

Parameter

Details

Headquarters

San Francisco, California, USA

CEO

Aravind Srinivas

Founded

2022

Website

https://www.perplexity.ai/

Core Focus

Frontier AI, AI agents, search

Perplexity AI is considered one of the fastest-growing search and answer engine companies with a native, unified experience in AI. The response positions the company as a direct competitor to traditional search engines around citation-based answers, research workflows, and AI-assisted information discovery.

Subscription Plans Pricing

Plan

Price

Perplexity AI (Free)

$0

Perplexity Pro

$17/month (billed annually)

Perplexity Enterprise Pro

$40/month per seat or $400/year

Perplexity Enterprise Max

$325/month per seat or $3,250/year

Selected Perplexity AI Token Pricing

Model

Input (per 1M tokens)

Output (per 1M tokens)

Sonar

$1

$1

Sonar Pro

$3

$15

Sonar Reasoning Pro

$2

$8

Sonar Deep Research

$2

$8

Products Offering

Product

Description

Perplexity Search

AI answer engine with source-grounded responses

Perplexity Computer

General-purpose task system coordinating models, browsers, files, and connected tools

Enterprise Search

Business-focused research environment using internal and external data

Recent Developments

  • Perplexity Computer, launched February 2026, orchestrates work across roughly 19 models, breaks complex projects into subtasks, and operates browsers, files, and connected services, initially for Max subscribers, later expanded to Pro and Enterprise.

  • In June 2026, Perplexity added Deep Research, a command panel, conversation forking, inline actions, analytics APIs, and enterprise credit controls to the Computer product.

  • Perplexity is a founding member of Mistral's newly announced open industry initiative alongside Cursor, LangChain, Reflection AI, Black Forest Labs, Sarvam, and Thinking Machines Lab, reflecting its deepening role in the broader open AI ecosystem.

8. xAI

Company Overview

Parameter

Details

Headquarters

Palo Alto, California

CEO

Elon Musk

Website

https://x.ai/

Founded

2023

Focus

Foundation models, generative AI, conversational AI, AI infrastructure

xAI joined SpaceX in February 2026, and its official product site now uses the SpaceXAI brand. The company develops large language models under the Grok family, integrated with X (formerly Twitter), enterprise APIs, and AI infrastructure powered by its Colossus supercomputer. xAI focuses on reasoning, real-time information retrieval, multimodal AI, and large-scale AI training capabilities.

Subscription Plans Pricing

Plan

Price

Grok Free

$0

X Premium

$8/month

X Premium+

$40/month

SuperGrok

$30/month

SuperGrok Annual

$300/year

Selected xAI API Pricing

Model

Processing Model

Input (per 1M tokens)

Output (per 1M tokens)

Application

Chat API

grok-build-0.1

$1.00

$2.00

Intelligent model optimized for speed, combining frontier-level reasoning with enhanced search, grounding, and real-time information retrieval capabilities.

grok-4.3

$1.25

$2.50

grok-4.20-multi-agent-0309

$1.25

$2.50

grok-4.20-0309-reasoning

$1.25

$2.50

Imagine API

grok-imagine-image

$0.002 / img

$0.02/img

 

 

 

 

Advanced flagship model designed for complex reasoning, coding, enterprise workflows, and agentic AI applications.

grok-imagine-image-quality

$0.01 / img

$0.05/ img

grok-imagine-video

$0.01 / sec

$0.002 / img

$0.05/ img

grok-imagine-video-1.5-preview

$0.01 / img

$0.08/ img

Products Offering

Product

Description

Grok

A conversational AI assistant designed for reasoning, real-time knowledge access, coding assistance, and creative tasks.

Grok API

Developer platform that enables businesses to integrate Grok models into applications, workflows, and AI-powered products.

Grok for X

AI assistant integrated within X, providing content analysis, search, summarization, and conversational capabilities.

Recent Developments

  • In June 2026, Grok Voice was integrated into Vapi, an enterprise voice-AI platform, and reportedly topped a blind evaluation of competing voice providers.

  • In July 2026, SpaceXAI launched a Voice Agent Builder and expanded its voice product portfolio.

  • On May 6, 2026, SpaceXAI signed a compute-infrastructure agreement giving Anthropic (not xAI itself) access to the Colossus 1 data center, an unusual cross-lab infrastructure deal.

9. Meta AI

Company Overview

Parameter

Details

Headquarters

Menlo Park, California, United States

CEO

Mark Zuckerberg

Website

https://www.meta.ai/

AI research lineage

FAIR established in 2013

Focus

Proprietary frontier models, selective open-weight models, generative AI, enterprise AI, and consumer AI assistants

Meta AI is the artificial intelligence division of Meta Platforms. Following the formation of Meta Superintelligence Labs (MSL) in mid-2025, a response to Llama 4's mixed reception, Meta has pivoted its frontier line from the open-weight Llama family toward the new, largely proprietary Muse model family, while continuing selective open releases.

Selected Subscription Plans Pricing

Plan

Price

Meta AI

Free

Meta AI Premium

Coming Soon

Meta Verified

Starting at $11.99/month

AI Studio

Free

Meta Business AI Solutions

Custom Enterprise Pricing

Products Offering

Product

Description

Meta AI

AI assistant integrated across WhatsApp, Instagram, Facebook, Messenger, and standalone applications.

Llama Models

Open-weight large language models designed for developers, enterprises, and researchers.

AI Studio

Platform that allows creators and businesses to build customized AI assistants and digital agents.

Latest Model: Muse Spark 1.3 (launched September 2, 2026)

Meta quietly shipped Muse Spark 1.3 on September 2, 2026, at an aggressive blended price near $0.10 per million tokens, positioning it as a routing default for high-volume, user-facing agent workloads rather than a peak-reasoning flagship.

Context on Meta's Llama-to-Muse Pivot

  • Muse Spark, unveiled in April 2026, was the first model from Meta Superintelligence Labs and marked a controversial shift away from Meta AI's “open science” roots toward a closed, proprietary model, as competitors like Zhipu AI's GLM-5 and Alibaba's Qwen 3.6/3.8 lines had begun outpacing Llama 4 Maverick on general and coding benchmarks.

  • Muse Glimmer, released August 10, 2026, is a 30B-parameter dense multimodal model distilled from Muse Spark, released under Apache 2.0 with a 128K context window, designed for local, privacy-preserving, agentic deployment (coding, document analysis, personal assistants). On Meta's own benchmarks, it trades wins with Alibaba's Qwen3.6-27B depending on the task.

  • Muse Image launched, and Muse Video was previewed on July 7, 2026, followed by Muse Spark 1.1 (July 9, 2026) with a one-million-token context window, available in Meta AI's “Thinking” mode and via the Meta Model API in public preview.

10. Sarvam AI

Company Overview

Parameter

Details

Headquarters

Bengaluru, Karnataka, India

Leadership

Vivek Raghavan and Pratyush Kumar, Co-founders

Website

https://www.sarvam.ai/

Founded

2023

Focus

Generative AI, LLMs, Indic AI, voice AI

Sarvam AI is an India-based artificial intelligence company focused on building foundational AI models and platforms tailored for Indian languages and enterprises. The company develops multilingual large language models, speech technologies, and enterprise AI solutions that support India's linguistic diversity.

Subscription Plans Pricing

Plan

Price

Sarvam API Access

Usage-Based

Enterprise AI Solutions

Custom Pricing

Voice AI Solutions

Starting at ?15 per 10,000 characters

Sarvam AI API Pricing

API Service

Pricing

Unit

Application

Sarvam-105B

?4 / ?2.5 / ?16

Per 1M input / cached input/output tokens

Flagship large language model designed for advanced reasoning, multilingual conversations, enterprise AI assistants, and complex workflow automation.

Sarvam-30B

?2.5 / ?1.5 / ?10

Per 1M input / cached input/output tokens

Cost-efficient language model optimized for conversational AI, content generation, and enterprise productivity applications.

Sarvam Vision

?0.5

Per page

Vision API for image understanding, document processing, OCR, and digitization workflows.

Text-to-Speech (Bulbul v3)

?30

Per 10K characters

Advanced speech synthesis model that generates natural-sounding speech across multiple Indian languages and dialects.

Text-to-Speech (Bulbul v2)

?15

Per 10K characters

Cost-effective text-to-speech model for voice assistants, IVR systems, and content narration.

Speech-to-Text

?30

Per hour

Converts spoken audio into text with support for multiple Indian languages and accents.

Speech-to-Text with Diarization

?45

Per hour

Transcribes audio while identifying and separating individual speakers in conversations.

Speech-to-Text & Translate

?30

Per hour

Simultaneously transcribes and translates speech into target languages for multilingual communication.

Speech-to-Text, Translate & Diarization

?45

Per hour

End-to-end speech intelligence service that transcribes, translates, and identifies speakers in audio content.

Sarvam Translate V1

?20

Per 10K characters

Neural machine translation service supporting high-quality translation across Indian languages.

Translate Mayura V1

?20

Per 10K characters

Advanced multilingual translation model optimized for enterprise and government language workflows.

Transliterate

?20

Per 10K characters

Converts text between different Indic scripts while preserving pronunciation and meaning.

Language Identification

?3.5

Per 10K characters

Automatically detects and classifies the language of text inputs across multiple Indian languages.

Products Offering

Product

Description

Sarvam-M

Foundational multilingual large language model developed for Indian languages and enterprise AI applications.

Sarvam Agents

AI agents designed for customer support, workflow automation, and enterprise productivity use cases.

Sarvam Speech Platform

Speech-to-text and text-to-speech platform supporting numerous Indian languages and dialects.

Sarvam Translate

AI-powered translation platform enabling seamless communication across India's multilingual ecosystem.

Recent Developments

  • On June 15, 2026, Sarvam announced a $234 million first close of a targeted $300 million Series B round at a $1.5 billion post-money valuation, aimed at frontier-model research, large-scale compute access, and enterprise/government deployment.

  • Sarvam moved Saaras v3 into its speech-to-text stack and promoted Bulbul v3 to stable production status.

  • Sarvam is a founding member of Mistral AI's newly announced open industry coalition (alongside Cursor, LangChain, Perplexity, Black Forest Labs, and others), reflecting deepening cross-lab collaboration.

  • In February 2026, Sarvam signed sovereign AI partnerships with the governments of Odisha and Tamil Nadu to build large-scale compute capacity and government-embedded AI systems.

11. Kimi AI (Moonshot AI)

Company Overview

Parameter

Details

Company

Moonshot AI

AI Assistant

Kimi

Headquarters

Beijing, China

Core Focus

Multimodal AI, reasoning, coding, AI agents, long-context processing

Key Products

Kimi K3, Kimi K2.6, Kimi Agent, Kimi Code, Kimi API

Website

https://www.kimi.com/

Kimi, developed by Moonshot AI, offers web search, deep thinking, multimodal reasoning, long-context conversation, agentic workflows, and developer APIs. Recent product expansion has focused on coding, office productivity, and multi-agent execution.

Subscription Plans Pricing

Plan

Monthly Price

Agent Credits

Concurrent Tasks

Adagio

$0

6

1

Moderato

$19

60

2

Allegretto

$39

150

2

Allegro

$99

360

4

Vivace

$199

720

4

Selected Products Offering

Product

Description

Kimi K3

Flagship 2.8-trillion-parameter, natively multimodal model with 1M-token context, built for long-horizon coding, knowledge work, and reasoning

Kimi K2.6

Open-source model with major upgrades in agentic coding, long-context reasoning, and Agent Swarm

Kimi Agent

Agentic capabilities for research, document processing, presentations, and spreadsheets

Agent Swarm

Multi-agent coordination for complex, parallel tasks (up to 100 sub-agents)

Kimi API

Developer platform with usage-based token pricing

Recent Developments

  • Kimi K3 launched in July 2026 as a 2.8-trillion-parameter, natively multimodal model with a 1-million-token context window, targeting long-horizon coding, knowledge work, and deep reasoning, positioning Moonshot AI competitively against Western frontier labs on cost-efficiency.

  • Kimi K2.6, released and open-sourced in April 2026, brought major upgrades to agentic coding, long-context reasoning, long-horizon execution, and Agent Swarm.

  • Kimi K2.5 (January 2026) was an open-source native multimodal model supporting up to 100 parallel sub-agents and 1,500 tool calls, and expanded Kimi Agent's office productivity capabilities (Word, PDF, Excel, presentations).

Kimi remains one of the most aggressively priced frontier-adjacent labs, competing with DeepSeek, Alibaba's Qwen, and Zhipu's GLM line in China's fast-moving open-weight ecosystem.

Have questions?

More Thought Articles

Aug 26, 2026

The Rise of Locally Manufactured Battery Separators in India’s EV Ecosystem

India’s battery industry is progressing from pack assembly toward domestic cell and material manufacturing, creating opportunities for lithium-ion battery separators. Policy support, ACC capacity expansion and R&D initiatives are strengthening the ecosystem. However, technical complexity, manufacturing yields, qualification requirements and economics mean localisation will likely advance gradually through coatings, testing and finishing.

Click here to read
Aug 25, 2026

Why Hub Motors Are Becoming a Game Changer for India’s Electric Two-Wheeler Manufacturers

India’s electric two-wheeler market is increasingly prioritizing cost, efficiency, localization, and reliability. Hub motors are gaining relevance for commuter, fleet, and urban scooters because their simpler drivetrain architecture can reduce mechanical complexity and maintenance exposure. However, mid-drive systems will remain important for premium and performance-focused electric two-wheelers, supporting differentiated applications.

Click here to read
Aug 24, 2026

How Advanced Battery Thermal Management Is Transforming Cold-Weather EV Performance in India

Advanced battery thermal management is becoming increasingly important for India's EV market as vehicles operate across diverse temperature conditions. Adaptive heating, cooling, preconditioning, liquid thermal systems and intelligent software can improve range consistency, charging performance and battery durability, helping automakers deliver reliable, cost-effective EV performance across regions and vehicle segments.

Click here to read