Kimi K2
Moonshot AI
Kimi K2 launched by Moonshot AI. Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.

Track model launches, capability updates, pricing changes, and benchmark milestones across the AI industry.
Moonshot AI
Kimi K2 launched by Moonshot AI. Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.
Moonshot AI
API endpoint `kimi-k2` available for Kimi K2.
Moonshot AI
Initial pricing set at $0.60/M input and $2.50/M output tokens.
Moonshot AI
Kimi K2 supports function calling and structured tool use.
Moonshot AI
Kimi K2 introduces extended reasoning and chain-of-thought capabilities.
Anthropic
Claude 4 Sonnet launched by Anthropic. Latest Sonnet with improved reasoning and tool use.
Anthropic
API endpoint `claude-sonnet-4-20250514` available for Claude 4 Sonnet.
Anthropic
Initial pricing set at $3.00/M input and $15.00/M output tokens.
Anthropic
Claude 4 Sonnet adds native vision and image understanding capabilities.
Anthropic
Claude 4 Sonnet supports function calling and structured tool use.
Anthropic
Claude 4 Sonnet introduces extended reasoning and chain-of-thought capabilities.
Anthropic
Claude 4 Sonnet achieves 96/100 coding benchmark score — top-tier performance.
Anthropic
Claude 4 Opus launched by Anthropic. Anthropic flagship for the most demanding tasks.
Anthropic
API endpoint `claude-opus-4-20250514` available for Claude 4 Opus.
Anthropic
Initial pricing set at $15.00/M input and $75.00/M output tokens.
Anthropic
Claude 4 Opus adds native vision and image understanding capabilities.
Anthropic
Claude 4 Opus supports function calling and structured tool use.
Anthropic
Claude 4 Opus introduces extended reasoning and chain-of-thought capabilities.
Anthropic
Claude 4 Opus achieves 97/100 coding benchmark score — top-tier performance.
Anthropic
Claude 4 Haiku launched by Anthropic. Next-gen Haiku with Claude 4 intelligence at low cost.
Anthropic
API endpoint `claude-haiku-4-20250514` available for Claude 4 Haiku.
Anthropic
Initial pricing set at $0.80/M input and $4.00/M output tokens.
Anthropic
Claude 4 Haiku adds native vision and image understanding capabilities.
Anthropic
Claude 4 Haiku supports function calling and structured tool use.
Gemini 2.5 Flash launched by Google. Flash model with thinking capabilities and top multimodal scores.
API endpoint `gemini-2.5-flash-preview-05-20` available for Gemini 2.5 Flash.
Initial pricing set at $0.15/M input and $0.60/M output tokens.
Gemini 2.5 Flash ships with 1M token context window.
Gemini 2.5 Flash adds native vision and image understanding capabilities.
Gemini 2.5 Flash supports audio input and multimodal voice workflows.
Gemini 2.5 Flash supports function calling and structured tool use.
Gemini 2.5 Flash introduces extended reasoning and chain-of-thought capabilities.
Gemini 2.5 Flash achieves 94/100 multimodal benchmark score — top-tier performance.
OpenAI
OpenAI o4 Mini launched by OpenAI. Latest compact reasoning model with vision support.
OpenAI
API endpoint `o4-mini` available for OpenAI o4 Mini.
OpenAI
Initial pricing set at $1.10/M input and $4.40/M output tokens.
OpenAI
OpenAI o4 Mini adds native vision and image understanding capabilities.
OpenAI
OpenAI o4 Mini supports function calling and structured tool use.
OpenAI
OpenAI o4 Mini introduces extended reasoning and chain-of-thought capabilities.
OpenAI
OpenAI o4 Mini achieves 93/100 reasoning benchmark score — top-tier performance.
OpenAI
OpenAI o3 launched by OpenAI. Full o3 reasoning model with vision and tool use.
OpenAI
API endpoint `o3` available for OpenAI o3.
OpenAI
Initial pricing set at $10.00/M input and $40.00/M output tokens.
OpenAI
OpenAI o3 supports function calling and structured tool use.
OpenAI
OpenAI o3 introduces extended reasoning and chain-of-thought capabilities.
OpenAI
OpenAI o3 achieves 97/100 reasoning benchmark score — top-tier performance.
OpenAI
GPT-4.1 launched by OpenAI. Next-gen GPT with 1M context and improved instruction following.
OpenAI
API endpoint `gpt-4.1` available for GPT-4.1.
OpenAI
Initial pricing set at $2.00/M input and $8.00/M output tokens.
OpenAI
GPT-4.1 ships with 1M token context window.
OpenAI
GPT-4.1 adds native vision and image understanding capabilities.
OpenAI
GPT-4.1 supports audio input and multimodal voice workflows.
OpenAI
GPT-4.1 supports function calling and structured tool use.
OpenAI
GPT-4.1 achieves 94/100 coding benchmark score — top-tier performance.
OpenAI
GPT-4.1 Mini launched by OpenAI. Mini variant of GPT-4.1 for scalable deployments.
OpenAI
API endpoint `gpt-4.1-mini` available for GPT-4.1 Mini.
OpenAI
Initial pricing set at $0.40/M input and $1.60/M output tokens.
OpenAI
GPT-4.1 Mini ships with 1M token context window.
OpenAI
GPT-4.1 Mini adds native vision and image understanding capabilities.
OpenAI
GPT-4.1 Mini supports audio input and multimodal voice workflows.
OpenAI
GPT-4.1 Mini supports function calling and structured tool use.
OpenAI
GPT-4.1 Nano launched by OpenAI. Ultra-low-cost GPT-4.1 tier for classification and extraction.
OpenAI
API endpoint `gpt-4.1-nano` available for GPT-4.1 Nano.
OpenAI
Initial pricing set at $0.10/M input and $0.40/M output tokens.
OpenAI
GPT-4.1 Nano ships with 1M token context window.
OpenAI
GPT-4.1 Nano adds native vision and image understanding capabilities.
OpenAI
GPT-4.1 Nano supports function calling and structured tool use.
Gemini 2.5 Pro launched by Google. Google flagship with 1M context and deep reasoning.
API endpoint `gemini-2.5-pro-preview-06-05` available for Gemini 2.5 Pro.
Initial pricing set at $1.25/M input and $10.00/M output tokens.
Gemini 2.5 Pro ships with 1M token context window.
Gemini 2.5 Pro adds native vision and image understanding capabilities.
Gemini 2.5 Pro supports audio input and multimodal voice workflows.
Gemini 2.5 Pro supports function calling and structured tool use.
Gemini 2.5 Pro introduces extended reasoning and chain-of-thought capabilities.
Gemini 2.5 Pro achieves 97/100 multimodal benchmark score — top-tier performance.
DeepSeek
DeepSeek V3 0324 launched by DeepSeek. Updated V3 with improved instruction following and 128K context.
DeepSeek
API endpoint `deepseek-chat` available for DeepSeek V3 0324.
DeepSeek
Initial pricing set at $0.27/M input and $1.10/M output tokens.
DeepSeek
DeepSeek V3 0324 supports function calling and structured tool use.
DeepSeek
DeepSeek V3 0324 achieves 92/100 coding benchmark score — top-tier performance.
DeepSeek
DeepSeek V3 0324 available for fine-tuning and custom deployment via DeepSeek.
Cohere
Command A launched by Cohere. Cohere flagship agentic model for tool use and search.
Cohere
API endpoint `command-a-03-2025` available for Command A.
Cohere
Initial pricing set at $2.50/M input and $10.00/M output tokens.
Cohere
Command A adds native vision and image understanding capabilities.
Cohere
Command A supports function calling and structured tool use.
Meta
Llama 4 Maverick launched by Meta. Open multimodal Llama with 1M context window.
Meta
API endpoint `llama-4-maverick` available for Llama 4 Maverick.
Meta
Initial pricing set at $0.20/M input and $0.60/M output tokens.
Meta
Llama 4 Maverick ships with 1M token context window.
Meta
Llama 4 Maverick adds native vision and image understanding capabilities.
Meta
Llama 4 Maverick supports function calling and structured tool use.
Meta
Llama 4 Maverick available for fine-tuning and custom deployment via Meta.
Meta
Llama 4 Scout launched by Meta. Llama 4 with industry-leading 10M context for documents.
Meta
API endpoint `llama-4-scout` available for Llama 4 Scout.
Meta
Initial pricing set at $0.15/M input and $0.50/M output tokens.
Meta
Llama 4 Scout ships with 10M token context window.
Meta
Llama 4 Scout adds native vision and image understanding capabilities.
Meta
Llama 4 Scout supports function calling and structured tool use.
Meta
Llama 4 Scout available for fine-tuning and custom deployment via Meta.
Anthropic
Claude 3.7 Sonnet launched by Anthropic. Sonnet with extended thinking for harder problems.
Anthropic
API endpoint `claude-3-7-sonnet-20250219` available for Claude 3.7 Sonnet.
Anthropic
Initial pricing set at $3.00/M input and $15.00/M output tokens.
Anthropic
Claude 3.7 Sonnet adds native vision and image understanding capabilities.
Anthropic
Claude 3.7 Sonnet supports function calling and structured tool use.
Anthropic
Claude 3.7 Sonnet introduces extended reasoning and chain-of-thought capabilities.
Anthropic
Claude 3.7 Sonnet achieves 95/100 coding benchmark score — top-tier performance.
xAI
Grok 3 launched by xAI. xAI flagship with strong reasoning and tool use.
xAI
API endpoint `grok-3` available for Grok 3.
xAI
Initial pricing set at $3.00/M input and $15.00/M output tokens.
xAI
Grok 3 adds native vision and image understanding capabilities.
xAI
Grok 3 supports function calling and structured tool use.
xAI
Grok 3 introduces extended reasoning and chain-of-thought capabilities.
xAI
Grok 3 achieves 92/100 reasoning benchmark score — top-tier performance.
xAI
Grok 3 Mini launched by xAI. Efficient Grok 3 variant for high-volume use.
xAI
API endpoint `grok-3-mini` available for Grok 3 Mini.
xAI
Initial pricing set at $0.30/M input and $0.50/M output tokens.
xAI
Grok 3 Mini adds native vision and image understanding capabilities.
xAI
Grok 3 Mini supports function calling and structured tool use.
xAI
Grok 3 Mini introduces extended reasoning and chain-of-thought capabilities.
Gemini 2.0 Flash Lite launched by Google. Ultra-cheap Gemini for simple tasks at scale.
API endpoint `gemini-2.0-flash-lite` available for Gemini 2.0 Flash Lite.
Initial pricing set at $0.075/M input and $0.30/M output tokens.
Gemini 2.0 Flash Lite ships with 1M token context window.
Gemini 2.0 Flash Lite adds native vision and image understanding capabilities.
Gemini 2.0 Flash Lite supports function calling and structured tool use.
OpenAI
OpenAI o3 Mini launched by OpenAI. Efficient reasoning model in the o3 family.
OpenAI
API endpoint `o3-mini` available for OpenAI o3 Mini.
OpenAI
Initial pricing set at $1.10/M input and $4.40/M output tokens.
OpenAI
OpenAI o3 Mini supports function calling and structured tool use.
OpenAI
OpenAI o3 Mini introduces extended reasoning and chain-of-thought capabilities.
Mistral
Mistral Small launched by Mistral. Compact Mistral for cost-sensitive applications.
Mistral
API endpoint `mistral-small-2501` available for Mistral Small.
Mistral
Initial pricing set at $0.10/M input and $0.30/M output tokens.
Mistral
Mistral Small supports function calling and structured tool use.
Mistral
Mistral Small available for fine-tuning and custom deployment via Mistral.
DeepSeek
DeepSeek R1 launched by DeepSeek. Open reasoning model with o1-class performance at fraction of cost.
DeepSeek
API endpoint `deepseek-reasoner` available for DeepSeek R1.
DeepSeek
Initial pricing set at $0.55/M input and $2.19/M output tokens.
DeepSeek
DeepSeek R1 supports function calling and structured tool use.
DeepSeek
DeepSeek R1 introduces extended reasoning and chain-of-thought capabilities.
DeepSeek
DeepSeek R1 achieves 95/100 reasoning benchmark score — top-tier performance.
DeepSeek
DeepSeek R1 available for fine-tuning and custom deployment via DeepSeek.
Mistral
Codestral launched by Mistral. Code-specialized Mistral model for IDE and agent use.
Mistral
API endpoint `codestral-2501` available for Codestral.
Mistral
Initial pricing set at $0.30/M input and $0.90/M output tokens.
Mistral
Codestral supports function calling and structured tool use.
Mistral
Codestral achieves 94/100 coding benchmark score — top-tier performance.
Mistral
Codestral available for fine-tuning and custom deployment via Mistral.