DeepSeek V3
DeepSeek
DeepSeek V3 launched by DeepSeek. Highly efficient MoE model rivaling frontier closed models.

Track model launches, capability updates, pricing changes, and benchmark milestones across the AI industry.
DeepSeek
DeepSeek V3 launched by DeepSeek. Highly efficient MoE model rivaling frontier closed models.
DeepSeek
API endpoint `deepseek-chat` available for DeepSeek V3.
DeepSeek
Initial pricing set at $0.27/M input and $1.10/M output tokens.
DeepSeek
DeepSeek V3 supports function calling and structured tool use.
DeepSeek
DeepSeek V3 available for fine-tuning and custom deployment via DeepSeek.
Gemini 2.0 Flash launched by Google. Fast multimodal Gemini with native audio and vision.
API endpoint `gemini-2.0-flash` available for Gemini 2.0 Flash.
Initial pricing set at $0.10/M input and $0.40/M output tokens.
Gemini 2.0 Flash ships with 1M token context window.
Gemini 2.0 Flash adds native vision and image understanding capabilities.
Gemini 2.0 Flash supports audio input and multimodal voice workflows.
Gemini 2.0 Flash supports function calling and structured tool use.
Gemini 2.0 Flash achieves 92/100 multimodal benchmark score — top-tier performance.
Meta
Llama 3.3 70B launched by Meta. Open-weight 70B model competitive with larger closed models.
Meta
API endpoint `llama-3.3-70b-instruct` available for Llama 3.3 70B.
Meta
Initial pricing set at $0.23/M input and $0.40/M output tokens.
Meta
Llama 3.3 70B supports function calling and structured tool use.
Meta
Llama 3.3 70B available for fine-tuning and custom deployment via Meta.
OpenAI
OpenAI o1 launched by OpenAI. Reasoning model using internal chain-of-thought.
OpenAI
API endpoint `o1` available for OpenAI o1.
OpenAI
Initial pricing set at $15.00/M input and $60.00/M output tokens.
OpenAI
OpenAI o1 supports function calling and structured tool use.
OpenAI
OpenAI o1 introduces extended reasoning and chain-of-thought capabilities.
OpenAI
OpenAI o1 achieves 96/100 reasoning benchmark score — top-tier performance.
Mistral
Mistral Large launched by Mistral. Mistral flagship for enterprise European deployments.
Mistral
API endpoint `mistral-large-2411` available for Mistral Large.
Mistral
Initial pricing set at $2.00/M input and $6.00/M output tokens.
Mistral
Mistral Large adds native vision and image understanding capabilities.
Mistral
Mistral Large supports function calling and structured tool use.
Mistral
Mistral Large available for fine-tuning and custom deployment via Mistral.
Mistral
Pixtral Large launched by Mistral. Mistral multimodal model with strong vision understanding.
Mistral
API endpoint `pixtral-large-2411` available for Pixtral Large.
Mistral
Initial pricing set at $2.00/M input and $6.00/M output tokens.
Mistral
Pixtral Large adds native vision and image understanding capabilities.
Mistral
Pixtral Large supports function calling and structured tool use.
Mistral
Pixtral Large achieves 93/100 multimodal benchmark score — top-tier performance.
Mistral
Pixtral Large available for fine-tuning and custom deployment via Mistral.
Anthropic
Claude 3.5 Haiku launched by Anthropic. Fast and affordable Claude for high-throughput workloads.
Anthropic
API endpoint `claude-3-5-haiku-20241022` available for Claude 3.5 Haiku.
Anthropic
Initial pricing set at $0.80/M input and $4.00/M output tokens.
Anthropic
Claude 3.5 Haiku adds native vision and image understanding capabilities.
Anthropic
Claude 3.5 Haiku supports function calling and structured tool use.
Mistral
Ministral 8B launched by Mistral. Tiny on-device capable Mistral for edge deployments.
Mistral
API endpoint `ministral-8b-2410` available for Ministral 8B.
Mistral
Initial pricing set at $0.10/M input and $0.10/M output tokens.
Mistral
Ministral 8B supports function calling and structured tool use.
Mistral
Ministral 8B available for fine-tuning and custom deployment via Mistral.
xAI
Grok 2 launched by xAI. xAI general model with real-time X data integration.
xAI
API endpoint `grok-2` available for Grok 2.
xAI
Initial pricing set at $2.00/M input and $10.00/M output tokens.
xAI
Grok 2 adds native vision and image understanding capabilities.
xAI
Grok 2 supports function calling and structured tool use.
Meta
Llama 3.1 405B launched by Meta. Largest open-weight Llama for maximum capability.
Meta
API endpoint `llama-3.1-405b-instruct` available for Llama 3.1 405B.
Meta
Initial pricing set at $0.35/M input and $0.70/M output tokens.
Meta
Llama 3.1 405B supports function calling and structured tool use.
Meta
Llama 3.1 405B available for fine-tuning and custom deployment via Meta.
OpenAI
GPT-4o Mini launched by OpenAI. Cost-efficient small model for high-volume applications.
OpenAI
API endpoint `gpt-4o-mini` available for GPT-4o Mini.
OpenAI
Initial pricing set at $0.15/M input and $0.60/M output tokens.
OpenAI
GPT-4o Mini adds native vision and image understanding capabilities.
OpenAI
GPT-4o Mini supports function calling and structured tool use.
Anthropic
Claude 3.5 Sonnet launched by Anthropic. Balanced Claude model excelling at coding and analysis.
Anthropic
API endpoint `claude-3-5-sonnet-20241022` available for Claude 3.5 Sonnet.
Anthropic
Initial pricing set at $3.00/M input and $15.00/M output tokens.
Anthropic
Claude 3.5 Sonnet adds native vision and image understanding capabilities.
Anthropic
Claude 3.5 Sonnet supports function calling and structured tool use.
Anthropic
Claude 3.5 Sonnet achieves 94/100 writing benchmark score — top-tier performance.
DeepSeek
DeepSeek Coder V2 launched by DeepSeek. Code-specialized DeepSeek for programming tasks.
DeepSeek
API endpoint `deepseek-coder` available for DeepSeek Coder V2.
DeepSeek
Initial pricing set at $0.14/M input and $0.28/M output tokens.
DeepSeek
DeepSeek Coder V2 supports function calling and structured tool use.
DeepSeek
DeepSeek Coder V2 achieves 93/100 coding benchmark score — top-tier performance.
DeepSeek
DeepSeek Coder V2 available for fine-tuning and custom deployment via DeepSeek.
OpenAI
GPT-4o launched by OpenAI. Flagship multimodal model balancing speed, intelligence, and cost.
OpenAI
API endpoint `gpt-4o` available for GPT-4o.
OpenAI
Initial pricing set at $2.50/M input and $10.00/M output tokens.
OpenAI
GPT-4o adds native vision and image understanding capabilities.
OpenAI
GPT-4o supports audio input and multimodal voice workflows.
OpenAI
GPT-4o supports function calling and structured tool use.
OpenAI
GPT-4o achieves 95/100 multimodal benchmark score — top-tier performance.
Cohere
Command R+ launched by Cohere. Cohere model optimized for RAG and enterprise search.
Cohere
API endpoint `command-r-plus` available for Command R+.
Cohere
Initial pricing set at $2.50/M input and $10.00/M output tokens.
Cohere
Command R+ supports function calling and structured tool use.
Cohere
Command R launched by Cohere. Affordable Cohere model for retrieval-augmented workflows.
Cohere
API endpoint `command-r` available for Command R.
Cohere
Initial pricing set at $0.15/M input and $0.60/M output tokens.
Cohere
Command R supports function calling and structured tool use.
Anthropic
Claude 3 Opus launched by Anthropic. Previous-gen Opus still strong for writing-heavy workloads.
Anthropic
API endpoint `claude-3-opus-20240229` available for Claude 3 Opus.
Anthropic
Initial pricing set at $15.00/M input and $75.00/M output tokens.
Anthropic
Claude 3 Opus adds native vision and image understanding capabilities.
Anthropic
Claude 3 Opus supports function calling and structured tool use.
Anthropic
Claude 3 Opus achieves 93/100 writing benchmark score — top-tier performance.
Gemini 1.5 Pro launched by Google. Proven Gemini with industry-leading 2M context window.
API endpoint `gemini-1.5-pro` available for Gemini 1.5 Pro.
Initial pricing set at $1.25/M input and $5.00/M output tokens.
Gemini 1.5 Pro ships with 2.1M token context window.
Gemini 1.5 Pro adds native vision and image understanding capabilities.
Gemini 1.5 Pro supports function calling and structured tool use.