SummitGovOnline

AI Release Timeline

Track model launches, capability updates, pricing changes, and benchmark milestones across the AI industry.

2025

MA

Kimi K2

Model Release

Moonshot AI

Kimi K2 launched by Moonshot AI. Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.

MA

Kimi K2

Pricing Update

Moonshot AI

Initial pricing set at $0.60/M input and $2.50/M output tokens.

MA

Kimi K2

Function Calling

Moonshot AI

Kimi K2 supports function calling and structured tool use.

MA

Kimi K2

Reasoning Model

Moonshot AI

Kimi K2 introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Sonnet

Model Release

Anthropic

Claude 4 Sonnet launched by Anthropic. Latest Sonnet with improved reasoning and tool use.

A

Claude 4 Sonnet

API Release

Anthropic

API endpoint `claude-sonnet-4-20250514` available for Claude 4 Sonnet.

A

Claude 4 Sonnet

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude 4 Sonnet

Vision Support

Anthropic

Claude 4 Sonnet adds native vision and image understanding capabilities.

A

Claude 4 Sonnet

Function Calling

Anthropic

Claude 4 Sonnet supports function calling and structured tool use.

A

Claude 4 Sonnet

Reasoning Model

Anthropic

Claude 4 Sonnet introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Sonnet

Benchmark Milestone

Anthropic

Claude 4 Sonnet achieves 96/100 coding benchmark score — top-tier performance.

A

Claude 4 Opus

Model Release

Anthropic

Claude 4 Opus launched by Anthropic. Anthropic flagship for the most demanding tasks.

A

Claude 4 Opus

API Release

Anthropic

API endpoint `claude-opus-4-20250514` available for Claude 4 Opus.

A

Claude 4 Opus

Pricing Update

Anthropic

Initial pricing set at $15.00/M input and $75.00/M output tokens.

A

Claude 4 Opus

Vision Support

Anthropic

Claude 4 Opus adds native vision and image understanding capabilities.

A

Claude 4 Opus

Function Calling

Anthropic

Claude 4 Opus supports function calling and structured tool use.

A

Claude 4 Opus

Reasoning Model

Anthropic

Claude 4 Opus introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Opus

Benchmark Milestone

Anthropic

Claude 4 Opus achieves 97/100 coding benchmark score — top-tier performance.

A

Claude 4 Haiku

Model Release

Anthropic

Claude 4 Haiku launched by Anthropic. Next-gen Haiku with Claude 4 intelligence at low cost.

A

Claude 4 Haiku

API Release

Anthropic

API endpoint `claude-haiku-4-20250514` available for Claude 4 Haiku.

A

Claude 4 Haiku

Pricing Update

Anthropic

Initial pricing set at $0.80/M input and $4.00/M output tokens.

A

Claude 4 Haiku

Vision Support

Anthropic

Claude 4 Haiku adds native vision and image understanding capabilities.

A

Claude 4 Haiku

Function Calling

Anthropic

Claude 4 Haiku supports function calling and structured tool use.

G

Gemini 2.5 Flash

Model Release

Google

Gemini 2.5 Flash launched by Google. Flash model with thinking capabilities and top multimodal scores.

G

Gemini 2.5 Flash

API Release

Google

API endpoint `gemini-2.5-flash-preview-05-20` available for Gemini 2.5 Flash.

G

Gemini 2.5 Flash

Pricing Update

Google

Initial pricing set at $0.15/M input and $0.60/M output tokens.

G

Gemini 2.5 Flash

Context Window Increase

Google

Gemini 2.5 Flash ships with 1M token context window.

G

Gemini 2.5 Flash

Vision Support

Google

Gemini 2.5 Flash adds native vision and image understanding capabilities.

G

Gemini 2.5 Flash

Audio Support

Google

Gemini 2.5 Flash supports audio input and multimodal voice workflows.

G

Gemini 2.5 Flash

Function Calling

Google

Gemini 2.5 Flash supports function calling and structured tool use.

G

Gemini 2.5 Flash

Reasoning Model

Google

Gemini 2.5 Flash introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.5 Flash

Benchmark Milestone

Google

Gemini 2.5 Flash achieves 94/100 multimodal benchmark score — top-tier performance.

O

OpenAI o4 Mini

Model Release

OpenAI

OpenAI o4 Mini launched by OpenAI. Latest compact reasoning model with vision support.

O

OpenAI o4 Mini

Pricing Update

OpenAI

Initial pricing set at $1.10/M input and $4.40/M output tokens.

O

OpenAI o4 Mini

Vision Support

OpenAI

OpenAI o4 Mini adds native vision and image understanding capabilities.

O

OpenAI o4 Mini

Function Calling

OpenAI

OpenAI o4 Mini supports function calling and structured tool use.

O

OpenAI o4 Mini

Reasoning Model

OpenAI

OpenAI o4 Mini introduces extended reasoning and chain-of-thought capabilities.

O

OpenAI o4 Mini

Benchmark Milestone

OpenAI

OpenAI o4 Mini achieves 93/100 reasoning benchmark score — top-tier performance.

O

OpenAI o3

Model Release

OpenAI

OpenAI o3 launched by OpenAI. Full o3 reasoning model with vision and tool use.

O

OpenAI o3

Pricing Update

OpenAI

Initial pricing set at $10.00/M input and $40.00/M output tokens.

O

OpenAI o3

Function Calling

OpenAI

OpenAI o3 supports function calling and structured tool use.

O

OpenAI o3

Reasoning Model

OpenAI

OpenAI o3 introduces extended reasoning and chain-of-thought capabilities.

O

OpenAI o3

Benchmark Milestone

OpenAI

OpenAI o3 achieves 97/100 reasoning benchmark score — top-tier performance.

O

GPT-4.1

Model Release

OpenAI

GPT-4.1 launched by OpenAI. Next-gen GPT with 1M context and improved instruction following.

O

GPT-4.1

Vision Support

OpenAI

GPT-4.1 adds native vision and image understanding capabilities.

O

GPT-4.1

Benchmark Milestone

OpenAI

GPT-4.1 achieves 94/100 coding benchmark score — top-tier performance.

O

GPT-4.1 Mini

Model Release

OpenAI

GPT-4.1 Mini launched by OpenAI. Mini variant of GPT-4.1 for scalable deployments.

O

GPT-4.1 Mini

Pricing Update

OpenAI

Initial pricing set at $0.40/M input and $1.60/M output tokens.

O

GPT-4.1 Mini

Vision Support

OpenAI

GPT-4.1 Mini adds native vision and image understanding capabilities.

O

GPT-4.1 Mini

Audio Support

OpenAI

GPT-4.1 Mini supports audio input and multimodal voice workflows.

O

GPT-4.1 Mini

Function Calling

OpenAI

GPT-4.1 Mini supports function calling and structured tool use.

O

GPT-4.1 Nano

Model Release

OpenAI

GPT-4.1 Nano launched by OpenAI. Ultra-low-cost GPT-4.1 tier for classification and extraction.

O

GPT-4.1 Nano

Pricing Update

OpenAI

Initial pricing set at $0.10/M input and $0.40/M output tokens.

O

GPT-4.1 Nano

Vision Support

OpenAI

GPT-4.1 Nano adds native vision and image understanding capabilities.

O

GPT-4.1 Nano

Function Calling

OpenAI

GPT-4.1 Nano supports function calling and structured tool use.

G

Gemini 2.5 Pro

Model Release

Google

Gemini 2.5 Pro launched by Google. Google flagship with 1M context and deep reasoning.

G

Gemini 2.5 Pro

API Release

Google

API endpoint `gemini-2.5-pro-preview-06-05` available for Gemini 2.5 Pro.

G

Gemini 2.5 Pro

Pricing Update

Google

Initial pricing set at $1.25/M input and $10.00/M output tokens.

G

Gemini 2.5 Pro

Context Window Increase

Google

Gemini 2.5 Pro ships with 1M token context window.

G

Gemini 2.5 Pro

Vision Support

Google

Gemini 2.5 Pro adds native vision and image understanding capabilities.

G

Gemini 2.5 Pro

Audio Support

Google

Gemini 2.5 Pro supports audio input and multimodal voice workflows.

G

Gemini 2.5 Pro

Function Calling

Google

Gemini 2.5 Pro supports function calling and structured tool use.

G

Gemini 2.5 Pro

Reasoning Model

Google

Gemini 2.5 Pro introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.5 Pro

Benchmark Milestone

Google

Gemini 2.5 Pro achieves 97/100 multimodal benchmark score — top-tier performance.

D

DeepSeek V3 0324

Model Release

DeepSeek

DeepSeek V3 0324 launched by DeepSeek. Updated V3 with improved instruction following and 128K context.

D

DeepSeek V3 0324

API Release

DeepSeek

API endpoint `deepseek-chat` available for DeepSeek V3 0324.

D

DeepSeek V3 0324

Pricing Update

DeepSeek

Initial pricing set at $0.27/M input and $1.10/M output tokens.

D

DeepSeek V3 0324

Function Calling

DeepSeek

DeepSeek V3 0324 supports function calling and structured tool use.

D

DeepSeek V3 0324

Benchmark Milestone

DeepSeek

DeepSeek V3 0324 achieves 92/100 coding benchmark score — top-tier performance.

D

DeepSeek V3 0324

Fine-tuning Release

DeepSeek

DeepSeek V3 0324 available for fine-tuning and custom deployment via DeepSeek.

C

Command A

Model Release

Cohere

Command A launched by Cohere. Cohere flagship agentic model for tool use and search.

C

Command A

Pricing Update

Cohere

Initial pricing set at $2.50/M input and $10.00/M output tokens.

C

Command A

Vision Support

Cohere

Command A adds native vision and image understanding capabilities.

C

Command A

Function Calling

Cohere

Command A supports function calling and structured tool use.

M

Llama 4 Maverick

Model Release

Meta

Llama 4 Maverick launched by Meta. Open multimodal Llama with 1M context window.

M

Llama 4 Maverick

API Release

Meta

API endpoint `llama-4-maverick` available for Llama 4 Maverick.

M

Llama 4 Maverick

Pricing Update

Meta

Initial pricing set at $0.20/M input and $0.60/M output tokens.

M

Llama 4 Maverick

Context Window Increase

Meta

Llama 4 Maverick ships with 1M token context window.

M

Llama 4 Maverick

Vision Support

Meta

Llama 4 Maverick adds native vision and image understanding capabilities.

M

Llama 4 Maverick

Function Calling

Meta

Llama 4 Maverick supports function calling and structured tool use.

M

Llama 4 Maverick

Fine-tuning Release

Meta

Llama 4 Maverick available for fine-tuning and custom deployment via Meta.

M

Llama 4 Scout

Model Release

Meta

Llama 4 Scout launched by Meta. Llama 4 with industry-leading 10M context for documents.

M

Llama 4 Scout

Pricing Update

Meta

Initial pricing set at $0.15/M input and $0.50/M output tokens.

M

Llama 4 Scout

Vision Support

Meta

Llama 4 Scout adds native vision and image understanding capabilities.

M

Llama 4 Scout

Function Calling

Meta

Llama 4 Scout supports function calling and structured tool use.

M

Llama 4 Scout

Fine-tuning Release

Meta

Llama 4 Scout available for fine-tuning and custom deployment via Meta.

A

Claude 3.7 Sonnet

Model Release

Anthropic

Claude 3.7 Sonnet launched by Anthropic. Sonnet with extended thinking for harder problems.

A

Claude 3.7 Sonnet

API Release

Anthropic

API endpoint `claude-3-7-sonnet-20250219` available for Claude 3.7 Sonnet.

A

Claude 3.7 Sonnet

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude 3.7 Sonnet

Vision Support

Anthropic

Claude 3.7 Sonnet adds native vision and image understanding capabilities.

A

Claude 3.7 Sonnet

Function Calling

Anthropic

Claude 3.7 Sonnet supports function calling and structured tool use.

A

Claude 3.7 Sonnet

Reasoning Model

Anthropic

Claude 3.7 Sonnet introduces extended reasoning and chain-of-thought capabilities.

A

Claude 3.7 Sonnet

Benchmark Milestone

Anthropic

Claude 3.7 Sonnet achieves 95/100 coding benchmark score — top-tier performance.

X

Grok 3

Model Release

xAI

Grok 3 launched by xAI. xAI flagship with strong reasoning and tool use.

X

Grok 3

Reasoning Model

xAI

Grok 3 introduces extended reasoning and chain-of-thought capabilities.

X

Grok 3

Benchmark Milestone

xAI

Grok 3 achieves 92/100 reasoning benchmark score — top-tier performance.

X

Grok 3 Mini

Model Release

xAI

Grok 3 Mini launched by xAI. Efficient Grok 3 variant for high-volume use.

X

Grok 3 Mini

Pricing Update

xAI

Initial pricing set at $0.30/M input and $0.50/M output tokens.

X

Grok 3 Mini

Vision Support

xAI

Grok 3 Mini adds native vision and image understanding capabilities.

X

Grok 3 Mini

Function Calling

xAI

Grok 3 Mini supports function calling and structured tool use.

X

Grok 3 Mini

Reasoning Model

xAI

Grok 3 Mini introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.0 Flash Lite

Model Release

Google

Gemini 2.0 Flash Lite launched by Google. Ultra-cheap Gemini for simple tasks at scale.

G

Gemini 2.0 Flash Lite

API Release

Google

API endpoint `gemini-2.0-flash-lite` available for Gemini 2.0 Flash Lite.

G

Gemini 2.0 Flash Lite

Pricing Update

Google

Initial pricing set at $0.075/M input and $0.30/M output tokens.

G

Gemini 2.0 Flash Lite

Context Window Increase

Google

Gemini 2.0 Flash Lite ships with 1M token context window.

G

Gemini 2.0 Flash Lite

Vision Support

Google

Gemini 2.0 Flash Lite adds native vision and image understanding capabilities.

G

Gemini 2.0 Flash Lite

Function Calling

Google

Gemini 2.0 Flash Lite supports function calling and structured tool use.

O

OpenAI o3 Mini

Model Release

OpenAI

OpenAI o3 Mini launched by OpenAI. Efficient reasoning model in the o3 family.

O

OpenAI o3 Mini

Pricing Update

OpenAI

Initial pricing set at $1.10/M input and $4.40/M output tokens.

O

OpenAI o3 Mini

Function Calling

OpenAI

OpenAI o3 Mini supports function calling and structured tool use.

O

OpenAI o3 Mini

Reasoning Model

OpenAI

OpenAI o3 Mini introduces extended reasoning and chain-of-thought capabilities.

M

Mistral Small

Model Release

Mistral

Mistral Small launched by Mistral. Compact Mistral for cost-sensitive applications.

M

Mistral Small

API Release

Mistral

API endpoint `mistral-small-2501` available for Mistral Small.

M

Mistral Small

Pricing Update

Mistral

Initial pricing set at $0.10/M input and $0.30/M output tokens.

M

Mistral Small

Function Calling

Mistral

Mistral Small supports function calling and structured tool use.

M

Mistral Small

Fine-tuning Release

Mistral

Mistral Small available for fine-tuning and custom deployment via Mistral.

D

DeepSeek R1

Model Release

DeepSeek

DeepSeek R1 launched by DeepSeek. Open reasoning model with o1-class performance at fraction of cost.

D

DeepSeek R1

Pricing Update

DeepSeek

Initial pricing set at $0.55/M input and $2.19/M output tokens.

D

DeepSeek R1

Function Calling

DeepSeek

DeepSeek R1 supports function calling and structured tool use.

D

DeepSeek R1

Reasoning Model

DeepSeek

DeepSeek R1 introduces extended reasoning and chain-of-thought capabilities.

D

DeepSeek R1

Benchmark Milestone

DeepSeek

DeepSeek R1 achieves 95/100 reasoning benchmark score — top-tier performance.

D

DeepSeek R1

Fine-tuning Release

DeepSeek

DeepSeek R1 available for fine-tuning and custom deployment via DeepSeek.

M

Codestral

Model Release

Mistral

Codestral launched by Mistral. Code-specialized Mistral model for IDE and agent use.

M

Codestral

Pricing Update

Mistral

Initial pricing set at $0.30/M input and $0.90/M output tokens.

M

Codestral

Function Calling

Mistral

Codestral supports function calling and structured tool use.

M

Codestral

Benchmark Milestone

Mistral

Codestral achieves 94/100 coding benchmark score — top-tier performance.

M

Codestral

Fine-tuning Release

Mistral

Codestral available for fine-tuning and custom deployment via Mistral.

AI Release Timeline | SummitGovOnline