SummitGovOnline

AI Release Timeline

Track model launches, capability updates, pricing changes, and benchmark milestones across the AI industry.

2026

G

Gemini 3.6 Flash

Model Release

Google

Gemini 3.6 Flash launched by Google. Google's speed-optimized Gemini 3.6 Flash with frontier-level intelligence at lower cost than prior Flash tiers.

G

Gemini 3.6 Flash

API Release

Google

API endpoint `gemini-3.6-flash` available for Gemini 3.6 Flash.

G

Gemini 3.6 Flash

Pricing Update

Google

Initial pricing set at $1.50/M input and $7.50/M output tokens.

G

Gemini 3.6 Flash

Context Window Increase

Google

Gemini 3.6 Flash ships with 1M token context window.

G

Gemini 3.6 Flash

Vision Support

Google

Gemini 3.6 Flash adds native vision and image understanding capabilities.

G

Gemini 3.6 Flash

Audio Support

Google

Gemini 3.6 Flash supports audio input and multimodal voice workflows.

G

Gemini 3.6 Flash

Function Calling

Google

Gemini 3.6 Flash supports function calling and structured tool use.

G

Gemini 3.6 Flash

Reasoning Model

Google

Gemini 3.6 Flash introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 3.6 Flash

Benchmark Milestone

Google

Gemini 3.6 Flash achieves 94/100 multimodal benchmark score — top-tier performance.

MA

Kimi K3

Model Release

Moonshot AI

Kimi K3 launched by Moonshot AI. Moonshot flagship open-frontier model with ~2.8T parameters, native vision, and a 1M-token context window.

MA

Kimi K3

Pricing Update

Moonshot AI

Initial pricing set at $3.00/M input and $15.00/M output tokens.

MA

Kimi K3

Vision Support

Moonshot AI

Kimi K3 adds native vision and image understanding capabilities.

MA

Kimi K3

Function Calling

Moonshot AI

Kimi K3 supports function calling and structured tool use.

MA

Kimi K3

Reasoning Model

Moonshot AI

Kimi K3 introduces extended reasoning and chain-of-thought capabilities.

MA

Kimi K3

Benchmark Milestone

Moonshot AI

Kimi K3 achieves 94/100 reasoning benchmark score — top-tier performance.

A

Claude Sonnet 4.5

Model Release

Anthropic

Claude Sonnet 4.5 launched by Anthropic. Anthropic's balanced Sonnet 4.5 for production coding and agent workflows.

A

Claude Sonnet 4.5

API Release

Anthropic

API endpoint `claude-sonnet-4-5` available for Claude Sonnet 4.5.

A

Claude Sonnet 4.5

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude Sonnet 4.5

Vision Support

Anthropic

Claude Sonnet 4.5 adds native vision and image understanding capabilities.

A

Claude Sonnet 4.5

Function Calling

Anthropic

Claude Sonnet 4.5 supports function calling and structured tool use.

A

Claude Sonnet 4.5

Reasoning Model

Anthropic

Claude Sonnet 4.5 introduces extended reasoning and chain-of-thought capabilities.

A

Claude Sonnet 4.5

Benchmark Milestone

Anthropic

Claude Sonnet 4.5 achieves 95/100 coding benchmark score — top-tier performance.

O

GPT-5

Model Release

OpenAI

GPT-5 launched by OpenAI. OpenAI flagship GPT-5 generation for general and agentic workloads.

O

GPT-5

Reasoning Model

OpenAI

GPT-5 introduces extended reasoning and chain-of-thought capabilities.

O

GPT-5

Benchmark Milestone

OpenAI

GPT-5 achieves 96/100 reasoning benchmark score — top-tier performance.

O

GPT-5 Mini

Model Release

OpenAI

GPT-5 Mini launched by OpenAI. Cost-efficient GPT-5 family model for high-volume production.

O

GPT-5 Mini

Pricing Update

OpenAI

Initial pricing set at $0.40/M input and $1.60/M output tokens.

O

GPT-5 Mini

Vision Support

OpenAI

GPT-5 Mini adds native vision and image understanding capabilities.

O

GPT-5 Mini

Function Calling

OpenAI

GPT-5 Mini supports function calling and structured tool use.

O

GPT-5 Mini

Reasoning Model

OpenAI

GPT-5 Mini introduces extended reasoning and chain-of-thought capabilities.

D

DeepSeek V4

Model Release

DeepSeek

DeepSeek V4 launched by DeepSeek. Next-gen DeepSeek general model emphasizing coding and cost efficiency.

D

DeepSeek V4

Pricing Update

DeepSeek

Initial pricing set at $0.30/M input and $1.20/M output tokens.

D

DeepSeek V4

Function Calling

DeepSeek

DeepSeek V4 supports function calling and structured tool use.

D

DeepSeek V4

Reasoning Model

DeepSeek

DeepSeek V4 introduces extended reasoning and chain-of-thought capabilities.

D

DeepSeek V4

Benchmark Milestone

DeepSeek

DeepSeek V4 achieves 92/100 coding benchmark score — top-tier performance.

D

DeepSeek V4

Fine-tuning Release

DeepSeek

DeepSeek V4 available for fine-tuning and custom deployment via DeepSeek.

A

Claude Opus 4.1

Model Release

Anthropic

Claude Opus 4.1 launched by Anthropic. Anthropic's latest Opus-class model for deep reasoning, coding, and long-form analysis.

A

Claude Opus 4.1

API Release

Anthropic

API endpoint `claude-opus-4-1` available for Claude Opus 4.1.

A

Claude Opus 4.1

Pricing Update

Anthropic

Initial pricing set at $15.00/M input and $75.00/M output tokens.

A

Claude Opus 4.1

Vision Support

Anthropic

Claude Opus 4.1 adds native vision and image understanding capabilities.

A

Claude Opus 4.1

Function Calling

Anthropic

Claude Opus 4.1 supports function calling and structured tool use.

A

Claude Opus 4.1

Reasoning Model

Anthropic

Claude Opus 4.1 introduces extended reasoning and chain-of-thought capabilities.

A

Claude Opus 4.1

Benchmark Milestone

Anthropic

Claude Opus 4.1 achieves 97/100 reasoning benchmark score — top-tier performance.

G

Gemini 3 Pro

Model Release

Google

Gemini 3 Pro launched by Google. Google Gemini 3 Pro frontier tier for complex multimodal reasoning and long-context workloads.

G

Gemini 3 Pro

Pricing Update

Google

Initial pricing set at $3.60/M input and $14.00/M output tokens.

G

Gemini 3 Pro

Vision Support

Google

Gemini 3 Pro adds native vision and image understanding capabilities.

G

Gemini 3 Pro

Audio Support

Google

Gemini 3 Pro supports audio input and multimodal voice workflows.

G

Gemini 3 Pro

Function Calling

Google

Gemini 3 Pro supports function calling and structured tool use.

G

Gemini 3 Pro

Reasoning Model

Google

Gemini 3 Pro introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 3 Pro

Benchmark Milestone

Google

Gemini 3 Pro achieves 96/100 multimodal benchmark score — top-tier performance.

X

Grok 4

Model Release

xAI

Grok 4 launched by xAI. xAI Grok 4 with real-time retrieval emphasis and strong general capability.

X

Grok 4

Reasoning Model

xAI

Grok 4 introduces extended reasoning and chain-of-thought capabilities.

2025

MA

Kimi K2

Model Release

Moonshot AI

Kimi K2 launched by Moonshot AI. Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.

MA

Kimi K2

Pricing Update

Moonshot AI

Initial pricing set at $0.60/M input and $2.50/M output tokens.

MA

Kimi K2

Function Calling

Moonshot AI

Kimi K2 supports function calling and structured tool use.

MA

Kimi K2

Reasoning Model

Moonshot AI

Kimi K2 introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Sonnet

Model Release

Anthropic

Claude 4 Sonnet launched by Anthropic. Latest Sonnet with improved reasoning and tool use.

A

Claude 4 Sonnet

API Release

Anthropic

API endpoint `claude-sonnet-4-20250514` available for Claude 4 Sonnet.

A

Claude 4 Sonnet

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude 4 Sonnet

Vision Support

Anthropic

Claude 4 Sonnet adds native vision and image understanding capabilities.

A

Claude 4 Sonnet

Function Calling

Anthropic

Claude 4 Sonnet supports function calling and structured tool use.

A

Claude 4 Sonnet

Reasoning Model

Anthropic

Claude 4 Sonnet introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Sonnet

Benchmark Milestone

Anthropic

Claude 4 Sonnet achieves 96/100 coding benchmark score — top-tier performance.

A

Claude 4 Opus

Model Release

Anthropic

Claude 4 Opus launched by Anthropic. Anthropic flagship for the most demanding tasks.

A

Claude 4 Opus

API Release

Anthropic

API endpoint `claude-opus-4-20250514` available for Claude 4 Opus.

A

Claude 4 Opus

Pricing Update

Anthropic

Initial pricing set at $15.00/M input and $75.00/M output tokens.

A

Claude 4 Opus

Vision Support

Anthropic

Claude 4 Opus adds native vision and image understanding capabilities.

A

Claude 4 Opus

Function Calling

Anthropic

Claude 4 Opus supports function calling and structured tool use.

A

Claude 4 Opus

Reasoning Model

Anthropic

Claude 4 Opus introduces extended reasoning and chain-of-thought capabilities.

A

Claude 4 Opus

Benchmark Milestone

Anthropic

Claude 4 Opus achieves 97/100 coding benchmark score — top-tier performance.

A

Claude 4 Haiku

Model Release

Anthropic

Claude 4 Haiku launched by Anthropic. Next-gen Haiku with Claude 4 intelligence at low cost.

A

Claude 4 Haiku

API Release

Anthropic

API endpoint `claude-haiku-4-20250514` available for Claude 4 Haiku.

A

Claude 4 Haiku

Pricing Update

Anthropic

Initial pricing set at $0.80/M input and $4.00/M output tokens.

A

Claude 4 Haiku

Vision Support

Anthropic

Claude 4 Haiku adds native vision and image understanding capabilities.

A

Claude 4 Haiku

Function Calling

Anthropic

Claude 4 Haiku supports function calling and structured tool use.

G

Gemini 2.5 Flash

Model Release

Google

Gemini 2.5 Flash launched by Google. Flash model with thinking capabilities and top multimodal scores.

G

Gemini 2.5 Flash

API Release

Google

API endpoint `gemini-2.5-flash-preview-05-20` available for Gemini 2.5 Flash.

G

Gemini 2.5 Flash

Pricing Update

Google

Initial pricing set at $0.15/M input and $0.60/M output tokens.

G

Gemini 2.5 Flash

Context Window Increase

Google

Gemini 2.5 Flash ships with 1M token context window.

G

Gemini 2.5 Flash

Vision Support

Google

Gemini 2.5 Flash adds native vision and image understanding capabilities.

G

Gemini 2.5 Flash

Audio Support

Google

Gemini 2.5 Flash supports audio input and multimodal voice workflows.

G

Gemini 2.5 Flash

Function Calling

Google

Gemini 2.5 Flash supports function calling and structured tool use.

G

Gemini 2.5 Flash

Reasoning Model

Google

Gemini 2.5 Flash introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.5 Flash

Benchmark Milestone

Google

Gemini 2.5 Flash achieves 94/100 multimodal benchmark score — top-tier performance.

O

OpenAI o4 Mini

Model Release

OpenAI

OpenAI o4 Mini launched by OpenAI. Latest compact reasoning model with vision support.

O

OpenAI o4 Mini

Pricing Update

OpenAI

Initial pricing set at $1.10/M input and $4.40/M output tokens.

O

OpenAI o4 Mini

Vision Support

OpenAI

OpenAI o4 Mini adds native vision and image understanding capabilities.

O

OpenAI o4 Mini

Function Calling

OpenAI

OpenAI o4 Mini supports function calling and structured tool use.

O

OpenAI o4 Mini

Reasoning Model

OpenAI

OpenAI o4 Mini introduces extended reasoning and chain-of-thought capabilities.

O

OpenAI o4 Mini

Benchmark Milestone

OpenAI

OpenAI o4 Mini achieves 93/100 reasoning benchmark score — top-tier performance.

O

OpenAI o3

Model Release

OpenAI

OpenAI o3 launched by OpenAI. Full o3 reasoning model with vision and tool use.

O

OpenAI o3

Pricing Update

OpenAI

Initial pricing set at $10.00/M input and $40.00/M output tokens.

O

OpenAI o3

Function Calling

OpenAI

OpenAI o3 supports function calling and structured tool use.

O

OpenAI o3

Reasoning Model

OpenAI

OpenAI o3 introduces extended reasoning and chain-of-thought capabilities.

O

OpenAI o3

Benchmark Milestone

OpenAI

OpenAI o3 achieves 97/100 reasoning benchmark score — top-tier performance.

O

GPT-4.1

Model Release

OpenAI

GPT-4.1 launched by OpenAI. Next-gen GPT with 1M context and improved instruction following.

O

GPT-4.1

Vision Support

OpenAI

GPT-4.1 adds native vision and image understanding capabilities.

O

GPT-4.1

Benchmark Milestone

OpenAI

GPT-4.1 achieves 94/100 coding benchmark score — top-tier performance.

O

GPT-4.1 Mini

Model Release

OpenAI

GPT-4.1 Mini launched by OpenAI. Mini variant of GPT-4.1 for scalable deployments.

O

GPT-4.1 Mini

Pricing Update

OpenAI

Initial pricing set at $0.40/M input and $1.60/M output tokens.

O

GPT-4.1 Mini

Vision Support

OpenAI

GPT-4.1 Mini adds native vision and image understanding capabilities.

O

GPT-4.1 Mini

Audio Support

OpenAI

GPT-4.1 Mini supports audio input and multimodal voice workflows.

O

GPT-4.1 Mini

Function Calling

OpenAI

GPT-4.1 Mini supports function calling and structured tool use.

O

GPT-4.1 Nano

Model Release

OpenAI

GPT-4.1 Nano launched by OpenAI. Ultra-low-cost GPT-4.1 tier for classification and extraction.

O

GPT-4.1 Nano

Pricing Update

OpenAI

Initial pricing set at $0.10/M input and $0.40/M output tokens.

O

GPT-4.1 Nano

Vision Support

OpenAI

GPT-4.1 Nano adds native vision and image understanding capabilities.

O

GPT-4.1 Nano

Function Calling

OpenAI

GPT-4.1 Nano supports function calling and structured tool use.

G

Gemini 2.5 Pro

Model Release

Google

Gemini 2.5 Pro launched by Google. Google flagship with 1M context and deep reasoning.

G

Gemini 2.5 Pro

API Release

Google

API endpoint `gemini-2.5-pro-preview-06-05` available for Gemini 2.5 Pro.

G

Gemini 2.5 Pro

Pricing Update

Google

Initial pricing set at $1.25/M input and $10.00/M output tokens.

G

Gemini 2.5 Pro

Context Window Increase

Google

Gemini 2.5 Pro ships with 1M token context window.

G

Gemini 2.5 Pro

Vision Support

Google

Gemini 2.5 Pro adds native vision and image understanding capabilities.

G

Gemini 2.5 Pro

Audio Support

Google

Gemini 2.5 Pro supports audio input and multimodal voice workflows.

G

Gemini 2.5 Pro

Function Calling

Google

Gemini 2.5 Pro supports function calling and structured tool use.

G

Gemini 2.5 Pro

Reasoning Model

Google

Gemini 2.5 Pro introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.5 Pro

Benchmark Milestone

Google

Gemini 2.5 Pro achieves 97/100 multimodal benchmark score — top-tier performance.

D

DeepSeek V3 0324

Model Release

DeepSeek

DeepSeek V3 0324 launched by DeepSeek. Updated V3 with improved instruction following and 128K context.

D

DeepSeek V3 0324

API Release

DeepSeek

API endpoint `deepseek-chat` available for DeepSeek V3 0324.

D

DeepSeek V3 0324

Pricing Update

DeepSeek

Initial pricing set at $0.27/M input and $1.10/M output tokens.

D

DeepSeek V3 0324

Function Calling

DeepSeek

DeepSeek V3 0324 supports function calling and structured tool use.

D

DeepSeek V3 0324

Benchmark Milestone

DeepSeek

DeepSeek V3 0324 achieves 92/100 coding benchmark score — top-tier performance.

D

DeepSeek V3 0324

Fine-tuning Release

DeepSeek

DeepSeek V3 0324 available for fine-tuning and custom deployment via DeepSeek.

C

Command A

Model Release

Cohere

Command A launched by Cohere. Cohere flagship agentic model for tool use and search.

C

Command A

Pricing Update

Cohere

Initial pricing set at $2.50/M input and $10.00/M output tokens.

C

Command A

Vision Support

Cohere

Command A adds native vision and image understanding capabilities.

C

Command A

Function Calling

Cohere

Command A supports function calling and structured tool use.

M

Llama 4 Maverick

Model Release

Meta

Llama 4 Maverick launched by Meta. Open multimodal Llama with 1M context window.

M

Llama 4 Maverick

API Release

Meta

API endpoint `llama-4-maverick` available for Llama 4 Maverick.

M

Llama 4 Maverick

Pricing Update

Meta

Initial pricing set at $0.20/M input and $0.60/M output tokens.

M

Llama 4 Maverick

Context Window Increase

Meta

Llama 4 Maverick ships with 1M token context window.

M

Llama 4 Maverick

Vision Support

Meta

Llama 4 Maverick adds native vision and image understanding capabilities.

M

Llama 4 Maverick

Function Calling

Meta

Llama 4 Maverick supports function calling and structured tool use.

M

Llama 4 Maverick

Fine-tuning Release

Meta

Llama 4 Maverick available for fine-tuning and custom deployment via Meta.

M

Llama 4 Scout

Model Release

Meta

Llama 4 Scout launched by Meta. Llama 4 with industry-leading 10M context for documents.

M

Llama 4 Scout

Pricing Update

Meta

Initial pricing set at $0.15/M input and $0.50/M output tokens.

M

Llama 4 Scout

Vision Support

Meta

Llama 4 Scout adds native vision and image understanding capabilities.

M

Llama 4 Scout

Function Calling

Meta

Llama 4 Scout supports function calling and structured tool use.

M

Llama 4 Scout

Fine-tuning Release

Meta

Llama 4 Scout available for fine-tuning and custom deployment via Meta.

A

Claude 3.7 Sonnet

Model Release

Anthropic

Claude 3.7 Sonnet launched by Anthropic. Sonnet with extended thinking for harder problems.

A

Claude 3.7 Sonnet

API Release

Anthropic

API endpoint `claude-3-7-sonnet-20250219` available for Claude 3.7 Sonnet.

A

Claude 3.7 Sonnet

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude 3.7 Sonnet

Vision Support

Anthropic

Claude 3.7 Sonnet adds native vision and image understanding capabilities.

A

Claude 3.7 Sonnet

Function Calling

Anthropic

Claude 3.7 Sonnet supports function calling and structured tool use.

A

Claude 3.7 Sonnet

Reasoning Model

Anthropic

Claude 3.7 Sonnet introduces extended reasoning and chain-of-thought capabilities.

A

Claude 3.7 Sonnet

Benchmark Milestone

Anthropic

Claude 3.7 Sonnet achieves 95/100 coding benchmark score — top-tier performance.

X

Grok 3

Model Release

xAI

Grok 3 launched by xAI. xAI flagship with strong reasoning and tool use.

X

Grok 3

Reasoning Model

xAI

Grok 3 introduces extended reasoning and chain-of-thought capabilities.

X

Grok 3

Benchmark Milestone

xAI

Grok 3 achieves 92/100 reasoning benchmark score — top-tier performance.

X

Grok 3 Mini

Model Release

xAI

Grok 3 Mini launched by xAI. Efficient Grok 3 variant for high-volume use.

X

Grok 3 Mini

Pricing Update

xAI

Initial pricing set at $0.30/M input and $0.50/M output tokens.

X

Grok 3 Mini

Vision Support

xAI

Grok 3 Mini adds native vision and image understanding capabilities.

X

Grok 3 Mini

Function Calling

xAI

Grok 3 Mini supports function calling and structured tool use.

X

Grok 3 Mini

Reasoning Model

xAI

Grok 3 Mini introduces extended reasoning and chain-of-thought capabilities.

G

Gemini 2.0 Flash Lite

Model Release

Google

Gemini 2.0 Flash Lite launched by Google. Ultra-cheap Gemini for simple tasks at scale.

G

Gemini 2.0 Flash Lite

API Release

Google

API endpoint `gemini-2.0-flash-lite` available for Gemini 2.0 Flash Lite.

G

Gemini 2.0 Flash Lite

Pricing Update

Google

Initial pricing set at $0.075/M input and $0.30/M output tokens.

G

Gemini 2.0 Flash Lite

Context Window Increase

Google

Gemini 2.0 Flash Lite ships with 1M token context window.

G

Gemini 2.0 Flash Lite

Vision Support

Google

Gemini 2.0 Flash Lite adds native vision and image understanding capabilities.

G

Gemini 2.0 Flash Lite

Function Calling

Google

Gemini 2.0 Flash Lite supports function calling and structured tool use.

O

OpenAI o3 Mini

Model Release

OpenAI

OpenAI o3 Mini launched by OpenAI. Efficient reasoning model in the o3 family.

O

OpenAI o3 Mini

Pricing Update

OpenAI

Initial pricing set at $1.10/M input and $4.40/M output tokens.

O

OpenAI o3 Mini

Function Calling

OpenAI

OpenAI o3 Mini supports function calling and structured tool use.

O

OpenAI o3 Mini

Reasoning Model

OpenAI

OpenAI o3 Mini introduces extended reasoning and chain-of-thought capabilities.

M

Mistral Small

Model Release

Mistral

Mistral Small launched by Mistral. Compact Mistral for cost-sensitive applications.

M

Mistral Small

API Release

Mistral

API endpoint `mistral-small-2501` available for Mistral Small.

M

Mistral Small

Pricing Update

Mistral

Initial pricing set at $0.10/M input and $0.30/M output tokens.

M

Mistral Small

Function Calling

Mistral

Mistral Small supports function calling and structured tool use.

M

Mistral Small

Fine-tuning Release

Mistral

Mistral Small available for fine-tuning and custom deployment via Mistral.

D

DeepSeek R1

Model Release

DeepSeek

DeepSeek R1 launched by DeepSeek. Open reasoning model with o1-class performance at fraction of cost.

D

DeepSeek R1

Pricing Update

DeepSeek

Initial pricing set at $0.55/M input and $2.19/M output tokens.

D

DeepSeek R1

Function Calling

DeepSeek

DeepSeek R1 supports function calling and structured tool use.

D

DeepSeek R1

Reasoning Model

DeepSeek

DeepSeek R1 introduces extended reasoning and chain-of-thought capabilities.

D

DeepSeek R1

Benchmark Milestone

DeepSeek

DeepSeek R1 achieves 95/100 reasoning benchmark score — top-tier performance.

D

DeepSeek R1

Fine-tuning Release

DeepSeek

DeepSeek R1 available for fine-tuning and custom deployment via DeepSeek.

M

Codestral

Model Release

Mistral

Codestral launched by Mistral. Code-specialized Mistral model for IDE and agent use.

M

Codestral

Pricing Update

Mistral

Initial pricing set at $0.30/M input and $0.90/M output tokens.

M

Codestral

Function Calling

Mistral

Codestral supports function calling and structured tool use.

M

Codestral

Benchmark Milestone

Mistral

Codestral achieves 94/100 coding benchmark score — top-tier performance.

M

Codestral

Fine-tuning Release

Mistral

Codestral available for fine-tuning and custom deployment via Mistral.

2024

D

DeepSeek V3

Model Release

DeepSeek

DeepSeek V3 launched by DeepSeek. Highly efficient MoE model rivaling frontier closed models.

D

DeepSeek V3

Pricing Update

DeepSeek

Initial pricing set at $0.27/M input and $1.10/M output tokens.

D

DeepSeek V3

Function Calling

DeepSeek

DeepSeek V3 supports function calling and structured tool use.

D

DeepSeek V3

Fine-tuning Release

DeepSeek

DeepSeek V3 available for fine-tuning and custom deployment via DeepSeek.

G

Gemini 2.0 Flash

Model Release

Google

Gemini 2.0 Flash launched by Google. Fast multimodal Gemini with native audio and vision.

G

Gemini 2.0 Flash

API Release

Google

API endpoint `gemini-2.0-flash` available for Gemini 2.0 Flash.

G

Gemini 2.0 Flash

Pricing Update

Google

Initial pricing set at $0.10/M input and $0.40/M output tokens.

G

Gemini 2.0 Flash

Context Window Increase

Google

Gemini 2.0 Flash ships with 1M token context window.

G

Gemini 2.0 Flash

Vision Support

Google

Gemini 2.0 Flash adds native vision and image understanding capabilities.

G

Gemini 2.0 Flash

Audio Support

Google

Gemini 2.0 Flash supports audio input and multimodal voice workflows.

G

Gemini 2.0 Flash

Function Calling

Google

Gemini 2.0 Flash supports function calling and structured tool use.

G

Gemini 2.0 Flash

Benchmark Milestone

Google

Gemini 2.0 Flash achieves 92/100 multimodal benchmark score — top-tier performance.

M

Llama 3.3 70B

Model Release

Meta

Llama 3.3 70B launched by Meta. Open-weight 70B model competitive with larger closed models.

M

Llama 3.3 70B

API Release

Meta

API endpoint `llama-3.3-70b-instruct` available for Llama 3.3 70B.

M

Llama 3.3 70B

Pricing Update

Meta

Initial pricing set at $0.23/M input and $0.40/M output tokens.

M

Llama 3.3 70B

Function Calling

Meta

Llama 3.3 70B supports function calling and structured tool use.

M

Llama 3.3 70B

Fine-tuning Release

Meta

Llama 3.3 70B available for fine-tuning and custom deployment via Meta.

O

OpenAI o1

Model Release

OpenAI

OpenAI o1 launched by OpenAI. Reasoning model using internal chain-of-thought.

O

OpenAI o1

Pricing Update

OpenAI

Initial pricing set at $15.00/M input and $60.00/M output tokens.

O

OpenAI o1

Reasoning Model

OpenAI

OpenAI o1 introduces extended reasoning and chain-of-thought capabilities.

O

OpenAI o1

Benchmark Milestone

OpenAI

OpenAI o1 achieves 96/100 reasoning benchmark score — top-tier performance.

M

Mistral Large

Model Release

Mistral

Mistral Large launched by Mistral. Mistral flagship for enterprise European deployments.

M

Mistral Large

API Release

Mistral

API endpoint `mistral-large-2411` available for Mistral Large.

M

Mistral Large

Pricing Update

Mistral

Initial pricing set at $2.00/M input and $6.00/M output tokens.

M

Mistral Large

Vision Support

Mistral

Mistral Large adds native vision and image understanding capabilities.

M

Mistral Large

Function Calling

Mistral

Mistral Large supports function calling and structured tool use.

M

Mistral Large

Fine-tuning Release

Mistral

Mistral Large available for fine-tuning and custom deployment via Mistral.

M

Pixtral Large

Model Release

Mistral

Pixtral Large launched by Mistral. Mistral multimodal model with strong vision understanding.

M

Pixtral Large

API Release

Mistral

API endpoint `pixtral-large-2411` available for Pixtral Large.

M

Pixtral Large

Pricing Update

Mistral

Initial pricing set at $2.00/M input and $6.00/M output tokens.

M

Pixtral Large

Vision Support

Mistral

Pixtral Large adds native vision and image understanding capabilities.

M

Pixtral Large

Function Calling

Mistral

Pixtral Large supports function calling and structured tool use.

M

Pixtral Large

Benchmark Milestone

Mistral

Pixtral Large achieves 93/100 multimodal benchmark score — top-tier performance.

M

Pixtral Large

Fine-tuning Release

Mistral

Pixtral Large available for fine-tuning and custom deployment via Mistral.

A

Claude 3.5 Haiku

Model Release

Anthropic

Claude 3.5 Haiku launched by Anthropic. Fast and affordable Claude for high-throughput workloads.

A

Claude 3.5 Haiku

API Release

Anthropic

API endpoint `claude-3-5-haiku-20241022` available for Claude 3.5 Haiku.

A

Claude 3.5 Haiku

Pricing Update

Anthropic

Initial pricing set at $0.80/M input and $4.00/M output tokens.

A

Claude 3.5 Haiku

Vision Support

Anthropic

Claude 3.5 Haiku adds native vision and image understanding capabilities.

A

Claude 3.5 Haiku

Function Calling

Anthropic

Claude 3.5 Haiku supports function calling and structured tool use.

M

Ministral 8B

Model Release

Mistral

Ministral 8B launched by Mistral. Tiny on-device capable Mistral for edge deployments.

M

Ministral 8B

Pricing Update

Mistral

Initial pricing set at $0.10/M input and $0.10/M output tokens.

M

Ministral 8B

Function Calling

Mistral

Ministral 8B supports function calling and structured tool use.

M

Ministral 8B

Fine-tuning Release

Mistral

Ministral 8B available for fine-tuning and custom deployment via Mistral.

X

Grok 2

Model Release

xAI

Grok 2 launched by xAI. xAI general model with real-time X data integration.

M

Llama 3.1 405B

Model Release

Meta

Llama 3.1 405B launched by Meta. Largest open-weight Llama for maximum capability.

M

Llama 3.1 405B

API Release

Meta

API endpoint `llama-3.1-405b-instruct` available for Llama 3.1 405B.

M

Llama 3.1 405B

Pricing Update

Meta

Initial pricing set at $0.35/M input and $0.70/M output tokens.

M

Llama 3.1 405B

Function Calling

Meta

Llama 3.1 405B supports function calling and structured tool use.

M

Llama 3.1 405B

Fine-tuning Release

Meta

Llama 3.1 405B available for fine-tuning and custom deployment via Meta.

O

GPT-4o Mini

Model Release

OpenAI

GPT-4o Mini launched by OpenAI. Cost-efficient small model for high-volume applications.

O

GPT-4o Mini

Pricing Update

OpenAI

Initial pricing set at $0.15/M input and $0.60/M output tokens.

O

GPT-4o Mini

Vision Support

OpenAI

GPT-4o Mini adds native vision and image understanding capabilities.

O

GPT-4o Mini

Function Calling

OpenAI

GPT-4o Mini supports function calling and structured tool use.

A

Claude 3.5 Sonnet

Model Release

Anthropic

Claude 3.5 Sonnet launched by Anthropic. Balanced Claude model excelling at coding and analysis.

A

Claude 3.5 Sonnet

API Release

Anthropic

API endpoint `claude-3-5-sonnet-20241022` available for Claude 3.5 Sonnet.

A

Claude 3.5 Sonnet

Pricing Update

Anthropic

Initial pricing set at $3.00/M input and $15.00/M output tokens.

A

Claude 3.5 Sonnet

Vision Support

Anthropic

Claude 3.5 Sonnet adds native vision and image understanding capabilities.

A

Claude 3.5 Sonnet

Function Calling

Anthropic

Claude 3.5 Sonnet supports function calling and structured tool use.

A

Claude 3.5 Sonnet

Benchmark Milestone

Anthropic

Claude 3.5 Sonnet achieves 94/100 writing benchmark score — top-tier performance.

D

DeepSeek Coder V2

Model Release

DeepSeek

DeepSeek Coder V2 launched by DeepSeek. Code-specialized DeepSeek for programming tasks.

D

DeepSeek Coder V2

API Release

DeepSeek

API endpoint `deepseek-coder` available for DeepSeek Coder V2.

D

DeepSeek Coder V2

Pricing Update

DeepSeek

Initial pricing set at $0.14/M input and $0.28/M output tokens.

D

DeepSeek Coder V2

Function Calling

DeepSeek

DeepSeek Coder V2 supports function calling and structured tool use.

D

DeepSeek Coder V2

Benchmark Milestone

DeepSeek

DeepSeek Coder V2 achieves 93/100 coding benchmark score — top-tier performance.

D

DeepSeek Coder V2

Fine-tuning Release

DeepSeek

DeepSeek Coder V2 available for fine-tuning and custom deployment via DeepSeek.

O

GPT-4o

Model Release

OpenAI

GPT-4o launched by OpenAI. Flagship multimodal model balancing speed, intelligence, and cost.

O

GPT-4o

Benchmark Milestone

OpenAI

GPT-4o achieves 95/100 multimodal benchmark score — top-tier performance.

C

Command R+

Model Release

Cohere

Command R+ launched by Cohere. Cohere model optimized for RAG and enterprise search.

C

Command R+

Pricing Update

Cohere

Initial pricing set at $2.50/M input and $10.00/M output tokens.

C

Command R+

Function Calling

Cohere

Command R+ supports function calling and structured tool use.

C

Command R

Model Release

Cohere

Command R launched by Cohere. Affordable Cohere model for retrieval-augmented workflows.

C

Command R

Pricing Update

Cohere

Initial pricing set at $0.15/M input and $0.60/M output tokens.

C

Command R

Function Calling

Cohere

Command R supports function calling and structured tool use.

A

Claude 3 Opus

Model Release

Anthropic

Claude 3 Opus launched by Anthropic. Previous-gen Opus still strong for writing-heavy workloads.

A

Claude 3 Opus

API Release

Anthropic

API endpoint `claude-3-opus-20240229` available for Claude 3 Opus.

A

Claude 3 Opus

Pricing Update

Anthropic

Initial pricing set at $15.00/M input and $75.00/M output tokens.

A

Claude 3 Opus

Vision Support

Anthropic

Claude 3 Opus adds native vision and image understanding capabilities.

A

Claude 3 Opus

Function Calling

Anthropic

Claude 3 Opus supports function calling and structured tool use.

A

Claude 3 Opus

Benchmark Milestone

Anthropic

Claude 3 Opus achieves 93/100 writing benchmark score — top-tier performance.

G

Gemini 1.5 Pro

Model Release

Google

Gemini 1.5 Pro launched by Google. Proven Gemini with industry-leading 2M context window.

G

Gemini 1.5 Pro

Pricing Update

Google

Initial pricing set at $1.25/M input and $5.00/M output tokens.

G

Gemini 1.5 Pro

Context Window Increase

Google

Gemini 1.5 Pro ships with 2.1M token context window.

G

Gemini 1.5 Pro

Vision Support

Google

Gemini 1.5 Pro adds native vision and image understanding capabilities.

G

Gemini 1.5 Pro

Function Calling

Google

Gemini 1.5 Pro supports function calling and structured tool use.

AI Release Timeline | SummitGovOnline