SummitGovOnline

AI Releases

Latest model launches and announcements from major AI providers.

Gemini 3.6 Flash
Google

Google's speed-optimized Gemini 3.6 Flash with frontier-level intelligence at lower cost than prior Flash tiers.

Kimi K3
Moonshot AI

Moonshot flagship open-frontier model with ~2.8T parameters, native vision, and a 1M-token context window.

Claude Sonnet 4.5
Anthropic

Anthropic's balanced Sonnet 4.5 for production coding and agent workflows.

GPT-5
OpenAI

OpenAI flagship GPT-5 generation for general and agentic workloads.

GPT-5 Mini
OpenAI

Cost-efficient GPT-5 family model for high-volume production.

DeepSeek V4
DeepSeek

Next-gen DeepSeek general model emphasizing coding and cost efficiency.

Claude Opus 4.1
Anthropic

Anthropic's latest Opus-class model for deep reasoning, coding, and long-form analysis.

Gemini 3 Pro
Google

Google Gemini 3 Pro frontier tier for complex multimodal reasoning and long-context workloads.

Grok 4
xAI

xAI Grok 4 with real-time retrieval emphasis and strong general capability.

Kimi K2
Moonshot AI

Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.

Claude 4 Sonnet
Anthropic

Latest Sonnet with improved reasoning and tool use.

Claude 4 Opus
Anthropic

Anthropic flagship for the most demanding tasks.

Claude 4 Haiku
Anthropic

Next-gen Haiku with Claude 4 intelligence at low cost.

Gemini 2.5 Flash
Google

Flash model with thinking capabilities and top multimodal scores.

OpenAI o4 Mini
OpenAI

Latest compact reasoning model with vision support.

OpenAI o3
OpenAI

Full o3 reasoning model with vision and tool use.

GPT-4.1
OpenAI

Next-gen GPT with 1M context and improved instruction following.

GPT-4.1 Mini
OpenAI

Mini variant of GPT-4.1 for scalable deployments.

GPT-4.1 Nano
OpenAI

Ultra-low-cost GPT-4.1 tier for classification and extraction.

Gemini 2.5 Pro
Google

Google flagship with 1M context and deep reasoning.

DeepSeek V3 0324
DeepSeek

Updated V3 with improved instruction following and 128K context.

Command A
Cohere

Cohere flagship agentic model for tool use and search.

Llama 4 Scout
Meta

Llama 4 with industry-leading 10M context for documents.

Claude 3.7 Sonnet
Anthropic

Sonnet with extended thinking for harder problems.

Grok 3
xAI

xAI flagship with strong reasoning and tool use.

Grok 3 Mini
xAI

Efficient Grok 3 variant for high-volume use.

OpenAI o3 Mini
OpenAI

Efficient reasoning model in the o3 family.

Mistral Small
Mistral

Compact Mistral for cost-sensitive applications.

DeepSeek R1
DeepSeek

Open reasoning model with o1-class performance at fraction of cost.

Codestral
Mistral

Code-specialized Mistral model for IDE and agent use.

DeepSeek V3
DeepSeek

Highly efficient MoE model rivaling frontier closed models.

Gemini 2.0 Flash
Google

Fast multimodal Gemini with native audio and vision.

Llama 3.3 70B
Meta

Open-weight 70B model competitive with larger closed models.

OpenAI o1
OpenAI

Reasoning model using internal chain-of-thought.

Mistral Large
Mistral

Mistral flagship for enterprise European deployments.

Pixtral Large
Mistral

Mistral multimodal model with strong vision understanding.

Claude 3.5 Haiku
Anthropic

Fast and affordable Claude for high-throughput workloads.

Ministral 8B
Mistral

Tiny on-device capable Mistral for edge deployments.

Grok 2
xAI

xAI general model with real-time X data integration.

Llama 3.1 405B
Meta

Largest open-weight Llama for maximum capability.

GPT-4o Mini
OpenAI

Cost-efficient small model for high-volume applications.

Claude 3.5 Sonnet
Anthropic

Balanced Claude model excelling at coding and analysis.

DeepSeek Coder V2
DeepSeek

Code-specialized DeepSeek for programming tasks.

GPT-4o
OpenAI

Flagship multimodal model balancing speed, intelligence, and cost.

Command R+
Cohere

Cohere model optimized for RAG and enterprise search.

Command R
Cohere

Affordable Cohere model for retrieval-augmented workflows.

Claude 3 Opus
Anthropic

Previous-gen Opus still strong for writing-heavy workloads.

Gemini 1.5 Pro
Google

Proven Gemini with industry-leading 2M context window.