Google's speed-optimized Gemini 3.6 Flash with frontier-level intelligence at lower cost than prior Flash tiers.
AI Releases
Latest model launches and announcements from major AI providers.
Moonshot flagship open-frontier model with ~2.8T parameters, native vision, and a 1M-token context window.
Anthropic's balanced Sonnet 4.5 for production coding and agent workflows.
OpenAI flagship GPT-5 generation for general and agentic workloads.
Cost-efficient GPT-5 family model for high-volume production.
Next-gen DeepSeek general model emphasizing coding and cost efficiency.
Anthropic's latest Opus-class model for deep reasoning, coding, and long-form analysis.
Google Gemini 3 Pro frontier tier for complex multimodal reasoning and long-context workloads.
xAI Grok 4 with real-time retrieval emphasis and strong general capability.
Prior Moonshot coding-strong MoE model, strong agent/tool use at competitive pricing.
Latest Sonnet with improved reasoning and tool use.
Anthropic flagship for the most demanding tasks.
Next-gen Haiku with Claude 4 intelligence at low cost.
Flash model with thinking capabilities and top multimodal scores.
Latest compact reasoning model with vision support.
Full o3 reasoning model with vision and tool use.
Next-gen GPT with 1M context and improved instruction following.
Mini variant of GPT-4.1 for scalable deployments.
Ultra-low-cost GPT-4.1 tier for classification and extraction.
Google flagship with 1M context and deep reasoning.
Updated V3 with improved instruction following and 128K context.
Cohere flagship agentic model for tool use and search.
Open multimodal Llama with 1M context window.
Llama 4 with industry-leading 10M context for documents.
Sonnet with extended thinking for harder problems.
xAI flagship with strong reasoning and tool use.
Efficient Grok 3 variant for high-volume use.
Ultra-cheap Gemini for simple tasks at scale.
Efficient reasoning model in the o3 family.
Compact Mistral for cost-sensitive applications.
Open reasoning model with o1-class performance at fraction of cost.
Code-specialized Mistral model for IDE and agent use.
Highly efficient MoE model rivaling frontier closed models.
Fast multimodal Gemini with native audio and vision.
Open-weight 70B model competitive with larger closed models.
Reasoning model using internal chain-of-thought.
Mistral flagship for enterprise European deployments.
Mistral multimodal model with strong vision understanding.
Fast and affordable Claude for high-throughput workloads.
Tiny on-device capable Mistral for edge deployments.
xAI general model with real-time X data integration.
Largest open-weight Llama for maximum capability.
Cost-efficient small model for high-volume applications.
Balanced Claude model excelling at coding and analysis.
Code-specialized DeepSeek for programming tasks.
Flagship multimodal model balancing speed, intelligence, and cost.
Cohere model optimized for RAG and enterprise search.
Affordable Cohere model for retrieval-augmented workflows.
Previous-gen Opus still strong for writing-heavy workloads.
Proven Gemini with industry-leading 2M context window.
