Grok 4
xAI Grok 4 with real-time retrieval emphasis and strong general capability.
Pricing
Specifications
Benchmark Scores
Strengths
- Real-time knowledge angle
- Fast
- Tool use
Weaknesses
- Smaller ecosystem than OpenAI/Anthropic
Historical Intelligence
Pricing, context, and benchmark trends over time
Pricing History
Input and output price per million tokens over time
Context Window Growth
Context window tokens over time
Coding Score Progress
Benchmark score over time (0–100)
Reasoning Score Progress
Benchmark score over time (0–100)
Grok 4 Evolution
Combined pricing trend and benchmark performance
Grok 4 vs related models
Capability radar vs semantically related catalog models
Intelligence Graph
Provider hub →Provider Overview
xAI develops Grok models with real-time data integration from X (Twitter) and competitive reasoning capabilities at accessible price points.
4 models tracked · Documentation
Related Models
Computed from benchmark, pricing, and capability similarity
Grok 3
xAI flagship with strong reasoning and tool use.
Why: same provider (xAI), similar coding performance
Confidence95%Grok 3 Mini
Efficient Grok 3 variant for high-volume use.
Why: same provider (xAI), both vision-capable
Confidence92%Claude Sonnet 4.5
Anthropic's balanced Sonnet 4.5 for production coding and agent workflows.
Why: similar coding performance, similar reasoning scores
Confidence92%GPT-5 Mini
Cost-efficient GPT-5 family model for high-volume production.
Why: similar coding performance, similar reasoning scores
Confidence90%Grok 2
xAI general model with real-time X data integration.
Why: same provider (xAI), both vision-capable
Confidence89%Kimi K3
Moonshot flagship open-frontier model with ~2.8T parameters, native vision, and a 1M-token context window.
Why: similar coding performance, similar reasoning scores
Confidence89%
Recommended Comparisons
Side-by-side analysis with semantically similar models
Grok 4 vs Grok 3
Side-by-side pricing, benchmarks, and capabilities
Why: same provider (xAI), similar coding performance
Confidence95%Grok 4 vs Grok 3 Mini
Side-by-side pricing, benchmarks, and capabilities
Why: same provider (xAI), both vision-capable
Confidence92%Grok 4 vs Claude Sonnet 4.5
Side-by-side pricing, benchmarks, and capabilities
Why: similar coding performance, similar reasoning scores
Confidence92%Grok 4 vs GPT-5 Mini
Side-by-side pricing, benchmarks, and capabilities
Why: similar coding performance, similar reasoning scores
Confidence90%Grok 4 vs Grok 2
Side-by-side pricing, benchmarks, and capabilities
Why: same provider (xAI), both vision-capable
Confidence89%Grok 4 vs Kimi K3
Side-by-side pricing, benchmarks, and capabilities
Why: similar coding performance, similar reasoning scores
Confidence89%
Best Alternatives
Higher-scoring models with similar capability profiles
Claude 4 Opus
Anthropic flagship for the most demanding tasks.
Why: +7 overall benchmark score, both vision-capable
Confidence72%GPT-5
OpenAI flagship GPT-5 generation for general and agentic workloads.
Why: +6 overall benchmark score, similar coding performance, similar reasoning scores
Confidence83%Claude Opus 4.1
Anthropic's latest Opus-class model for deep reasoning, coding, and long-form analysis.
Why: +6 overall benchmark score, both vision-capable
Confidence74%Claude 4 Sonnet
Latest Sonnet with improved reasoning and tool use.
Why: +5 overall benchmark score, similar reasoning scores, both vision-capable
Confidence88%
Cheaper Alternatives
Lower-cost models with comparable attributes
Grok 3 Mini
Efficient Grok 3 variant for high-volume use.
Why: Saves $17.20/M combined tokens vs Grok 4
Confidence92%GPT-5 Mini
Cost-efficient GPT-5 family model for high-volume production.
Why: Saves $16.00/M combined tokens vs Grok 4
Confidence90%Grok 2
xAI general model with real-time X data integration.
Why: Saves $6.00/M combined tokens vs Grok 4
Confidence89%Claude 4 Haiku
Next-gen Haiku with Claude 4 intelligence at low cost.
Why: Saves $13.20/M combined tokens vs Grok 4
Confidence85%
Related Calculators
AI Cost Calculator
Estimate daily, monthly, and yearly API costs with provider and model selection.
Why: Estimate costs for Grok 4
Inference Latency Estimator
Project response times based on output tokens and throughput.
Why: fast latency tier estimation
Token Counter
Convert between words and estimated token counts.
Why: Token estimation for prompt workloads
Context Window Utilization
Calculate how much of a model's context window you're consuming.
Why: 256K context window analysis
Related Tools
AI Model Finder
Find models matching your priorities
Why: Relevant for Grok 4 workloads
AI Stack Builder
Build a tailored AI tool stack
Why: Relevant for Grok 4 workloads
Prompt Optimizer
Optimize prompts for cost and quality
Why: Relevant for Grok 4 workloads
Subscription Optimizer
Optimize AI subscription spend
Why: Relevant for Grok 4 workloads
Timeline Highlights
Latest Releases
Recent launches from related providers
