Meta
Creator of the open-weight Llama model family.
Meta develops the open-weight Llama model family, enabling self-hosting, fine-tuning, and cost-effective deployment at scale with competitive benchmark performance.
Models
4
in catalog
Avg Input Price
$0.2325
per 1M tokens
Avg Output Price
$0.55
per 1M tokens
Price Range
$0.15 – $0.35
input / 1M
Meta Pricing History
Average input and output pricing across model portfolio
Meta Release Timeline
Recent releases and capability milestones
Functions support added
capability
Llama 3.3 70B released
release
Vision support added
capability
Vision support added
capability
Functions support added
capability
Functions support added
capability
Llama 4 Maverick released
release
Llama 4 Scout released
release
Model Catalog
4 models currently tracked
| Model | Context | Input | Coding | Tier |
|---|---|---|---|---|
| Llama 4 Maverick | 1M | $0.20 | 87 | standard |
| Llama 4 Scout | 10M | $0.15 | 83 | standard |
| Llama 3.3 70B | 128K | $0.23 | 84 | standard |
| Llama 3.1 405B | 128K | $0.35 | 85 | slow |
Related Timeline Events
Recent activity from the AI release timeline
Llama 4 Maverick
Model Release — Llama 4 Maverick launched by Meta. Open multimodal Llama with 1M context window.…
Llama 4 Maverick
API Release — API endpoint `llama-4-maverick` available for Llama 4 Maverick.…
Llama 4 Maverick
Pricing Update — Initial pricing set at $0.20/M input and $0.60/M output tokens.…
Llama 4 Maverick
Context Window Increase — Llama 4 Maverick ships with 1M token context window.…
Llama 4 Maverick
Vision Support — Llama 4 Maverick adds native vision and image understanding capabilities.…
Llama 4 Maverick
Function Calling — Llama 4 Maverick supports function calling and structured tool use.…
Compare with Competitors
- Llama 4 Maverick vs GPT-5(Meta vs OpenAI)
- Llama 4 Maverick vs Claude 4 Opus(Meta vs Anthropic)
- Llama 4 Maverick vs Gemini 3 Pro(Meta vs Google)
Related Calculators
- AI Cost Calculator
Estimate daily, monthly, and yearly API costs with provider and model selection.
- Context Window Utilization
Calculate how much of a model's context window you're consuming.
- Inference Latency Estimator
Project response times based on output tokens and throughput.
- Estimate Meta API costs →
Frequently Asked Questions
- Are Llama models open source?
- Llama models are released under Meta's open license, allowing commercial use with certain restrictions.
- What is Llama 4 Scout's context window?
- Llama 4 Scout offers an industry-leading 10M token context window for document RAG applications.
