SummitGovOnline

Fast Models

Low-latency models optimized for real-time applications and interactive experiences.

Low-latency models optimized for real-time applications and interactive experiences.

Models

21

in category

Providers

8

represented

Avg. API Cost

$3.92

combined in+out / 1M

Top Model

Gemini 3.6 Flash

Latency

Fast Models Rankings

Sorted by latency

Full leaderboard →
#ModelProviderContextInput / 1M
1Gemini 3.6 FlashGoogle1M$1.50
2Claude Sonnet 4.5Anthropic200K$3.00
3GPT-5 MiniOpenAI400K$0.40
4DeepSeek V4DeepSeek163.8K$0.30
5Grok 4xAI256K$3.00
6Kimi K2Moonshot AI131.1K$0.60
7Claude 4 HaikuAnthropic200K$0.80
8Gemini 2.5 FlashGoogle1M$0.15
9GPT-4.1 MiniOpenAI1M$0.40
10GPT-4.1 NanoOpenAI1M$0.10

Timeline Events

Frequently Asked Questions

What is the fast models?
Based on SGO's dataset, GPT-4o by OpenAI ranks highly for latency with a latency score of 90. Rankings are computed from structured benchmark and pricing data.
How are fast models ranked?
Rankings are derived from the centralized SGO model dataset — benchmark scores, API pricing, context windows, and capability flags. No recommendations are hardcoded.
How many models match this category?
21 models in the SGO catalog currently match this category.

Explore More

Catalog updated 2d ago

Last updated:

Source: SGO Intelligence Catalog

50 models · 384 snapshots

Methodology

Recommendations are computed from structured model attributes in the SGO catalog. No models are hardcoded. Similarity uses weighted benchmarks, pricing, context, capabilities, and release metadata.

21 models · Data derived from the SGO intelligence catalog