Gemini 3.7 Flash vs Kimi K3

Side-by-side pricing per provider, context window, benchmark scores and speed, all on one OpenAI-compatible API.

Metric
Gemini 3.7 FlashGooglegoogle/gemini-3.7-flash
Kimi K3MoonshotAImoonshotai/kimi-k3
Creator
Google
MoonshotAI
Released
Aug 13, 2026
Jul 16, 2026
Context window
1.05M
1.05M
Max output
66K
1.05M (best)
Input / 1MLowest provider price
$0.75 (best)
$2.70$2.70–$3.00 across 4
Output / 1MLowest provider price
$3.75 (best)
$13.50$13.50–$15.00 across 4
Providers
2
4 (best)
IntelligenceArtificial Analysis index
39.1
43.6 (best)
CodingArtificial Analysis index
76.1
76.2 (best)
LatencyMedian time to first token
—
7.4s
ThroughputMedian tokens per second
—
<1 t/s
Accepts
Text, Image, Audio, Video, Files
Text, Image
Produces
Text
Text
Tool calling
Yes
Yes
Structured output
No
Yes
Reasoning
No
No

Gemini 3.7 Flash vs Kimi K3: summary

Gemini 3.7 Flash and Kimi K3 are both available through FastRouter's OpenAI-compatible API, so switching between them is a change of model id, not a new integration.

Gemini 3.7 Flash, from Google, has a 1,048,576-token context window, costs from $0.75/1M input and $3.75/1M output tokens across 2 providers and scores 39.1 on the Artificial Analysis Intelligence Index.

Kimi K3, from MoonshotAI, has a 1,048,576-token context window, costs from $2.70/1M input and $13.50/1M output tokens across 4 providers and scores 43.6 on the Artificial Analysis Intelligence Index.

Gemini 3.7 Flash is the cheapest on input, 3.6× cheaper than the next model, Kimi K3 scores highest for intelligence and Kimi K3 leads on coding.

Switch with one line

client = OpenAI(base_url="https://api.fastrouter.ai/api/v1", api_key="<FASTROUTER_API_KEY>") client.chat.completions.create(model="google/gemini-3.7-flash", messages=[...]) # or client.chat.completions.create(model="moonshotai/kimi-k3", messages=[...])
Full API docs

Frequently asked questions

More comparisons

Benchmark scores sourced from ArtificialAnalysis