Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
Head to head

Google: Gemini 3.7 Flash vs OpenAI: GPT-5.6 Luna

Live prices, context windows, capabilities, and measured quality — updated continuously.

Side by side

MetricGoogle: Gemini 3.7 Flash
Google
OpenAI: GPT-5.6 Luna
Openai
Input $/M tokens$0.375$0.200
Output $/M tokens$1.88$1.20
Context window1.0M1.1M
Max output66K128K
WM Score8887
Benchmark quality8783.7
Hosts serving it23
Throughput (tok/s)23099.5
Latency (ms)1561865
Vision
Tools
Reasoning
JSON mode

Which should you pick?

OpenAI: GPT-5.6 Luna costs less per million input tokens — about 1.9× cheaper than Google: Gemini 3.7 Flash. Google: Gemini 3.7 Flash scores higher on measured quality. OpenAI: GPT-5.6 Luna takes a larger context window, which matters for long documents or big codebases. Prices here are live list prices; the same model can cost less on a different host, so check the model page for the cheapest offering.