Baseline seeded — waiting for the first market change.
GPU market syncing — no offers yet. Admin → Bootstrap GPU market.
The WhatModel rating

WM Score

WM Score rates every model 0–100 on price, quality, context, capabilities, reliability and host redundancy — not intelligence alone. Currently v1.3.

Benchmarks measure intelligence. The WM Score measures whether a model is actually the right choice — one number that folds in what you pay, what you get, how much context you can use, which capabilities are supported, how reliably it serves, and how many independent hosts can serve it if one goes down.

Quality
Independent benchmark standing across reasoning, coding and general tasks.
Price
Blended input/output cost per million tokens, relative to the whole market.
Context
Usable context window — how much you can actually feed the model.
Capabilities
Vision, tool calling, structured output and reasoning support.
Reliability
Observed uptime and throughput from live host probes.
Host redundancy
How many independent providers serve the model — your fallback depth.

Factor weights are proprietary and versioned. Scores refresh with every market sync (about every 10 minutes), and we only publish a score move once it clears 5 points, so the number stays meaningful.

Score distribution

Loading…
90–100
EliteBest-in-market on nearly every factor.
80–89
StrongSafe default choice for production.
70–79
SolidGood, with one clear trade-off.
55–69
SituationalWorks when its strength matches your task.
Below 55
WeakUsually beaten on price or reliability.

Recent score moves

Only moves of 5 points or more are published

    Top 100 by WM Score

    0 rated models · updated every 10 minutes
    #ModelScoreBandIn / MOut / MContextHostsBadge