OpenRouter
Managed routerScore 97/100- Flash, per 1M
- $0.044 · p50 0.76s · p95 2.59s
- GLM-5.2, per 1M
- $0.46 · p50 1.20s · p95 3.69s
- Needle · tool calls
- 0.94 / 0.93 · 0.98 / 1.00
- Catalog
- 418 model IDs
Fastest median on both reference models and the lowest invoiced cost on Flash; no alternative beat it on both.
OpenRouter posted the fastest median latency on both reference models, the lowest invoiced cost per million tokens on DeepSeek Flash, and GLM-5.2 within 23% of the cheapest platform. No alternative in this run was both cheaper and faster on either model.
Served deepseek-v4-flash-0731, the current snapshot, with GLM-5.2 available on the run date.
- Loses points on
- GLM-5.2 cost: DeepInfra’s automatic caching undercut it by 19% per blended token. The credit top-up fee (documented at about 5.5%; not measured here) sits outside the per-token figures.
- Suits
- Workloads where one account, the widest live catalog we measured (418 model IDs), and invoiced pricing that tracked provider list rates matter most.
- Does not suit
- Cases where the workload is one large open-weight model at steady volume; DeepInfra billed 19% less per GLM-5.2 token.