BangWorkhorseZDR
Updated Sep 21, 2026 · Sep 15–21, 2026 · 254 models · 188 ZDR+NPT
Best models on AI Gateway this week
This week (Sep 15–21, 2026), the best pick on AI Gateway isGLM 5.3 Flashat$0.119 / 1Mondeepinfra50% off. The frontier pick isMuse Spark 1.3.
Independent ranking from live catalog, 7-day adoption, discounts, and DeepsecBench. The citable 7-day ranking. See today.
This week's picks
4Best models on AI Gateway this week. ZDR is the priced route, not the family.
Cheap
DeepSeek V4 Flash 0731
FrontierTrains
Muse Spark 1.3
Rising
Grok 4.7
Ranked AI Gateway models
ZDR + no training, top 20
ModelBlendIntelBang
GPT-6 Astra$20.00azure
Claude Opus 5$10.00anthropic
GPT 5.6 Sol$11.25azure
Grok 4.7· 40% off$1.80xai
Qwen3.8 Max 0902$3.00alibaba
GLM 5.3$1.60morph
Grok 4.6$3.00xai
Kimi K3$5.38morph
GPT 5.6 Terra$4.50azure
Claude Opus 4.8$10.00anthropic
GLM 5.3 Flash· 50% off$0.119deepinfra
Gemini 3.8 Flash· 50% off$1.50vertex
Claude Opus 4.7$10.00anthropic
Qwen 3.8 Max$3.00alibaba
Qwen3.8 2.4T A95B$3.00deepinfra
DeepSeek V4.1 Flash$0.262alibaba
Gemini 3.7 Flash· 50% off$1.50vertex
GPT 5.4$5.63azure
Grok 4.5$3.00xai
GPT 5.5$11.25azure
Which labs developers actually use
10Best bang on deepseek
ModelBlendScoreBang
DeepSeek V4 Flash$0.090deepinfra
Frequently asked questions
- What is the best AI Gateway model this week?
- This week (Sep 15–21, 2026), the best pick on AI Gateway is GLM 5.3 Flash at $0.119 / 1M blended on deepinfra (50% off). The frontier pick is Muse Spark 1.3.
- What does ZDR + no training mean?
- ZDR is zero data retention: the provider does not keep prompts. No training means prompts are not used to train models. The weekly cards badge ZDR when the priced route qualifies — not when the model family merely has a ZDR provider somewhere. A privacy route requires the catalog to mark both `zdr` and `no_training` as all or some, matching ?zdr=true and ?npt=true. `some` means only certain providers behind a model id qualify, so the ZDR price is the cheapest ZDR endpoint, even when a cheaper non-ZDR route exists.
- How is bang-for-buck calculated?
- Bang-for-buck is a DeepsecBench run's score divided by that same run's cost — each run is kept intact, so a model's best score and its most efficient run are reported separately instead of mixed together. Adopted models win the homepage slot when any qualify, so a 0-token catalog row cannot. Value score is 7-day mean token share (missing days count as zero) divided by blended $/1M, using a 3× input + 1× output mix. On sale matches Vercel's official list-vs-sale promo on the models page — a cheaper third-party endpoint is a routing price, not a discount.
- Which models count as capable?
- Capable models support tool use and have at least 128,000 tokens of context. Vision is not required. Cheap routers blend at or under $0.50 / 1M and must clear a quality floor: Deepsec ≥ 12 or an Artificial Analysis intelligence/coding index ≥ 40. Workhorses are the capable models people actually run (at least 3% token share of the 7-day mean, on both pages) at or under $6.00 / 1M, excluding the cheap-router family, ranked on everyday Deepsec — AA intelligence is the fallback when nobody in that pool is benchmarked. Frontier is the highest AA intelligence among capable models at or under $6.00 / 1M. Rising is the leftover capable model in that same usable band, ranked on AA first so a high-quality catalog row the other roles missed can surface. Weekly fallback is week-over-week token-share growth; daily fallback is complete-day share minus the 7-day mean.
- What does the AA number mean?
- AA intel / coding / agentic are Artificial Analysis headline indices, fetched from the Artificial Analysis Free API (OpenRouter is the fallback). They are never averaged with DeepsecBench. Frontier and rising rank on AA intelligence inside the usable price band (at or under $6 / 1M) — coding, then Deepsec, then cheaper blend break ties. A $10 model does not take a default slot. Workhorse still uses the everyday Deepsec run first, with AA as the fallback. Cheap routers may qualify on an AA floor when they have no Deepsec run. Ranked boards show the top 20 of each metric: intelligence, coding, bang, Deepsec score, on sale, cheap, tokens, and spend.
- Is this an official Vercel product?
- No. bestmodels.dev is an independent ranking (daily home, weekly at /week). Catalog, adoption, and DeepsecBench numbers come from Vercel AI Gateway data licensed CC BY 4.0. We are not affiliated with Vercel.
Weekly archives
Sep 15–21, 2026 · Sep 8–14, 2026 · Sep 1–7, 2026 · Aug 26–Sep 1, 2026 · Aug 25–31, 2026