2026-08-28
150 change(s) in the model market
added to the catalog
- Mistral: Codestral 2508 (batch) (mistralai/codestral-2508:batch) added to the catalog: 256,000 ctx, $0.3/M in / $0.9/M out
- Mistral: Ministral 3 8B 2512 (batch) (mistralai/ministral-8b-2512:batch) added to the catalog: 262,144 ctx, $0.15/M in / $0.15/M out
- Mistral: Mistral Large 3 2512 (batch) (mistralai/mistral-large-2512:batch) added to the catalog: 262,144 ctx, $0.5/M in / $1.5/M out
- Mistral: Mistral Medium 3.5 (batch) (mistralai/mistral-medium-3-5:batch) added to the catalog: 262,144 ctx, $0.75/M in / $3.75/M out
- Mistral: Mistral Medium 3.1 (batch) (mistralai/mistral-medium-3.1:batch) added to the catalog: 131,072 ctx, $0.4/M in / $2/M out
- Mistral: Mistral Small 4 (batch) (mistralai/mistral-small-2603:batch) added to the catalog: 262,144 ctx, $0.15/M in / $0.6/M out
- Qwen: Qwen3.8 2.4T A95B (batch) (qwen/qwen3.8-2.4t-a95b:batch) added to the catalog: 1,010,000 ctx, $2/M in / $6/M out
- Tencent: Hy4 preview (tencent/hy4-preview) added to the catalog: 1,048,576 ctx, $0.834/M in / $2.5/M out
removed from the catalog
- Mistral: Ministral 8B (mistralai/ministral-8b) removed from the catalog
- MoonshotAI: Kimi K2.7 Code (batch) (moonshotai/kimi-k2.7-code:batch) removed from the catalog
- OpenAI: GPT-3.5 Turbo (batch) (openai/gpt-3.5-turbo:batch) removed from the catalog
- OpenAI: GPT-4 Turbo (batch) (openai/gpt-4-turbo:batch) removed from the catalog
- OpenAI: GPT-4.1 Mini (batch) (openai/gpt-4.1-mini:batch) removed from the catalog
- OpenAI: GPT-4.1 Nano (batch) (openai/gpt-4.1-nano:batch) removed from the catalog
- OpenAI: GPT-4.1 (batch) (openai/gpt-4.1:batch) removed from the catalog
- OpenAI: GPT-4o-mini (batch) (openai/gpt-4o-mini:batch) removed from the catalog
- OpenAI: GPT-4o (batch) (openai/gpt-4o:batch) removed from the catalog
- OpenAI: GPT-5 Codex (batch) (openai/gpt-5-codex:batch) removed from the catalog
- OpenAI: GPT-5 Mini (batch) (openai/gpt-5-mini:batch) removed from the catalog
- OpenAI: GPT-5 Nano (batch) (openai/gpt-5-nano:batch) removed from the catalog
- OpenAI: GPT-5 Pro (batch) (openai/gpt-5-pro:batch) removed from the catalog
- OpenAI: GPT-5 (batch) (openai/gpt-5:batch) removed from the catalog
- OpenAI: GPT-5.1 (batch) (openai/gpt-5.1:batch) removed from the catalog
- OpenAI: GPT-5.2 Pro (batch) (openai/gpt-5.2-pro:batch) removed from the catalog
- OpenAI: GPT-5.2 (batch) (openai/gpt-5.2:batch) removed from the catalog
- OpenAI: GPT-5.4 Mini (batch) (openai/gpt-5.4-mini:batch) removed from the catalog
- OpenAI: GPT-5.4 Nano (batch) (openai/gpt-5.4-nano:batch) removed from the catalog
- OpenAI: GPT-5.4 Pro (batch) (openai/gpt-5.4-pro:batch) removed from the catalog
- OpenAI: GPT-5.4 (batch) (openai/gpt-5.4:batch) removed from the catalog
- OpenAI: GPT-5.5 Pro (batch) (openai/gpt-5.5-pro:batch) removed from the catalog
- OpenAI: GPT-5.5 (batch) (openai/gpt-5.5:batch) removed from the catalog
- OpenAI: GPT-5.6 Luna Pro (batch) (openai/gpt-5.6-luna-pro:batch) removed from the catalog
- OpenAI: GPT-5.6 Luna (batch) (openai/gpt-5.6-luna:batch) removed from the catalog
- OpenAI: GPT-5.6 Sol Pro (batch) (openai/gpt-5.6-sol-pro:batch) removed from the catalog
- OpenAI: GPT-5.6 Sol (batch) (openai/gpt-5.6-sol:batch) removed from the catalog
- OpenAI: GPT-5.6 Terra Pro (batch) (openai/gpt-5.6-terra-pro:batch) removed from the catalog
- OpenAI: GPT-5.6 Terra (batch) (openai/gpt-5.6-terra:batch) removed from the catalog
- OpenAI: o1-pro (batch) (openai/o1-pro:batch) removed from the catalog
- OpenAI: o1 (batch) (openai/o1:batch) removed from the catalog
- OpenAI: o3 Mini High (batch) (openai/o3-mini-high:batch) removed from the catalog
- OpenAI: o3 Mini (batch) (openai/o3-mini:batch) removed from the catalog
- OpenAI: o3 Pro (batch) (openai/o3-pro:batch) removed from the catalog
- OpenAI: o3 (batch) (openai/o3:batch) removed from the catalog
- OpenAI: o4 Mini High (batch) (openai/o4-mini-high:batch) removed from the catalog
- OpenAI: o4 Mini (batch) (openai/o4-mini:batch) removed from the catalog
- TheDrummer: Rocinante 12B (thedrummer/rocinante-12b) removed from the catalog
model price changes
- ~z-ai/glm-latest: prompt price $1.4/M → $1.25/M (−10.7%)
- deepseek/deepseek-v3.2: prompt price $0.26/M → $0.269/M (+3.5%)
- deepseek/deepseek-v3.2: completion price $0.38/M → $0.4/M (+5.3%)
- deepseek/deepseek-v4-flash: prompt price $0.078/M → $0.087/M (+11.3%)
- deepseek/deepseek-v4-flash: completion price $0.156/M → $0.174/M (+11.3%)
- deepseek/deepseek-v4-pro: prompt price $0.87/M → $0.762/M (−12.4%)
- deepseek/deepseek-v4-pro: completion price $1.74/M → $1.52/M (−12.4%)
- deepseek/deepseek-v4-pro-0813: prompt price $1.12/M → $0.66/M (−41.2%)
- deepseek/deepseek-v4-pro-0813: completion price $3.37/M → $1.98/M (−41.2%)
- nvidia/nemotron-3-ultra-550b-a55b: prompt price $0.6/M → $0.5/M (−16.7%)
- nvidia/nemotron-3-ultra-550b-a55b: completion price $3.6/M → $2.2/M (−38.9%)
- nvidia/nemotron-3.5-lightning: prompt price $0.08/M → $0.1/M (+25.0%)
- nvidia/nemotron-3.5-lightning: completion price $0.2/M → $0.25/M (+25.0%)
- qwen/qwen3.5-122b-a10b: prompt price $0.26/M → $0.29/M (+11.5%)
- qwen/qwen3.5-122b-a10b: completion price $2.08/M → $2.4/M (+15.4%)
model context changes
- ~z-ai/glm-latest: context length 1,048,576 → 1,310,720
- google/gemma-3-27b-it: context length 262,144 → 131,072
- kwaipilot/kat-coder-pro-v2.5: context length 256,000 → 262,144
- mistralai/voxtral-small-24b-2507: context length 32,000 → 32,768
- nvidia/nemotron-3-ultra-550b-a55b: context length 512,288 → 262,144
- z-ai/glm-5.3: context length 1,048,576 → 1,310,720
quantization changes
- AkashML: qwen/qwen3.8-27b quantization bf16 → fp8
provider price changes
- DigitalOcean: deepseek/deepseek-v3.2 prompt price $0.25/M → $0.5/M (+100.0%)
- DigitalOcean: deepseek/deepseek-v3.2 completion price $0.8/M → $1.6/M (+100.0%)
- Baidu: deepseek/deepseek-v4-flash prompt price $0.079/M → $0.087/M (+9.5%)
- Baidu: deepseek/deepseek-v4-flash completion price $0.159/M → $0.174/M (+9.5%)
- DigitalOcean: deepseek/deepseek-v4-flash prompt price $0.068/M → $0.14/M (+106.2%)
- DigitalOcean: deepseek/deepseek-v4-flash completion price $0.168/M → $0.28/M (+66.7%)
- Mancer 2: deepseek/deepseek-v4-flash prompt price $0.15/M → $0.16/M (+6.7%)
- StreamLake: deepseek/deepseek-v4-flash prompt price $0.078/M → $0.087/M (+11.3%)
- StreamLake: deepseek/deepseek-v4-flash completion price $0.156/M → $0.174/M (+11.3%)
- AkashML: deepseek/deepseek-v4-flash-0731 prompt price $0.14/M → $0.1/M (−28.6%)
- Baidu: deepseek/deepseek-v4-flash-0731 prompt price $0.06/M → $0.07/M (+16.6%)
- Baidu: deepseek/deepseek-v4-flash-0731 completion price $0.12/M → $0.14/M (+16.6%)
- DigitalOcean: deepseek/deepseek-v4-flash-0731 prompt price $0.08/M → $0.14/M (+75.0%)
- DigitalOcean: deepseek/deepseek-v4-flash-0731 completion price $0.252/M → $0.28/M (+11.1%)
- Mancer 2: deepseek/deepseek-v4-flash-0731 prompt price $0.15/M → $0.16/M (+6.7%)
- Relace: deepseek/deepseek-v4-flash-0731 prompt price $0.05/M → $0.06/M (+20.0%)
- Relace: deepseek/deepseek-v4-flash-0731 completion price $0.1/M → $0.12/M (+20.0%)
- StreamLake: deepseek/deepseek-v4-flash-0731 prompt price $0.22/M → $0.088/M (−60.0%)
- StreamLake: deepseek/deepseek-v4-flash-0731 completion price $0.66/M → $0.264/M (−60.0%)
- Baidu: deepseek/deepseek-v4-pro prompt price $0.791/M → $0.771/M (−2.6%)
- Baidu: deepseek/deepseek-v4-pro completion price $1.58/M → $1.54/M (−2.6%)
- DigitalOcean: deepseek/deepseek-v4-pro prompt price $0.87/M → $1.74/M (+100.0%)
- DigitalOcean: deepseek/deepseek-v4-pro completion price $1.74/M → $3.48/M (+100.0%)
- StreamLake: deepseek/deepseek-v4-pro prompt price $0.87/M → $0.762/M (−12.4%)
- StreamLake: deepseek/deepseek-v4-pro completion price $1.74/M → $1.52/M (−12.4%)
- StreamLake: deepseek/deepseek-v4-pro-0813 prompt price $0.66/M → $0.579/M (−12.2%)
- StreamLake: deepseek/deepseek-v4-pro-0813 completion price $1.98/M → $1.74/M (−12.2%)
- DigitalOcean: meta-llama/llama-4-maverick prompt price $0.2/M → $0.25/M (+25.0%)
- DigitalOcean: meta-llama/llama-4-maverick completion price $0.696/M → $0.87/M (+25.0%)
- Baidu: moonshotai/kimi-k2.6 prompt price $0.548/M → $0.529/M (−3.5%)
- Baidu: moonshotai/kimi-k2.6 completion price $2.31/M → $2.23/M (−3.5%)
- Decart: moonshotai/kimi-k2.6 prompt price $0.548/M → $0.537/M (−2.0%)
- Decart: moonshotai/kimi-k2.6 completion price $2.31/M → $2.26/M (−2.0%)
- Inceptron: moonshotai/kimi-k2.6 prompt price $0.54/M → $0.53/M (−1.9%)
- DigitalOcean: moonshotai/kimi-k3 prompt price $2.85/M → $3/M (+5.3%)
- DigitalOcean: moonshotai/kimi-k3 completion price $14.3/M → $15/M (+5.3%)
- Amazon Bedrock: openai/gpt-5.6-sol prompt price $5.5/M → $4.4/M (−20.0%)
- Amazon Bedrock: openai/gpt-5.6-sol completion price $33/M → $22/M (−33.3%)
- DigitalOcean: openai/gpt-oss-120b prompt price $0.055/M → $0.1/M (+81.8%)
- DigitalOcean: openai/gpt-oss-120b completion price $0.385/M → $0.7/M (+81.8%)
- Mancer 2: openai/gpt-oss-120b prompt price $0.08/M → $0.085/M (+6.2%)
- Darkbloom: qwen/qwen3.6-35b-a3b prompt price $0.07/M → $0.05/M (−28.6%)
- DigitalOcean: xiaomi/mimo-v2.5-pro prompt price $0.4/M → $0.8/M (+100.0%)
- DigitalOcean: xiaomi/mimo-v2.5-pro completion price $1.5/M → $3/M (+100.0%)
- DeepInfra: z-ai/glm-5.2 prompt price $0.75/M → $0.487/M (−35.0%)
- DeepInfra: z-ai/glm-5.2 completion price $2.4/M → $1.56/M (−35.0%)
- DigitalOcean: z-ai/glm-5.2 prompt price $0.7/M → $1.4/M (+100.0%)
- DigitalOcean: z-ai/glm-5.2 completion price $2.2/M → $4.4/M (+100.0%)
- Inceptron: z-ai/glm-5.2 prompt price $0.75/M → $0.71/M (−5.3%)
- Inceptron: z-ai/glm-5.2 completion price $2.4/M → $2.35/M (−2.1%)
providers added
- NextBit now serves deepseek/deepseek-v4-flash (fp8, 1,048,576 ctx, $0.14/M in / $0.28/M out)
- OpenInference now serves google/gemma-4-31b-it (bf16, 262,144 ctx, $0.08/M in / $0.35/M out)
- AtlasCloud now serves kwaipilot/kat-coder-pro-v2.5 (fp8, 262,144 ctx, $0.74/M in / $2.96/M out)
- Mistral now serves mistralai/voxtral-small-24b-2507 (unknown, 32,768 ctx, $0.1/M in / $0.3/M out)
- Mistral now serves mistralai/voxtral-small-24b-2507 (unknown, 32,000 ctx, $0.11/M in / $0.33/M out)
- Sail Research now serves moonshotai/kimi-k3 (fp4, 974,842 ctx, $2.6/M in / $13/M out)
- Novita now serves qwen/qwen3.8-2.4t-a95b (unknown, 1,000,000 ctx, $2/M in / $6/M out)
- Novita now serves qwen/qwen3.8-27b (unknown, 1,000,000 ctx, $0.42/M in / $3/M out)
- BaseTen now serves z-ai/glm-5.3 (fp4, 1,048,576 ctx, $1.4/M in / $4.4/M out)
- Cloudflare now serves z-ai/glm-5.3 (unknown, 1,310,720 ctx, $1.4/M in / $4.4/M out)
- DeepInfra now serves z-ai/glm-5.3 (bf16, 1,048,576 ctx, $1.4/M in / $4.4/M out)
- Fireworks now serves z-ai/glm-5.3 (unknown, 1,048,576 ctx, $1.4/M in / $4.4/M out)
- Friendli now serves z-ai/glm-5.3 (unknown, 1,048,576 ctx, $1.4/M in / $4.4/M out)
- GMICloud now serves z-ai/glm-5.3 (fp8, 1,048,576 ctx, $1.4/M in / $4.4/M out)
- Io Net now serves z-ai/glm-5.3 (fp8, 262,144 ctx, $1.25/M in / $4.4/M out)
- DigitalOcean now serves z-ai/glm-5.3-flash (unknown, 1,048,576 ctx, $0.15/M in / $0.5/M out)
- Friendli now serves z-ai/glm-5.3-flash (unknown, 1,048,576 ctx, $0.15/M in / $0.5/M out)
- Morph now serves z-ai/glm-5.3-flash (fp8, 1,048,576 ctx, $0.15/M in / $0.5/M out)
- Phala now serves z-ai/glm-5.3-flash (fp8, 1,048,576 ctx, $0.15/M in / $0.5/M out)
- Wafer now serves z-ai/glm-5.3-flash (unknown, 1,048,576 ctx, $0.15/M in / $0.5/M out)
providers removed
- Together no longer serves deepseek/deepseek-v4-pro (was unknown)
- Phala no longer serves google/gemma-3-27b-it (was unknown)
- Novita no longer serves inclusionai/ling-3.0-flash (was unknown)
- Together no longer serves minimax/minimax-m3 (was unknown)
- Mistral no longer serves mistralai/mistral-small-2603 (was unknown)
- Together no longer serves nvidia/nemotron-3-ultra-550b-a55b (was unknown)
- Morph no longer serves qwen/qwen3.5-397b-a17b (was fp4)
- Morph no longer serves qwen/qwen3.6-27b (was fp4)
- Morph no longer serves qwen/qwen3.8-27b (was fp4)
- AkashML no longer serves z-ai/glm-5.2 (was fp8)
- Makora no longer serves z-ai/glm-5.2 (was fp8)
- Wafer no longer serves z-ai/glm-5.2 (was fp4)
openrouter catalog snapshots 2026-08-27 → 2026-08-28 · schema 1 · generated 2026-08-28T17:27:01.768Z
$ curl variables.md — this site speaks markdown to your terminal.