2026-07-15
86 change(s) in the model market
removed from the catalog
- Sao10K: Llama 3.1 70B Hanami x1 (sao10k/l3.1-70b-hanami-x1) removed from the catalog
model price changes
- deepseek/deepseek-chat-v3-0324: prompt price $0.24/M → $0.27/M (+12.5%)
- deepseek/deepseek-chat-v3-0324: completion price $0.9/M → $1.12/M (+24.4%)
- deepseek/deepseek-v3.1-terminus: completion price $0.95/M → $1/M (+5.3%)
- deepseek/deepseek-v3.2: prompt price $0.214/M → $0.269/M (+25.4%)
- deepseek/deepseek-v3.2: completion price $0.322/M → $0.4/M (+24.3%)
- deepseek/deepseek-v4-flash: prompt price $0.09/M → $0.098/M (+8.9%)
- deepseek/deepseek-v4-flash: completion price $0.18/M → $0.196/M (+8.9%)
- google/gemma-3-27b-it: completion price $0.16/M → $0.45/M (+181.2%)
- google/gemma-4-26b-a4b-it: prompt price $0.06/M → $0.1/M (+66.7%)
- google/gemma-4-26b-a4b-it: completion price $0.33/M → $0.3/M (−9.1%)
- google/gemma-4-31b-it: prompt price $0.06/M → $0.12/M (+100.0%)
- google/gemma-4-31b-it: completion price $0.35/M → $0.37/M (+5.7%)
- meta-llama/llama-3.1-8b-instruct: prompt price $0.02/M → $0.05/M (+150.0%)
- meta-llama/llama-3.1-8b-instruct: completion price $0.03/M → $0.08/M (+166.7%)
- meta-llama/llama-3.2-3b-instruct: prompt price $0.05/M → $0.051/M (+1.8%)
- meta-llama/llama-3.2-3b-instruct: completion price $0.33/M → $0.335/M (+1.5%)
- meta-llama/llama-3.3-70b-instruct: prompt price $0.1/M → $0.13/M (+30.0%)
- meta-llama/llama-3.3-70b-instruct: completion price $0.32/M → $0.4/M (+25.0%)
- minimax/minimax-m1: prompt price $0.4/M → $0.55/M (+37.5%)
- minimax/minimax-m2.7: prompt price $0.24/M → $0.3/M (+25.0%)
- minimax/minimax-m2.7: completion price $0.96/M → $1.2/M (+25.0%)
- mistralai/mistral-nemo: completion price $0.03/M → $0.04/M (+33.3%)
- mistralai/mistral-small-3.2-24b-instruct: prompt price $0.075/M → $0.1/M (+33.3%)
- mistralai/mistral-small-3.2-24b-instruct: completion price $0.2/M → $0.3/M (+50.0%)
- moonshotai/kimi-k2.5: prompt price $0.375/M → $0.57/M (+52.0%)
- moonshotai/kimi-k2.5: completion price $2.02/M → $2.85/M (+40.7%)
- nvidia/nemotron-3-super-120b-a12b: prompt price $0.08/M → $0.21/M (+162.5%)
- nvidia/nemotron-3-super-120b-a12b: completion price $0.45/M → $0.455/M (+1.1%)
- nvidia/nemotron-3-ultra-550b-a55b: prompt price $0.5/M → $0.6/M (+20.0%)
- nvidia/nemotron-3-ultra-550b-a55b: completion price $2.2/M → $3.6/M (+63.6%)
- openai/gpt-oss-120b: prompt price $0.03/M → $0.037/M (+23.3%)
- openai/gpt-oss-120b: completion price $0.15/M → $0.17/M (+13.3%)
- openai/gpt-oss-20b: prompt price $0.029/M → $0.03/M (+3.4%)
- openai/gpt-oss-20b: completion price $0.14/M → $0.13/M (−7.1%)
- qwen/qwen2.5-vl-72b-instruct: prompt price $0.25/M → $0.8/M (+220.0%)
- qwen/qwen2.5-vl-72b-instruct: completion price $0.75/M → $1/M (+33.3%)
- qwen/qwen3-30b-a3b-instruct-2507: prompt price $0.048/M → $0.1/M (+107.7%)
- qwen/qwen3-30b-a3b-instruct-2507: completion price $0.193/M → $0.3/M (+55.4%)
- qwen/qwen3-coder: prompt price $0.22/M → $0.3/M (+36.4%)
- qwen/qwen3-coder: completion price $1.8/M → $1/M (−44.4%)
- qwen/qwen3-coder-next: prompt price $0.11/M → $0.12/M (+9.1%)
- qwen/qwen3-next-80b-a3b-instruct: prompt price $0.09/M → $0.1/M (+11.1%)
- qwen/qwen3-vl-235b-a22b-instruct: prompt price $0.2/M → $0.21/M (+5.0%)
- qwen/qwen3-vl-235b-a22b-instruct: completion price $0.88/M → $1.9/M (+115.9%)
- qwen/qwen3.5-397b-a17b: prompt price $0.385/M → $0.45/M (+16.9%)
- qwen/qwen3.5-397b-a17b: completion price $2.45/M → $3/M (+22.4%)
- qwen/qwen3.6-27b: prompt price $0.289/M → $0.45/M (+55.7%)
- qwen/qwen3.6-27b: completion price $2.4/M → $2.7/M (+12.5%)
- tencent/hy3: prompt price $0.14/M → $0.2/M (+42.9%)
- tencent/hy3: completion price $0.58/M → $0.8/M (+37.9%)
- xiaomi/mimo-v2.5: prompt price $0.105/M → $0.14/M (+33.3%)
- z-ai/glm-4.6: prompt price $0.43/M → $0.5/M (+16.3%)
- z-ai/glm-4.6: completion price $1.75/M → $2/M (+14.3%)
- z-ai/glm-4.7-flash: prompt price $0.06/M → $0.061/M (+0.8%)
- z-ai/glm-5: prompt price $0.6/M → $0.95/M (+58.3%)
- z-ai/glm-5: completion price $1.92/M → $3.15/M (+64.1%)
- z-ai/glm-5.2: prompt price $0.928/M → $0.895/M (−3.6%)
- z-ai/glm-5.2: completion price $2.92/M → $2.81/M (−3.6%)
model context changes
- deepseek/deepseek-v3.1-terminus: context length 163,840 → 131,072
- deepseek/deepseek-v3.2: context length 131,072 → 163,840
- mistralai/mistral-small-3.2-24b-instruct: context length 128,000 → 131,072
- qwen/qwen3-30b-a3b-instruct-2507: context length 131,072 → 262,144
- qwen/qwen3-vl-235b-a22b-instruct: context length 262,144 → 131,072
- qwen/qwen3.5-397b-a17b: context length 256,000 → 262,144
- z-ai/glm-4.6: context length 200,000 → 202,752
- z-ai/glm-4.7-flash: context length 202,752 → 200,000
- z-ai/glm-5-turbo: context length 262,144 → 202,752
provider price changes
- NextBit: deepseek/deepseek-v4-flash prompt price $0.14/M → $0.18/M (+28.6%)
- NextBit: deepseek/deepseek-v4-flash completion price $0.28/M → $0.35/M (+25.0%)
- Io Net: qwen/qwen3.6-27b prompt price $0.314/M → $0.345/M (+10.0%)
- Io Net: qwen/qwen3.6-27b completion price $2.64/M → $2.9/M (+10.0%)
- Ambient: z-ai/glm-5.2 prompt price $1.4/M → $1.05/M (−25.0%)
- Io Net: z-ai/glm-5.2 prompt price $1.59/M → $1.43/M (−10.0%)
- Io Net: z-ai/glm-5.2 completion price $3.52/M → $3.17/M (−10.0%)
- Novita: z-ai/glm-5.2 prompt price $0.928/M → $0.893/M (−3.8%)
- Novita: z-ai/glm-5.2 completion price $2.92/M → $2.81/M (−3.8%)
- StreamLake: z-ai/glm-5.2 prompt price $0.928/M → $0.895/M (−3.6%)
- StreamLake: z-ai/glm-5.2 completion price $2.92/M → $2.81/M (−3.6%)
providers added
- GMICloud now serves xiaomi/mimo-v2.5 (fp8, 1,050,000 ctx, $0.32/M in / $1.6/M out)
- Novita now serves xiaomi/mimo-v2.5 (fp8, 1,048,576 ctx, $0.168/M in / $0.336/M out)
providers removed
- OpenInference no longer serves google/gemma-4-31b-it (was bf16)
- Nebius no longer serves minimax/minimax-m2.5 (was fp4)
- Nebius no longer serves moonshotai/kimi-k2.6 (was int4)
- Infermatic no longer serves thedrummer/rocinante-12b (was bf16)
- AtlasCloud no longer serves z-ai/glm-5-turbo (was fp8)
openrouter catalog snapshots 2026-07-14 → 2026-07-15 · schema 1 · generated 2026-07-15T07:24:37.732Z
$ curl variables.md — this site speaks markdown to your terminal.