2026-07-18
48 change(s) in the model market
added to the catalog
- Auto Router (Beta) (openrouter/auto-beta) added to the catalog: 2,000,000 ctx, $-1000000/M in / $-1000000/M out
- Thinking Machines: Inkling (thinkingmachines/inkling) added to the catalog: 1,048,576 ctx, $1/M in / $4.05/M out
removed from the catalog
- Meta: Llama 3.2 11B Vision Instruct (meta-llama/llama-3.2-11b-vision-instruct) removed from the catalog
- NVIDIA: Llama 3.3 Nemotron Super 49B V1.5 (nvidia/llama-3.3-nemotron-super-49b-v1.5) removed from the catalog
model price changes
- mancer/weaver: prompt price $0.75/M → $0.5/M (−33.3%)
- mancer/weaver: completion price $1/M → $0.75/M (−25.0%)
- moonshotai/kimi-k2.7-code: prompt price $0.75/M → $1/M (+33.3%)
- moonshotai/kimi-k2.7-code: completion price $3.5/M → $4.4/M (+25.7%)
- qwen/qwen3-14b: prompt price $0.1/M → $0.12/M (+20.0%)
- qwen/qwen3-30b-a3b: prompt price $0.12/M → $0.13/M (+8.3%)
- qwen/qwen3-30b-a3b: completion price $0.5/M → $0.52/M (+4.0%)
- qwen/qwen3.5-27b: prompt price $0.195/M → $0.26/M (+33.3%)
- qwen/qwen3.5-27b: completion price $1.56/M → $2.6/M (+66.7%)
- z-ai/glm-4.7-flash: prompt price $0.06/M → $0.061/M (+0.8%)
- z-ai/glm-5.2: prompt price $0.967/M → $0.343/M (−64.5%)
- z-ai/glm-5.2: completion price $3.04/M → $1.08/M (−64.5%)
model context changes
- z-ai/glm-4.7-flash: context length 202,752 → 200,000
quantization changes
- Ambient: stepfun/step-3.7-flash quantization fp8 → fp4
provider price changes
- Mancer 2: gryphe/mythomax-l2-13b prompt price $0.5/M → $0.45/M (−10.0%)
- Mancer 2: gryphe/mythomax-l2-13b completion price $0.75/M → $0.65/M (−13.3%)
- Mancer 2: mancer/weaver prompt price $0.75/M → $0.5/M (−33.3%)
- Mancer 2: mancer/weaver completion price $1/M → $0.75/M (−25.0%)
- DekaLLM: mistralai/mistral-nemo prompt price $0.02/M → $0.018/M (−10.0%)
- Ambient: moonshotai/kimi-k2.7-code prompt price $0.75/M → $1/M (+33.3%)
- Ambient: moonshotai/kimi-k2.7-code completion price $3.5/M → $4.4/M (+25.7%)
- Mancer 2: openai/gpt-oss-120b prompt price $0.045/M → $0.04/M (−11.1%)
- Mancer 2: openai/gpt-oss-120b completion price $0.375/M → $0.35/M (−6.7%)
- Io Net: qwen/qwen3.6-27b prompt price $0.31/M → $0.341/M (+10.0%)
- Io Net: qwen/qwen3.6-27b completion price $2.61/M → $2.87/M (+10.0%)
- Mancer 2: undi95/remm-slerp-l2-13b prompt price $0.5/M → $0.45/M (−10.0%)
- Mancer 2: undi95/remm-slerp-l2-13b completion price $0.75/M → $0.65/M (−13.3%)
- Baidu: z-ai/glm-5.2 prompt price $0.97/M → $0.42/M (−56.7%)
- Baidu: z-ai/glm-5.2 completion price $3.07/M → $1.32/M (−57.0%)
- Io Net: z-ai/glm-5.2 prompt price $1.94/M → $2.13/M (+10.0%)
- Io Net: z-ai/glm-5.2 completion price $4.24/M → $4.66/M (+10.0%)
- Novita: z-ai/glm-5.2 prompt price $0.965/M → $0.343/M (−64.4%)
- Novita: z-ai/glm-5.2 completion price $3.03/M → $1.08/M (−64.4%)
- StreamLake: z-ai/glm-5.2 prompt price $0.967/M → $0.342/M (−64.6%)
- StreamLake: z-ai/glm-5.2 completion price $3.04/M → $1.08/M (−64.6%)
provider context changes
- AkashML: z-ai/glm-5.2 context length 131,072 → 96,890
providers added
- Ambient now serves deepseek/deepseek-v4-flash (fp4, 1,048,576 ctx, $0.25/M in / $0.4/M out)
- Ionstream now serves deepseek/deepseek-v4-pro (fp4, 1,048,576 ctx, $1.13/M in / $2.26/M out)
- Ambient now serves moonshotai/kimi-k2.5 (int4, 262,144 ctx, $0.6/M in / $3/M out)
- BaseTen now serves nvidia/nemotron-3-ultra-550b-a55b (fp4, 202,800 ctx, $0.6/M in / $2.4/M out)
providers removed
- NextBit no longer serves deepseek/deepseek-v4-flash (was fp8)
- DeepInfra no longer serves minimax/minimax-m2.5 (was fp8)
- ModelRun no longer serves moonshotai/kimi-k2.5 (was fp4)
- ModelRun no longer serves moonshotai/kimi-k2.7-code (was fp4)
openrouter catalog snapshots 2026-07-17 → 2026-07-18 · schema 1 · generated 2026-07-18T07:09:42.623Z
$ curl variables.md — this site speaks markdown to your terminal.