Every AI model price, context and end-of-life change detected on 2026-08-01.
| Change | Model | Detail |
|---|---|---|
| END OF LIFE MOVED | z-ai/glm-4.5v | 2026-12-31 -> - |
| PRICE CHANGE | deepseek/deepseek-v4-flash-0731 | output -35.7% · cache read +542.9% · input -35.7% |
| PRICE CHANGE | z-ai/glm-5.2 | output -41.4% · cache read -41.4% · input -41.4% |
| PRICE CHANGE | moonshotai/kimi-k2.6 | output +37.5% · cache read +101.6% · input +1.9% |
| PRICE CHANGE | openai/gpt-oss-20b | output -7.1% · cache read now charged |
| PRICE CHANGE | qwen/qwen3-coder-30b-a3b-instruct | output +3.7% |
| PRICE CHANGE | thinkingmachines/inkling-small | output -16.7% · cache read -13.8% · input -13.8% |
| PRICE CHANGE | z-ai/glm-5.2 | output -36.0% · cache read -36.0% · input -36.0% |
| PRICE CHANGE | mistralai/mistral-small-3.2-24b-instruct | output -33.3% · cache read no longer charged · input -25.0% |
| PRICE CHANGE | moonshotai/kimi-k2.6 | output -38.0% · cache read -38.0% · input -38.0% |
| PRICE CHANGE | nvidia/nemotron-3-nano-30b-a3b | cache read +20.0% |
| PRICE CHANGE | openai/gpt-oss-20b | output +7.7% · cache read no longer charged |
| PRICE CHANGE | poolside/laguna-s-2.1 | output -10.0% · cache read -10.0% · input -10.0% |
| PRICE CHANGE | qwen/qwen3-235b-a22b-thinking-2507 | output -23.3% · input -23.3% |
| PRICE CHANGE | qwen/qwen3-vl-30b-a3b-instruct | output -13.3% · input -13.3% |
| PRICE CHANGE | z-ai/glm-5.2 | output +15.9% · cache read +15.9% · input +15.9% |
| NEW MODEL | ~deepseek/deepseek-v4-flash-latest | context 1,048,576 |
| NEW MODEL | deepseek/deepseek-v4-flash-0731 | context 1,048,576 |
| NEW MODEL | thinkingmachines/inkling-small | context 524,288 |
| MODEL REMOVED | mistralai/devstral-2512 | |
| MODEL REMOVED | openai/gpt-5.1-chat | |
| FROM THE SOURCE | Tibo @thsottiaux | @RyanEls4 Codex |
| FROM THE SOURCE | Tibo @thsottiaux | @_chenglou I was part of that team. Basically ChatGPT one year before it came out. Called LMChat and then another codename. Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. I th |
| RELEASE | b10219 | |
| OUTAGE | Incident with Copilot AI Model Providers | |
| PARAMETERS CHANGED | deepseek/deepseek-v4-flash-0731 | added top_a |
| FROM THE SOURCE | Elon Musk @elonmusk | Grok 4.5 is Pareto #1 when considering speed & cost https://t.co/MoIM9hpHWr |
| FROM THE SOURCE | Tibo @thsottiaux | @NanneWielinga @gdb All of them actually. The subscription is very generous in tokens. |
| RELEASE | b10218 | |
| ANNOUNCEMENT | Gemini 2.5 Pro and Gemini 3 Flash deprecated | |
| OUTAGE | Degraded availability GPT 5.6 Luna | |
| ANNOUNCEMENT | Ten advances in mathematics and theoretical computer science | |
| PARAMETERS CHANGED | sao10k/l3-lunaris-8b | added logprobs, top_logprobs |
| PARAMETERS CHANGED | thinkingmachines/inkling-small | added logprobs, tool_choice, tools, top_logprobs |
| CATALOG_VARIANT_DELISTED | ||
| PARAMETERS CHANGED | poolside/laguna-s-2.1 | removed stop |
| CONTEXT CHANGED | thedrummer/unslopnemo-12b | 32768 -> 1024000 |
| PARAMETERS CHANGED | thedrummer/unslopnemo-12b | added logit_bias, top_k |