Every AI model price, context and end-of-life change detected on 2026-08-19.
| Change | Model | Detail |
|---|---|---|
| END OF LIFE | inclusionai/ling-2.6-1t | shuts down 2026-08-24 · 5 days left |
| END OF LIFE | inclusionai/ling-2.6-flash | shuts down 2026-08-24 · 5 days left |
| END OF LIFE | inclusionai/ring-2.6-1t | shuts down 2026-08-24 · 5 days left |
| END OF LIFE | nvidia/nemotron-nano-12b-v2-vl:free | shuts down 2026-08-24 · 5 days left |
| END OF LIFE | z-ai/glm-4.5 | shuts down 2026-12-31 · 134 days left |
| END OF LIFE ANNOUNCED | inclusionai/ling-2.6-1t | - -> 2026-08-24 |
| END OF LIFE ANNOUNCED | inclusionai/ling-2.6-flash | - -> 2026-08-24 |
| END OF LIFE ANNOUNCED | inclusionai/ring-2.6-1t | - -> 2026-08-24 |
| PRICE CHANGE | deepseek/deepseek-chat-v3-0324 | output -10.7% · cache read no longer charged · input -7.4% |
| PRICE CHANGE | deepseek/deepseek-v4-flash | output +3.5% · cache read +3.5% · input +3.5% |
| PRICE CHANGE | deepseek/deepseek-v4-pro | output +45.5% · cache read +452.3% · overrides no longer charged · input +118.2% |
| PRICE CHANGE | google/gemma-4-31b-it | cache read -50.0% · input -10.0% |
| PRICE CHANGE | minimax/minimax-m2.5 | cache read +20.0% · input +2.3% |
| PRICE CHANGE | minimax/minimax-m3:batch | output +100.0% · cache read +100.0% · input +100.0% |
| PRICE CHANGE | moonshotai/kimi-k2.5 | output -21.1% · cache read -26.3% · input -21.1% |
| PRICE CHANGE | moonshotai/kimi-k2.6 | output -3.4% · cache read -3.4% · input -3.4% |
| PRICE CHANGE | moonshotai/kimi-k2.7-code:batch | output +100.0% · cache read +100.0% · input +100.0% |
| PRICE CHANGE | nvidia/nemotron-3-ultra-550b-a55b:batch | output +100.0% · cache read +100.0% · input +100.0% |
| PRICE CHANGE | qwen/qwen3-next-80b-a3b-instruct | cache read no longer charged · input -10.0% |
| PRICE CHANGE | qwen/qwen3.5-122b-a10b | output -13.3% · input -10.3% |
| PRICE CHANGE | qwen/qwen3.5-35b-a3b | output -30.6% · cache read +11.1% · input +11.1% |
| PRICE CHANGE | qwen/qwen3.6-27b | output -16.7% · cache read now charged · input +3.8% |
| PRICE CHANGE | thinkingmachines/inkling:batch | output +100.0% · cache read +100.0% · input +100.0% |
| PRICE CHANGE | z-ai/glm-5.2 | output +102.9% · cache read +118.5% · input +102.9% |
| PRICE CHANGE | z-ai/glm-5.2:batch | output +100.0% · cache read +100.0% · input +100.0% |
| PRICE CHANGE | ~deepseek/deepseek-v4-flash-latest | output -2.7% · cache read -2.7% · input -2.7% |
| NEW MODEL | z-ai/glm-5.3 | context 1,048,576 |
| NEW MODEL | ~z-ai/glm-latest | context 1,048,576 |
| MODEL REMOVED | ai21/jamba-large-1.7 | |
| FROM THE SOURCE | Sam Altman @sama | (We still expect to ship great new models soon; this impacts further-out releases.) |
| FROM THE SOURCE | Tibo @thsottiaux | It has not been used yet, but would you look at that. Codex for scale. https://t.co/o1pulwoifd |
| FROM THE SOURCE | Tibo @thsottiaux | @ClaudeDevs Do not worry, we have compute |
| FROM THE SOURCE | Anthropic @AnthropicAI | For more on how Claude ran this experiment and the full results, see our blog: https://t.co/48COrpUcJn |
| FROM THE SOURCE | Anthropic @AnthropicAI | One of our highest priorities remains launching an access program for scientists to use our most capable models. We expect to share more on this soon. Opus 5 remains our most capable model available for life science research. |
| FROM THE SOURCE | Anthropic @AnthropicAI | Designing a binder is an easier process than designing a drug, but it’s a useful proxy. The typical success rate in the field today is between 10% and 15%. Between 22% and 35% of Claude's designs bound successfully, depending on the setup. |
| FROM THE SOURCE | Google AI Developers @googleaidevs | From a single prompt to a fully animated parallax landing page. 🤯 See how Gemini 3.7 Flash, combined with Nano Banana and Omni, builds interactive websites in one shot. Under the hood, 3.7 Flash calls the right tools to generate the copy, i |
| OUTAGE | Degraded performance for Claude Opus 5 and Claude Haiku 4.5 | |
| RELEASE | v0.124.0 | |
| RELEASE | v2.1.235 | |
| RELEASE | 0.148.0 | |
| RELEASE | langchain-core==1.6.0 | |
| RELEASE | b10502 | |
| RELEASE | v3.3.1 | |
| RELEASE | Patch release: v5.15.1 | |
| ANNOUNCEMENT | Credential revocation and deauthorization by token type | |
| ANNOUNCEMENT | Enterprise managed settings in GitHub Copilot for JetBrains | |
| ANNOUNCEMENT | Track organization code quality trends | |
| ANNOUNCEMENT | GitHub Copilot app for Beginners: Managing your work | |
| ANNOUNCEMENT | LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation | |
| ANNOUNCEMENT | ChatGPT Ads expands across Europe | |
| ANNOUNCEMENT | How NVIDIA scales expertise with ChatGPT Work | |
| ANNOUNCEMENT | Offering Zero Data Retention for frontier models | |
| ANNOUNCEMENT | Replit expands access to software creation with GPT-5.6 Luna | |
| ANNOUNCEMENT | Strengthening democratic oversight in national security | |
| OUTAGE | Elevated errors deploying Sites | |
| PARAMETERS CHANGED | deepseek/deepseek-chat-v3-0324 | added logprobs, top_logprobs |
| PARAMETERS CHANGED | inclusionai/ling-3.0-flash | removed structured_outputs |
| PARAMETERS CHANGED | minimax/minimax-m2.5 | removed parallel_tool_calls |