2026-09-02

Every AI model price, context and end-of-life change detected on 2026-09-02.

ChangeModelDetail
END OF LIFEnex-agi/nex-n2-minishuts down 2026-09-08 · 6 days left
END OF LIFEnex-agi/nex-n2-proshuts down 2026-09-08 · 6 days left
END OF LIFEdots-studio/dots-3-note-preview:freeshuts down 2026-09-30 · 28 days left
END OF LIFEz-ai/glm-4.5shuts down 2026-12-31 · 120 days left
END OF LIFEz-ai/glm-4.5vshuts down 2026-12-31 · 120 days left
END OF LIFE ANNOUNCEDz-ai/glm-4.5v- -> 2026-12-31
PRICE CHANGEdeepseek/deepseek-chat-v3.1output -42.4% · cache read -76.4% · input -54.5%
PRICE CHANGEdeepseek/deepseek-v4-flashoutput +12.6% · cache read +12.6% · input +12.6%
PRICE CHANGEdeepseek/deepseek-v4-prooutput -35.6% · cache read -36.4% · input -35.6%
PRICE CHANGEdeepseek/deepseek-v4-pro-0813overrides no longer charged
PRICE CHANGEgoogle/gemini-3.7-flash:batchaudio +100.0% · output +100.0% · image +100.0% · input_audio_cache +100.0% · cache read +100.0% · cache write +100.0% · internal_reasoning +100.0% · input +100.0%
PRICE CHANGEnvidia/nemotron-3-ultra-550b-a55boutput +42.0% · cache read +87.5% · input +25.0%
PRICE CHANGEqwen/qwen3.8-2.4t-a95bcache read -20.0%
PRICE CHANGEqwen/qwen3.8-2.4t-a95b:batchoutput -4.0% · cache read -50.0% · input -20.0%
PRICE CHANGEz-ai/glm-5.2output -18.8% · cache read -12.6% · input -18.8%
PRICE CHANGE~deepseek/deepseek-v4-flash-latestoutput +60.1% · cache read +30.1% · input +0.0%
PRICE CHANGE~z-ai/glm-latestoutput -11.6% · cache read -57.3% · input -1.7%
NEW MODELanthropic/claude-fable-5.1:batchcontext 1,000,000
NEW MODELgoogle/gemini-3.8-flashcontext 1,048,576
NEW MODELgoogle/gemini-3.8-flash:batchcontext 1,048,576
NEW MODEL~z-ai/glm-flash-latestcontext 1,310,720
FROM THE SOURCESam Altman @samaOver the summer, we have been sprinting on safety priorities; it's more important than ever for capabilities and safeguards to advance together. We have more to do but have made a lot of progress. We are also going to be launching our next
FROM THE SOURCETibo @thsottiaux@TokenGremlin Given how little we sleep during releases, that would be very welcome
FROM THE SOURCEOpenAI @OpenAIAs we prepare to release Astra, we’re focused on making increasingly capable AI safe and broadly accessible. Astra represents a significant advance in cybersecurity capability, reaching the Critical threshold under our Preparedness Framewor
FROM THE SOURCEGoogle AI Developers @googleaidevsGemini 3.8 Flash is hardwired for complex reasoning. ⚙️ To test its skills, we built an interactive 3D visualizer with 3.8 Flash and @ThreeJS in @GoogleAIStudio. Watch the model generate realistic, physically-proportioned teardowns for hard
OUTAGEDelays in credit purchases
RECOVERYAll Systems Operational
ANNOUNCEMENTIntroducing Gemini 3.8 Flash and 3.8 Flash Cyber
ANNOUNCEMENTProactive cyber defense for governments and enterprises
RELEASEv2.1.258
RELEASE0.152.1
RELEASEv3.7.0
RELEASEv2.22.0
ANNOUNCEMENTGitHub CLI: Media in issues, pull requests, and comments
ANNOUNCEMENTSelected GitHub Copilot models deprecated
ANNOUNCEMENTClaude Fable 5.1 is generally available in GitHub Copilot
ANNOUNCEMENTCopilot code review can now approve pull requests
ANNOUNCEMENTEnterprise Live Migrations from GHES to ghe.com generally available
ANNOUNCEMENTSet an expiration date for individual user budgets
ANNOUNCEMENTHow we make AI coding more cost efficient without sacrificing task quality
ANNOUNCEMENTBenchMIRT: What are LLM benchmarks actually measuring?
ANNOUNCEMENTReal-Time Intelligence with IBM Time Series Models on Confluent
ANNOUNCEMENTHow law firm Gilbert + Tobin governs and scales AI with OpenAI
ANNOUNCEMENTPath to Astra: critical capabilities and frontier safeguards
OUTAGEElevated errors creating new accounts
PARAMETERS CHANGEDxiaomi/mimo-v2.5removed logprobs, top_logprobs

Previous 2026-09-01