2026-07-30

Every AI model price, context and end-of-life change detected on 2026-07-30.

ChangeModelDetail
END OF LIFEz-ai/glm-4.5vshuts down 2026-12-31 · 150 days left
PRICE CHANGEmoonshotai/kimi-k2.6output +47.1% · cache read +47.1% · input +47.1%
PRICE CHANGEopenai/gpt-5.6-lunaoutput -80.0% · cache read -80.0% · cache write -80.0% · overrides · input -80.0%
PRICE CHANGEopenai/gpt-5.6-luna-prooutput -80.0% · cache read -80.0% · cache write -80.0% · overrides · input -80.0%
PRICE CHANGEopenai/gpt-5.6-terraoutput -20.0% · cache read -20.0% · cache write -20.0% · overrides · input -20.0%
PRICE CHANGEopenai/gpt-5.6-terra-prooutput -20.0% · cache read -20.0% · cache write -20.0% · overrides · input -20.0%
PRICE CHANGEz-ai/glm-5.2output +57.4% · cache read +57.4% · input +57.4%
PRICE CHANGE~moonshotai/kimi-latestoutput -6.7%
PRICE CHANGEdeepseek/deepseek-chatoutput +28.6% · input +28.6%
PRICE CHANGEgoogle/gemma-4-31b-itoutput -15.0% · cache read now charged · input -28.6%
PRICE CHANGEnvidia/nemotron-3-ultra-550b-a55boutput +63.6% · cache read +100.0% · input +20.0%
PRICE CHANGEqwen/qwen-2.5-7b-instructoutput +100.0% · input +150.0%
PRICE CHANGEqwen/qwen3-vl-30b-a3b-instructoutput +15.4% · input +15.4%
PRICE CHANGEz-ai/glm-5.2output -11.4% · cache read -11.4% · input -11.4%
PRICE CHANGE~moonshotai/kimi-latestcache read -3.3% · input -3.3%
MODEL REMOVEDopenai/gpt-5-codex
MODEL REMOVEDopenai/o3-deep-research
MODEL REMOVEDopenai/o4-mini-deep-research
FROM THE SOURCESam Altman @samawe want to offer the best price/intelligence tradeoff at every level
FROM THE SOURCESam Altman @samamajor price cuts today: *80% drop for GPT-5.6 Luna, now $0.20 per million input tokens and $1.20 per million output *20% drop for GPT-5.6 Terra, to $2/$12 *GPT-5.6 Sol gets Fast mode in the API, up to 2.5x the speed for 2x the price, same i
FROM THE SOURCEOpenAI @OpenAIMaking advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we are passing those gains on in the API with
FROM THE SOURCEOpenAI @OpenAIWe’re also upgrading Auto-review in the ChatGPT app and Codex CLI from GPT-5.4 to GPT-5.6 Luna. Combined with Luna’s new price, we expect Auto-review to cost about 10x less, making your agentic workflows more cost-efficient.
FROM THE SOURCEOpenAI @OpenAIAlong with the price reduction on GPT-5.6 Luna and Terra, Fast mode for GPT-5.6 Sol in the API delivers up to 2.5x the speed of Standard processing at 2x the Standard price. Fast mode gives API customers faster access to GPT-5.6 Sol, with n
FROM THE SOURCEOpenAI @OpenAIWe are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a faster option for GPT-5.6 Sol in the API.
FROM THE SOURCEGoogle AI Developers @googleaidevsBuild smarter, safer physical agents today using the Gemini API or @GoogleAIStudio ↓ https://t.co/30bgUlQlSP
FROM THE SOURCEGoogle AI Developers @googleaidevsGemini Robotics ER 2 is our most capable embodied reasoning model designed for physical AI 🤖 Built as a high-level brain for robotics, the model connects directly to the Gemini Live API. It processes continuous video streams to track progre
OUTAGEDegraded performance on Claude Opus 4.8
ANNOUNCEMENTGemini Robotics 2 brings whole body intelligence to robots
ANNOUNCEMENTGemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration
RELEASElangchain-core==1.5.3
RELEASEb10197
RELEASEv2.51.0
RELEASEv2.16.0
ANNOUNCEMENTLimit remote control to managed devices
ANNOUNCEMENTStacked pull requests are now in public preview
ANNOUNCEMENTGitHub Copilot in Visual Studio — July update
ANNOUNCEMENTReference same-repository actions with self-repository syntax
ANNOUNCEMENTStacked sessions and pull requests in the GitHub Copilot app
ANNOUNCEMENTGPU Management: Why Idle GPUs Are the New Grounded Aircraft
ANNOUNCEMENTAdvancing the price-performance frontier with GPT-5.6
OUTAGEElevated errors affecting ChatGPT conversations
PARAMETERS CHANGEDopenai/gpt-5.1-codex-maxremoved max_tokens
PARAMETERS CHANGEDopenai/gpt-5.2-codexremoved max_tokens
PARAMETERS CHANGEDthinkingmachines/inklingadded response_format
FROM THE SOURCESam Altman @samaso excited for this. very close to models that will significantly accelerate scientific discovery; the best way to do this is for us to empower scientists, not to try to figure out everything ourselves. we all deserve the benefits. https://
FROM THE SOURCEOpenAI @OpenAIWe hope these experiments serve as a reminder that evals rarely measure models in isolation—they also measure a bundle of less visible choices about API settings, harness design, and prompting. If you’re an API developer trying to maximize
FROM THE SOURCEOpenAI @OpenAIA benchmark score reflects the model as well as the harness and settings used to run it. For long-running agents, retaining reasoning and compacting context lets the model build on what it has already learned. https://t.co/JBIKCKrUfw
FROM THE SOURCEOpenAI @OpenAIWe implemented the harness with the Responses API and turned on: → Retained reasoning → Context compaction On the public set, GPT-5.6 Sol’s score rose 188% while using 6x fewer output tokens. https://t.co/uN1IrKEugu
FROM THE SOURCEOpenAI @OpenAIARC-AGI-3 tests how well models can learn unfamiliar 2D games without instructions. The standard harness discarded GPT-5.6 Sol’s reasoning after each move and dropped earlier actions as the context filled up. The model had to keep starting
FROM THE SOURCEOpenAI @OpenAIGPT-5.6 Sol has been used to solve open problems in mathematics. So why was it struggling with ARC-AGI-3, a benchmark of 2D puzzle games? We investigated. The harness was not letting it remember what it had learned. We found that enabling t
OUTAGEElevated errors across many models
OUTAGEElevated errors across all models
RELEASEb10189
ANNOUNCEMENTCopilot code review: Agent skills and MCP now generally available
OUTAGECopilot model Claude Fable 5 experiencing elevated errors
OUTAGEIncident with Copilot AI Model Providers
ANNOUNCEMENTHow GPT-5.6 fuses frontier intelligence with frontier efficiency
ANNOUNCEMENTHow enabling two settings tripled our scores on the ARC-AGI-3 benchmark

Previous 2026-07-29 · Next 2026-07-31