Every AI model price, context and end-of-life change detected on 2026-07-30.
| Change | Model | Detail |
|---|---|---|
| END OF LIFE | z-ai/glm-4.5v | shuts down 2026-12-31 · 150 days left |
| PRICE CHANGE | moonshotai/kimi-k2.6 | output +47.1% · cache read +47.1% · input +47.1% |
| PRICE CHANGE | openai/gpt-5.6-luna | output -80.0% · cache read -80.0% · cache write -80.0% · overrides · input -80.0% |
| PRICE CHANGE | openai/gpt-5.6-luna-pro | output -80.0% · cache read -80.0% · cache write -80.0% · overrides · input -80.0% |
| PRICE CHANGE | openai/gpt-5.6-terra | output -20.0% · cache read -20.0% · cache write -20.0% · overrides · input -20.0% |
| PRICE CHANGE | openai/gpt-5.6-terra-pro | output -20.0% · cache read -20.0% · cache write -20.0% · overrides · input -20.0% |
| PRICE CHANGE | z-ai/glm-5.2 | output +57.4% · cache read +57.4% · input +57.4% |
| PRICE CHANGE | ~moonshotai/kimi-latest | output -6.7% |
| PRICE CHANGE | deepseek/deepseek-chat | output +28.6% · input +28.6% |
| PRICE CHANGE | google/gemma-4-31b-it | output -15.0% · cache read now charged · input -28.6% |
| PRICE CHANGE | nvidia/nemotron-3-ultra-550b-a55b | output +63.6% · cache read +100.0% · input +20.0% |
| PRICE CHANGE | qwen/qwen-2.5-7b-instruct | output +100.0% · input +150.0% |
| PRICE CHANGE | qwen/qwen3-vl-30b-a3b-instruct | output +15.4% · input +15.4% |
| PRICE CHANGE | z-ai/glm-5.2 | output -11.4% · cache read -11.4% · input -11.4% |
| PRICE CHANGE | ~moonshotai/kimi-latest | cache read -3.3% · input -3.3% |
| MODEL REMOVED | openai/gpt-5-codex | |
| MODEL REMOVED | openai/o3-deep-research | |
| MODEL REMOVED | openai/o4-mini-deep-research | |
| FROM THE SOURCE | Sam Altman @sama | we want to offer the best price/intelligence tradeoff at every level |
| FROM THE SOURCE | Sam Altman @sama | major price cuts today: *80% drop for GPT-5.6 Luna, now $0.20 per million input tokens and $1.20 per million output *20% drop for GPT-5.6 Terra, to $2/$12 *GPT-5.6 Sol gets Fast mode in the API, up to 2.5x the speed for 2x the price, same i |
| FROM THE SOURCE | OpenAI @OpenAI | Making advanced intelligence more abundant and affordable is central to our mission to ensure AGI benefits all of humanity. With the help of GPT-5.6 Sol, we have made leaps in efficiency. Today, we are passing those gains on in the API with |
| FROM THE SOURCE | OpenAI @OpenAI | We’re also upgrading Auto-review in the ChatGPT app and Codex CLI from GPT-5.4 to GPT-5.6 Luna. Combined with Luna’s new price, we expect Auto-review to cost about 10x less, making your agentic workflows more cost-efficient. |
| FROM THE SOURCE | OpenAI @OpenAI | Along with the price reduction on GPT-5.6 Luna and Terra, Fast mode for GPT-5.6 Sol in the API delivers up to 2.5x the speed of Standard processing at 2x the Standard price. Fast mode gives API customers faster access to GPT-5.6 Sol, with n |
| FROM THE SOURCE | OpenAI @OpenAI | We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a faster option for GPT-5.6 Sol in the API. |
| FROM THE SOURCE | Google AI Developers @googleaidevs | Build smarter, safer physical agents today using the Gemini API or @GoogleAIStudio ↓ https://t.co/30bgUlQlSP |
| FROM THE SOURCE | Google AI Developers @googleaidevs | Gemini Robotics ER 2 is our most capable embodied reasoning model designed for physical AI 🤖 Built as a high-level brain for robotics, the model connects directly to the Gemini Live API. It processes continuous video streams to track progre |
| OUTAGE | Degraded performance on Claude Opus 4.8 | |
| ANNOUNCEMENT | Gemini Robotics 2 brings whole body intelligence to robots | |
| ANNOUNCEMENT | Gemini Robotics ER 2: powering robotics with video understanding, task orchestration, and multi-robot collaboration | |
| RELEASE | langchain-core==1.5.3 | |
| RELEASE | b10197 | |
| RELEASE | v2.51.0 | |
| RELEASE | v2.16.0 | |
| ANNOUNCEMENT | Limit remote control to managed devices | |
| ANNOUNCEMENT | Stacked pull requests are now in public preview | |
| ANNOUNCEMENT | GitHub Copilot in Visual Studio — July update | |
| ANNOUNCEMENT | Reference same-repository actions with self-repository syntax | |
| ANNOUNCEMENT | Stacked sessions and pull requests in the GitHub Copilot app | |
| ANNOUNCEMENT | GPU Management: Why Idle GPUs Are the New Grounded Aircraft | |
| ANNOUNCEMENT | Advancing the price-performance frontier with GPT-5.6 | |
| OUTAGE | Elevated errors affecting ChatGPT conversations | |
| PARAMETERS CHANGED | openai/gpt-5.1-codex-max | removed max_tokens |
| PARAMETERS CHANGED | openai/gpt-5.2-codex | removed max_tokens |
| PARAMETERS CHANGED | thinkingmachines/inkling | added response_format |
| FROM THE SOURCE | Sam Altman @sama | so excited for this. very close to models that will significantly accelerate scientific discovery; the best way to do this is for us to empower scientists, not to try to figure out everything ourselves. we all deserve the benefits. https:// |
| FROM THE SOURCE | OpenAI @OpenAI | We hope these experiments serve as a reminder that evals rarely measure models in isolation—they also measure a bundle of less visible choices about API settings, harness design, and prompting. If you’re an API developer trying to maximize |
| FROM THE SOURCE | OpenAI @OpenAI | A benchmark score reflects the model as well as the harness and settings used to run it. For long-running agents, retaining reasoning and compacting context lets the model build on what it has already learned. https://t.co/JBIKCKrUfw |
| FROM THE SOURCE | OpenAI @OpenAI | We implemented the harness with the Responses API and turned on: → Retained reasoning → Context compaction On the public set, GPT-5.6 Sol’s score rose 188% while using 6x fewer output tokens. https://t.co/uN1IrKEugu |
| FROM THE SOURCE | OpenAI @OpenAI | ARC-AGI-3 tests how well models can learn unfamiliar 2D games without instructions. The standard harness discarded GPT-5.6 Sol’s reasoning after each move and dropped earlier actions as the context filled up. The model had to keep starting |
| FROM THE SOURCE | OpenAI @OpenAI | GPT-5.6 Sol has been used to solve open problems in mathematics. So why was it struggling with ARC-AGI-3, a benchmark of 2D puzzle games? We investigated. The harness was not letting it remember what it had learned. We found that enabling t |
| OUTAGE | Elevated errors across many models | |
| OUTAGE | Elevated errors across all models | |
| RELEASE | b10189 | |
| ANNOUNCEMENT | Copilot code review: Agent skills and MCP now generally available | |
| OUTAGE | Copilot model Claude Fable 5 experiencing elevated errors | |
| OUTAGE | Incident with Copilot AI Model Providers | |
| ANNOUNCEMENT | How GPT-5.6 fuses frontier intelligence with frontier efficiency | |
| ANNOUNCEMENT | How enabling two settings tripled our scores on the ARC-AGI-3 benchmark |