Deprecated since February and available to existing customers only, Claude 3 Haiku clears off Google Cloud on August 23 — the last Claude 3-generation model on the platform. The absurd part: Anthropic itself retired claude-3-haiku back on April 20, 2026. Google kept serving a model its maker had already killed for another four months.
Persistent AI agents with memory, tools, and file handling. Replaced by the Responses API. Developers face a full rewrite.
The official preconfigured DALL·E GPT — the last place the DALL·E brand still lived inside ChatGPT — retires August 30. OpenAI won't say what happens to images stored in it, and instead just recommends downloading anything you want to keep before the deadline. Image generation itself stays, under the ChatGPT Images banner.
Released April 14, 2026. Scheduled for deletion August 31, 2026. A 139-day lifespan, and it never left preview. It killed its own predecessor, ER 1.5, on April 30 — meaning Google will have retired two embodied-reasoning robotics models inside five months, neither of which ever reached GA.
Both retire the same day. Medium 3.1 shipped in August 2025 and dies almost exactly a year later.
While the Sora app died in April, the API endpoint survives until September. Then it's gone completely.
The end of the GPT-3 base-model lineage. These three shipped in August 2023, inherited the job when the original ada/babbage/curie/davinci died in January 2024, and are the last models served over the legacy /v1/completions endpoint. Their own fine-tuned versions outlive them by 25 days.
Mistral announced this Lean 4 formal-proof model's death in the same sentence that announced its birth. The changelog reads: 'We released Leanstral 1.5... This model will be retired on September 30, 2026.' A 92-day advertised lifespan, dead on arrival by design.
The image model that went viral as 'Nano Banana' and briefly made Gemini the #1 free iPhone app shuts down exactly one year after release, to the day. Google's officially recommended replacement, gemini-3.1-flash-image-preview, was itself shut down in June 2026. The migration target is dead; the model isn't yet.
GPT-4.1 launched April 2025 as OpenAI's 'best model for coding' with a 1M-token context window. Azure retires it, plus mini and nano, after 18 months. Fine-tuned deployments on the same bases get a one-year stay of execution.
Google's entire Gemini 2.5 family retires on Vertex AI no earlier than October 16, 2026. Note this is Vertex-only — the Gemini Developer API still lists no shutdown date for the same three models.
Azure deletes an entire generation of reasoning models in under 90 days: o3-mini October 1, o4-mini October 16, o1/o1-pro/o3 October 21, codex-mini November 15, o3-pro December 17. Four have no named replacement — Microsoft declares one only 90 to 120 days out, because 'declaring a replacement too early risks directing customers to a model that is no longer the best available option.'
New organizations can no longer create fine-tuning jobs. The first wave executed on schedule July 23, 2026; the second takes every remaining fine-tuned model — ft-gpt-3.5-turbo, ft-gpt-4, ft-gpt-4.1-nano, ft-babbage-002 and ft-davinci-002 — on October 23.
One date kills three eras. On October 23, OpenAI removes gpt-3.5-turbo (the engine that launched ChatGPT), gpt-4 and gpt-4-0613 (the model that defined 2023), and its entire first generation of reasoning models — o1, o1-pro, o3-mini, o4-mini — plus gpt-4o-2024-05-13 and gpt-image-1. Every fine-tune built on them dies the same day. Seventeen table rows, six months' notice.
Older GPT Image model snapshots were flagged for removal from the API on December 1, 2026, just months after they replaced DALL·E. The endpoint that buried DALL·E is already on its own deprecation clock.
The original GPT-5 launch snapshots — gpt-5-2025-08-07 and its mini, nano and pro siblings — plus o3 and o3-pro shut down December 11. GPT-5 launched in August 2025 as the 'PhD-level expert' model. Its launch snapshot gets 16 months before it starts returning errors.
OpenAI put its entire GA voice stack on death row: gpt-realtime, gpt-audio, gpt-realtime-mini, gpt-audio-mini and all four gpt-4o audio/realtime variants. gpt-realtime and gpt-audio went GA in August 2025 — the models that headlined 'production-ready voice agents' get roughly 17 months. OpenAI is skipping its own intermediate generation: the 1.5 series shipped in February 2026 and the named replacement is already 2.1.