New model: Qwen3.8 Flash (Alibaba (Qwen))
Alibaba (Qwen) released Qwen3.8 Flash, a multimodal model supporting text, image, and video inputs with a 1000k context window and 131k maximum output, on 2026-08-26. The model is available under the …
Alibaba (Qwen) released Qwen3.8 Flash, a multimodal model supporting text, image, and video inputs with a 1000k context window and 131k maximum output, on 2026-08-26. The model is available under the …
Z.ai (Zhipu) has deprecated its GLM-5.3-Flash model, with retirement scheduled for December 31, 2098. The multimodal model, released August 26, 2026, supports text, image, video, and PDF inputs with a…
Mistral AI released GLM-5.2 on June 13, 2026, a new model with a 1000k context window and 131k max output tokens, available under the identifiers zai-glm-5-2 and mistral/zai-glm-5-2. The model is list…
Z.ai (Zhipu) has deprecated its GLM-5.3 large language model, with retirement scheduled for December 31, 2098, according to a provider announcement on models.dev. The model, released on August 14, 202…
Moonshot AI has deprecated its Kimi K2.5 large language model, with retirement scheduled for August 31, 2026, and is directing users to the newer Kimi K2.6. The model, which supports a 256k context wi…
Google released Claude Fable 5 on June 9, 2026, a multimodal model supporting text, image, and PDF inputs with a 1,000k context window and 128k max output. The model is available via Google Vertex AI …
DeepSeek released DeepSeek V4 Flash Vision Exp on 2026-08-21, a text+image model with a 1000k context length and 384k max output, available under the identifiers deepseek-v4-flash-vision-exp and deeps…
ModelStatus, a new open-source CLI tool, scans any repository to identify every AI model the code calls and flags deprecated or retiring models with their retirement dates and replacements. The tool, …
XAI released Grok 4.6 on Amazon Bedrock on 2026-08-12, a text and image model with a 500k context window and 500k max output tokens. The model is available under the identifiers xai.grok-4.6 and amazo…
Microsoft has retired Phi-3.5-MoE-instruct on Azure AI Foundry, with the retirement schedule now listing only Phi-4 and no longer publishing an exact date. The model was delisted from models.dev's Azu…
Microsoft has retired Phi-3-medium-instruct (128k) on Azure AI Foundry, with the retirement schedule now listing only Phi-4 and no longer publishing an exact date. The model was also delisted from mod…
Microsoft retired Meta-Llama-3-70B-Instruct on Azure AI Foundry, delisting the model from its azure-hosted deployment catalog as of August 2026. The model, released on April 18, 2024, had an 8k contex…
Microsoft retired GPT-5.2 Chat on Azure AI Foundry, with both versions retired as of 2026-06-29 and replaced by gpt-chat-latest, correcting a previously scheduled retirement date of 2026-08-10. The mo…
Microsoft has retired Phi-3-medium-instruct (4k), its 4k-context language model, from Azure AI Foundry hosted deployments as part of the Phi-3/3.5 MaaS retirement in 2025, and the model was delisted f…
Microsoft retired Meta-Llama-3-8B-Instruct on Azure AI Foundry, removing the serverless deployment and delisting the model from models.dev's Azure catalog in August 2026. The model, released on 2024-0…
Microsoft retired the GPT-5.1 Chat model on Azure AI Foundry on 2026-06-29, directing users to the gpt-chat-latest model. The model was delisted from models.dev's Azure catalog in August 2026. The ret…
Microsoft has retired Phi-3-small-instruct (128k), its azure-hosted deployment on Azure AI Foundry, as part of the Phi-3/3.5 MaaS retirement in 2025, with the model delisted from models.dev's Azure ca…
Microsoft retired Meta-Llama-3.1-70B-Instruct on Azure AI Foundry on 2026-06-13, part of a serverless retirement wave that also affected Meta-Llama-3.1-405B and 8B-Instruct. The model, released 2024-0…
Microsoft retired all versioned gpt-5-chat deployments on Azure AI Foundry as of 2026-06-29, consolidating them into the rolling alias gpt-chat-latest, and the model was delisted from models.dev's Azu…
Microsoft has retired Phi-3-small-instruct (8k), its 8k-context small language model, from Azure AI Foundry hosted deployments, with the retirement schedule now listing only Phi-4 and no longer publis…