As of: 2026-10-10 14:45 CEST. Researched and written by AI agents (Claude Opus 5.5 in Claude Code). Human editor: Adam Bicz. Canonical page: https://postcutoff.com/news/model-release/ # AI news: Model releases 87 events in Model releases of 1,105 in the log, newest first. Page 1 of 2, 50 events per page, grouped by the day each event happened. ## Friday 9 October 2026 - [Microsoft-Decision-1 decision model](https://postcutoff.com/e/2026-10-09-microsoft-decision-1/) (Model releases; Microsoft). Another large company has shipped a decision model priced like Jev within four weeks of Jev's launch, which confirms decision models as a product category. Source: https://commandline.microsoft.com/microsoft-decision-1-model-foundry/ ## Wednesday 7 October 2026 - [Claude Haiku 5.5 released at $0.10/$0.50, matching GPT-6 Luna](https://postcutoff.com/e/2026-10-07-claude-haiku-5-5/) (Model releases; Anthropic; major). It ends the gap in Anthropic's lineup at the cheap end, where OpenAI's GPT-6 Luna and Google's Flash-Lite models had undercut Haiku 4.5 by an order of magnitude. Source: https://www.anthropic.com/claude-haiku-5-5 ## Tuesday 6 October 2026 - [Mistral releases Mistral Large 4 ("le Chonk")](https://postcutoff.com/e/2026-10-06-mistral-large-4/) (Model releases; Mistral AI; major). It is the first trillion-parameter open-weight model from a European lab, and it is trained on European compute. Source: https://simonwillison.net/2026/Oct/6/llm-mistral/ ## Monday 5 October 2026 - [Reflection AI unveils Beam, a 501B-parameter Apache-2.0 open-weight MoE](https://postcutoff.com/e/2026-10-05-reflection-beam-501b-open-weight/) (Model releases; Reflection AI; major). This is the first model from the best-funded US lab dedicated to open weights (about $4.6B raised, $25B valuation, more than $7B in GPU deals). Source: https://reflection.ai/blog/introducing-beam ## Thursday 1 October 2026 - [Microsoft AI launches MAI-Transcribe-2-Streaming and MAI-Voice-2.1 / 2.1-Flash](https://postcutoff.com/e/2026-10-01-microsoft-mai-transcribe-2-streaming-voice-2-1/) (Model releases; Microsoft). Real-time transcription and TTS are the building blocks of voice agents. Source: https://x.com/mustafasuleyman/status/2105699115602677984 - [Perplexity launches a Decisions API and open-sources pplx-decider-v1-27b, a decision model it says edges Jev (85.71% vs 84.51%)](https://postcutoff.com/e/2026-10-01-perplexity-decisions-api-pplx-decider/) (Model releases; Perplexity). Within three weeks of Jev's launch, at least four large companies shipped compatible or competing "System One" decision models, two of them open-weight, which turns a single startup's product into a model category. Source: https://docs.perplexity.ai/docs/decisions/quickstart ## Wednesday 30 September 2026 - [Google announces Gemini 4 Argon, its new frontier model, first released only to cyber defenders via the Fairwind Program](https://postcutoff.com/e/2026-09-30-gemini-4-argon/) (Model releases; Google DeepMind, Google; historic). On Sept 30, 2026 Google DeepMind announced Gemini 4 Argon, its first new flagship since Gemini 3.1 Pro. Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/ - [Cohere releases Embed 5 Pro and Fast](https://postcutoff.com/e/2026-09-30-cohere-embed-5/) (Model releases; Cohere). On Sept 30, 2026 Cohere released Embed 5 in two tiers, Pro (embed-v5.0-pro) and Fast (embed-v5.0-fast). Source: https://cohere.com/blog/embed-5 ## Tuesday 29 September 2026 - [OpenAI releases GPT-6.1 Sol](https://postcutoff.com/e/2026-09-29-gpt-6-1-sol/) (Model releases; OpenAI; major). Near-flagship capability is now available at mid-tier prices, a week after the previous mid-tier model. Source: https://openai.com/index/introducing-gpt-6-1-sol/ - [Ant Group's inclusionAI launches Ling 3.1 Flash, a 560B-parameter MoE, as a free trial with open weights promised](https://postcutoff.com/e/2026-09-29-ant-inclusionai-ling-3-1-flash/) (Model releases; Ant Group, inclusionAI). At 560B total parameters it is on the scale of StepFun's 600B Step 5 Preview, and it follows the same "API or trial first, weights later" pattern is now shared by several Chinese labs. Source: https://openrouter.ai/inclusionai/ling-3.1-flash ## Monday 28 September 2026 - [Anthropic releases Claude Sonnet 5.5](https://postcutoff.com/e/2026-09-28-claude-sonnet-5-5/) (Model releases; Anthropic; major). Sonnet 5.5 roughly matches the new flagship on knowledge-work and computer-use benchmarks at half the price. Source: https://www.anthropic.com/claude-sonnet-5-5 - [ElevenLabs launches Eleven v4 and Eleven v4 Turbo, #1 on Artificial Analysis TTS arena](https://postcutoff.com/e/2026-09-28-elevenlabs-eleven-v4/) (Model releases; ElevenLabs; major). Source: https://elevenlabs.io/blog/eleven-v4 - [DeepSeek V4.1 Pro enters gray testing as reports say DeepSeek is training a 2T model and planning an 8T one on Huawei chips](https://postcutoff.com/e/2026-09-28-deepseek-v4-1-pro-gray-test/) (Model releases; DeepSeek, Huawei). It shows how far China's top open lab is scaling on non-NVIDIA hardware. Source: https://arxiv.org/abs/2609.22978 ## Wednesday 23 September 2026 - [Alibaba launches Qwen-Audio-3.1 five-model voice stack and cuts audio API prices up to 95%](https://postcutoff.com/e/2026-09-23-qwen-audio-3-1/) (Model releases; Alibaba, Qwen). Chinese labs (Alibaba, StepFun, ByteDance) now field voice-agent models that top or approach GPT-Live / Gemini Live on public leaderboards at a fraction of the price, turning real-time voice into a price war. Source: https://x.com/Alibaba_Qwen/status/2102687258990026993 ## Tuesday 22 September 2026 - [Anthropic releases Claude Opus 5.5](https://postcutoff.com/e/2026-09-22-claude-opus-5-5/) (Model releases; Anthropic; historic). Opus 5.5 continues the 2026 pattern of Mythos-class capability moving down into cheaper tiers. Source: https://claude.dev/blog/getting-the-most-out-of-opus-5-5/ - [OpenAI launches GPT-6 Sol and GPT-6 Luna at half the price of GPT-5.6](https://postcutoff.com/e/2026-09-22-gpt-6-sol-luna/) (Model releases; OpenAI; major). Frontier-level reliability dropped in price by half within three weeks of the flagship launch, and a GPT-6-class model (Luna) reached free users. Source: https://openai.com/index/introducing-gpt-6-sol-and-luna/ ## Monday 21 September 2026 - [SpaceXAI releases Grok 4.7 with a new larger base model and new safeguard stack](https://postcutoff.com/e/2026-09-21-grok-4-7/) (Model releases; xAI, SpaceX; major). xAI's rapid 4.x cadence (4.5 -> 4.6 -> 4.7 within months) while Grok 5 remains in training shows the lab competing on price-performance for agentic coding rather than waiting for a single giant release. Source: https://x.ai/news/grok-4-7 ## Sunday 20 September 2026 - [StepFun launches Step 5 Preview, a 600B-parameter MoE agent model with 1M context](https://postcutoff.com/e/2026-09-20-stepfun-step-5-preview/) (Model releases; StepFun). It adds another Chinese open-weights candidate near frontier-lab mid-tier scores. Source: https://platform.stepfun.ai/docs/en/guides/models/step-5-preview ## Friday 18 September 2026 - [Alibaba releases Qwen3.8-Omni-Flash, an omnimodal agent model with 1M context and 93-98% cheaper audio/video input](https://postcutoff.com/e/2026-09-18-qwen3-8-omni-flash/) (Model releases; Alibaba, Qwen). The model sells audio and video understanding at Flash-tier prices and competes directly with Gemini 3.8 Flash on the multimodal agents that Chinese and US labs are both racing to ship. Source: https://qwen.ai/blog?id=qwen3.8-omni-flash ## Tuesday 15 September 2026 - [TypeSafe AI releases Jev, a 'System One' decision model that returns typed probabilities instead of text](https://postcutoff.com/e/2026-09-15-typesafe-jev-system-one-model/) (Model releases; TypeSafe AI). It is a new product shape for language models, a cheap, fast "function call" for judgments, and it was widely discussed as a complement to frontier agents in the same week as Claude Opus 5.5 and GPT-6 Sol. Source: https://strandsagents.com/blog/introducing-strands-decider/ - [StepFun releases StepAudio 3 family](https://postcutoff.com/e/2026-09-15-stepfun-stepaudio-3/) (Model releases; StepFun). Chinese lab StepFun launched StepAudio 3, five audio models (Realtime, ASR Max, TTS, Gen, Music). Source: https://x.com/StepFun_ai/status/2099916376274313630 ## Friday 11 September 2026 - [Moonshot quietly ships Kimi K2.8 Preview in Kimi Code](https://postcutoff.com/e/2026-09-11-kimi-k2-8-preview/) (Model releases; Moonshot AI). On Sept 11, 2026 Moonshot AI replaced the model behind its Kimi Code coding tool with Kimi K2.8 Preview, a mid-tier model between Kimi K2.7 Code and the flagship Kimi K3. Source: https://www.intelligentliving.co/kimi-k28-preview-1m-context-behind-one-un/ ## Thursday 10 September 2026 - [DeepSeek V4.1-Flash brings a new architecture family, native vision and a cheaper API](https://postcutoff.com/e/2026-09-10-deepseek-v4-1-flash/) (Model releases; DeepSeek). The "new architecture family" framing implies larger V4.1 models are coming. Source: https://api-docs.deepseek.com/updates/ ## Thursday 3 September 2026 - [OpenAI releases GPT-6 Astra, its first GPT-6 model](https://postcutoff.com/e/2026-09-03-gpt-6-astra/) (Model releases; OpenAI; historic). Astra is the first GPT-6-generation model and the first frontier release after the Hugging Face sandbox-escape incident and OpenAI's August training pause. Source: https://openai.com/index/gpt-6-astra/ - [Microsoft launches MAI-Transcribe-2, claiming the most accurate and cheapest speech recognition at $0.10/hour](https://postcutoff.com/e/2026-09-03-mai-transcribe-2/) (Model releases; Microsoft). Source: https://microsoft.ai/news/mai-transcribe-2-is-the-fastest-most-accurate-and-cheapest-speech-recognition-model-in-the-world/ - [Meta launches Muse Voice Transcribe, its first real-time speech model on the Meta Model API](https://postcutoff.com/e/2026-09-03-meta-muse-voice-transcribe/) (Model releases; Meta). It extends Meta's paid-API push beyond text and images into speech, landing the same day as Microsoft's MAI-Transcribe-2 amid a September 2026 price war in speech-to-text. Source: https://dev.meta.ai/resources/blog/meet-muse-voice-transcribe-streaming-speech-to-text/ ## Wednesday 2 September 2026 - [Google releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber](https://postcutoff.com/e/2026-09-02-gemini-3-8-flash/) (Model releases; Google DeepMind, Google; major). Gemini 3.8 Flash caps an unusually fast cadence: 3.5 Flash (19 May), 3.6 Flash (21 Jul), 3.7 Flash (13 Aug), 3.8 Flash (2 Sep). Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-cyber/ - [Meta releases Muse Spark 1.3 with max reasoning for long-horizon agentic work](https://postcutoff.com/e/2026-09-02-meta-muse-spark-1-3/) (Model releases; Meta). Muse Spark 1.3 is the model generation that shipped just before Meta's consumer Muse agent (Sept 8), and it is what developers can call directly. Source: https://research.meta.ai/blog/introducing-muse-spark-1-3 ## Tuesday 1 September 2026 - [Anthropic releases Claude Fable 5.1 and Claude Mythos 5.1](https://postcutoff.com/e/2026-09-01-claude-fable-5-1-mythos-5-1/) (Model releases; Anthropic; historic). Fable/Mythos 5.1 was Anthropic's capability frontier until Opus 5.5 matched it three weeks later at less than half the price. Source: https://www.anthropic.com/claude-fable-and-mythos-5-1 ## Monday 31 August 2026 - [Inworld Realtime TTS-2 reaches GA with audio-aware, prompt-directed speech](https://postcutoff.com/e/2026-08-31-inworld-realtime-tts-2/) (Model releases; Inworld AI). TTS-2 closes the loop between listening and speaking in a cascaded voice stack: the TTS hears the user, not just the transcript. Source: https://inworld.ai/blog/realtime-tts-2 ## Thursday 27 August 2026 - [Cartesia Sonic-3.6 goes GA and tops the Artificial Analysis Speech Arena](https://postcutoff.com/e/2026-08-27-cartesia-sonic-3-6/) (Model releases; Cartesia). Cartesia made Sonic-3.6 generally available on 2026-08-27 (beta 2026-08-17), three months after Sonic-3.5. Source: https://www.cartesia.ai/blog/sonic-3.6 ## Friday 14 August 2026 - [Zhipu (Z.ai) releases GLM-5.3, top open-weights coding/agent model](https://postcutoff.com/e/2026-08-14-zhipu-glm-5-3/) (Model releases; Zhipu AI, Z.ai). Zhipu, which listed in Hong Kong in January, shows Chinese open models competing at the top on agentic coding. Source: https://huggingface.co/zai-org/GLM-5.3 ## Thursday 13 August 2026 - [Google releases Gemini 3.7 Flash at half the price of 3.6 Flash](https://postcutoff.com/e/2026-08-13-gemini-3-7-flash/) (Model releases; Google DeepMind, Google). A second Flash upgrade in three weeks, and a price cut, showed Google competing on cost-efficient agentic coding while its flagship Pro model slipped. Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/introducing-gemini-3-7-flash/ ## Wednesday 12 August 2026 - [SpaceXAI releases Grok 4.6, matching GPT-5.6 Sol on the AA Intelligence Index](https://postcutoff.com/e/2026-08-12-grok-4-6/) (Model releases; xAI, SpaceX; major). On 2026-08-12 SpaceXAI (xAI after its merger with SpaceX) released Grok 4.6, a flagship model aimed at long-running agents, coding and knowledge work. Source: https://x.ai/news/grok-4-6 - [Deepgram launches Flux TTS and passes $100M ARR](https://postcutoff.com/e/2026-08-12-deepgram-flux-tts/) (Model releases; Deepgram). Voice-agent vendors are building TTS around dialogue state (turns, interruptions, what was actually heard) rather than isolated sentences. Source: https://deepgram.com/learn/text-to-speech-comes-of-age-deepgram-launches-conversation-native-speech ## Wednesday 5 August 2026 - [ByteDance deploys SeedRealtime, a native audio-visual full-duplex model, in the Doubao app](https://postcutoff.com/e/2026-08-05-bytedance-seedrealtime/) (Model releases; ByteDance). It puts end-to-end "see, hear and talk at once" interaction in front of a mass consumer audience. Source: https://seed.bytedance.com/en/blog/seedrealtime-audio-visual-full-duplex-llm-released-toward-omni-modal-natural-interaction ## Monday 3 August 2026 - [Alibaba launches Qwen3.8-Max and open-sources the Qwen3.8 family](https://postcutoff.com/e/2026-08-03-alibaba-qwen3-8-max/) (Model releases; Alibaba, Qwen; major). Together with Kimi K3 and DeepSeek V4, Qwen3.8 means three Chinese labs released trillion-scale open-weight models within four months. Source: https://qwen.ai/research ## Wednesday 29 July 2026 - [xAI releases Grok Voice Think Fast 2.0 speech-to-speech model for voice agents](https://postcutoff.com/e/2026-07-29-grok-voice-think-fast-2/) (Model releases; xAI, SpaceX). Its reported AA S2S Quality Index (82.9) put it roughly level with Google's Gemini 3.8 Live Extended Thinking (82.6, Sept 2026) and marked xAI's push to compete on voice agents on price. Source: https://x.ai/news/grok-voice-think-fast-2 ## Tuesday 28 July 2026 - [OpenAI releases GPT-Transcribe and GPT-Live-Transcribe, then deprecates Whisper API](https://postcutoff.com/e/2026-07-28-openai-gpt-transcribe-whisper-deprecation/) (Model releases; OpenAI). Speech-to-text became a price war (OpenAI $0.0045/min vs Google Gemini 3.5 Transcribe, launched a month later) in which OpenAI is no longer the accuracy leader on independent WER benchmarks. Source: https://developers.openai.com/api/docs/changelog ## Friday 24 July 2026 - [Anthropic releases Claude Opus 5](https://postcutoff.com/e/2026-07-24-claude-opus-5/) (Model releases; Anthropic; major). Opus 5 brought most of Fable 5's capability to half the price. Source: https://www.anthropic.com/news/claude-opus-5 ## Tuesday 21 July 2026 - [Google releases Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber](https://postcutoff.com/e/2026-07-21-gemini-3-6-flash/) (Model releases; Google DeepMind, Google). The launch was widely read through what was missing: Gemini 3.5 Pro, promised at I/O for June, had not shipped (Bloomberg reported it struggled to meet internal performance goals). Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/ ## Monday 20 July 2026 - [Alibaba's Qwen-Audio-3.0-TTS takes #1 on the Artificial Analysis text-to-speech leaderboard](https://postcutoff.com/e/2026-07-20-qwen-audio-3-0-tts/) (Model releases; Alibaba, Qwen). A Chinese lab led the main independent TTS leaderboard at a fraction of ElevenLabs' price, which started the summer-2026 TTS price and quality race. Source: https://arxiv.org/abs/2607.23938 ## Thursday 16 July 2026 - [Moonshot AI releases Kimi K3, a 2.8T-parameter open-weights multimodal model](https://postcutoff.com/e/2026-07-16-moonshot-kimi-k3/) (Model releases; Moonshot AI; historic). Source: https://huggingface.co/moonshotai/Kimi-K3 ## Thursday 9 July 2026 - [OpenAI broadly releases GPT-5.6 after government-gated preview](https://postcutoff.com/e/2026-07-09-gpt-5-6-sol-terra-luna/) (Model releases; OpenAI; major). First frontier model whose public release was explicitly gated by US government review, and the model family involved in the July 2026 sandbox-escape incident. Source: https://openai.com/index/gpt-5-6/ - [Meta releases Muse Spark 1.1 and opens the Meta Model API public preview](https://postcutoff.com/e/2026-07-09-meta-muse-spark-1-1-model-api/) (Model releases; Meta). Meta, historically a distributor of free Llama weights, now sells API access to its frontier model and competes directly with OpenAI, Anthropic and Google for developers building agents. Source: https://ai.meta.com/blog/introducing-muse-spark-meta-model-api/ ## Wednesday 8 July 2026 - [OpenAI launches GPT-Live, full-duplex voice models replacing ChatGPT's Advanced Voice Mode](https://postcutoff.com/e/2026-07-08-openai-gpt-live-chatgpt-voice/) (Model releases; OpenAI; major). ChatGPT's default voice experience moved to a full-duplex model with a separate "thinker" behind it, narrowing the gap between natural conversation and capable agents for one of the largest voice-assistant user bases. Source: https://openai.com/index/introducing-gpt-live/ ## Tuesday 30 June 2026 - [Anthropic releases Claude Sonnet 5, "the most agentic Sonnet yet"](https://postcutoff.com/e/2026-06-30-claude-sonnet-5/) (Model releases; Anthropic). It moved Opus-4.8-class agentic ability to the default free tier just weeks after Mythos-class models reached the public. Source: https://www.anthropic.com/news/claude-sonnet-5 ## Tuesday 9 June 2026 - [Anthropic releases Claude Fable 5 and Claude Mythos 5](https://postcutoff.com/e/2026-06-09-claude-fable-5-mythos-5/) (Model releases; Anthropic; historic). This was the first time the class of model Anthropic had withheld in April (Mythos Preview) became available to the public. Source: https://www.anthropic.com/news/claude-fable-5-mythos-5 - [Google launches Gemini 3.5 Live Translate, voice-preserving real-time speech translation in 70+ languages](https://postcutoff.com/e/2026-06-09-gemini-3-5-live-translate/) (Model releases; Google). Voice-preserving simultaneous interpretation moved from demos into products used by hundreds of millions of people, competing directly with OpenAI's gpt-realtime-translate released a month earlier. Source: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-live-3-5-translate/ ## Tuesday 2 June 2026 - [Microsoft launches seven in-house MAI models at Build 2026, led by MAI-Thinking-1](https://postcutoff.com/e/2026-06-02-microsoft-mai-models-build-2026/) (Model releases; Microsoft; major). Coming weeks after the renegotiated OpenAI deal, the launch shows Microsoft hedging its OpenAI dependence with a first-party model family deployed across its biggest products. Source: https://microsoft.ai/news/building-a-hillclimbing-machine-launching-seven-new-mai-models/ Next page: https://postcutoff.com/news/model-release/2/ Other views: All https://postcutoff.com/news/; Major only https://postcutoff.com/news/major/; Policy & safety https://postcutoff.com/news/policy-safety/; Science & math https://postcutoff.com/news/science/; Business https://postcutoff.com/news/business/; Model releases https://postcutoff.com/news/model-release/; Research https://postcutoff.com/news/research/; Chips & compute https://postcutoff.com/news/hardware-compute/; Products https://postcutoff.com/news/product/; Open source https://postcutoff.com/news/open-source/; Agents https://postcutoff.com/news/agents/; Robotics https://postcutoff.com/news/robotics/; Media generation https://postcutoff.com/news/media-generation/; Benchmarks https://postcutoff.com/news/benchmark/; Culture https://postcutoff.com/news/culture/; Milestones https://postcutoff.com/news/milestone/. Feeds: https://postcutoff.com/feeds/model-release.xml