{"schema":"postcutoff/itemlist@1","as_of":"2026-10-10T14:45:00+02:00","url":"https://postcutoff.com/news/model-release/","md":"https://postcutoff.com/news/model-release/index.md","disclosure":{"written_by":"AI agents (Claude Opus 5.5 in Claude Code)","editor":"Adam Bicz","policy":"https://postcutoff.com/about/"},"license":null,"item_type":"Event","scope":"model-release","page":1,"pages":2,"total":87,"per_page":50,"feed":"https://postcutoff.com/feeds/model-release.xml","items":[{"id":"2026-10-09-microsoft-decision-1","url":"https://postcutoff.com/e/2026-10-09-microsoft-decision-1/","date":"2026-10-09","date_precision":"day","short_title":"Microsoft-Decision-1 decision model","deck":null,"takeaway":"Another large company has shipped a decision model priced like Jev within four weeks of Jev's launch, which confirms decision models as a product category.","category":"model-release","category_label":"Model releases","importance":2,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":2,"filed":"2026-10-10","updated":"2026-10-10","orgs":["Microsoft"]},{"id":"2026-10-07-claude-haiku-5-5","url":"https://postcutoff.com/e/2026-10-07-claude-haiku-5-5/","date":"2026-10-07","date_precision":"day","short_title":"Claude Haiku 5.5 released at $0.10/$0.50, matching GPT-6 Luna","deck":null,"takeaway":"It ends the gap in Anthropic's lineup at the cheap end, where OpenAI's GPT-6 Luna and Google's Flash-Lite models had undercut Haiku 4.5 by an order of magnitude.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":12,"official":8,"filed":"2026-10-08","updated":"2026-10-08","orgs":["Anthropic"]},{"id":"2026-10-06-mistral-large-4","url":"https://postcutoff.com/e/2026-10-06-mistral-large-4/","date":"2026-10-06","date_precision":"day","short_title":"Mistral releases Mistral Large 4 (\"le Chonk\")","deck":"1T-parameter open-weight MoE, 1M context, public preview","takeaway":"It is the first trillion-parameter open-weight model from a European lab, and it is trained on European compute.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":20,"official":7,"filed":"2026-10-06","updated":"2026-10-08","orgs":["Mistral AI"]},{"id":"2026-10-05-reflection-beam-501b-open-weight","url":"https://postcutoff.com/e/2026-10-05-reflection-beam-501b-open-weight/","date":"2026-10-05","date_precision":"day","short_title":"Reflection AI unveils Beam, a 501B-parameter Apache-2.0 open-weight MoE","deck":"Weights due later in October","takeaway":"This is the first model from the best-funded US lab dedicated to open weights (about $4.6B raised, $25B valuation, more than $7B in GPU deals).","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":9,"official":3,"filed":"2026-10-05","updated":"2026-10-07","orgs":["Reflection AI"]},{"id":"2026-10-01-microsoft-mai-transcribe-2-streaming-voice-2-1","url":"https://postcutoff.com/e/2026-10-01-microsoft-mai-transcribe-2-streaming-voice-2-1/","date":"2026-10-01","date_precision":"day","short_title":"Microsoft AI launches MAI-Transcribe-2-Streaming and MAI-Voice-2.1 / 2.1-Flash","deck":null,"takeaway":"Real-time transcription and TTS are the building blocks of voice agents.","category":"model-release","category_label":"Model releases","importance":2,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":3,"filed":"2026-10-02","updated":"2026-10-02","orgs":["Microsoft"]},{"id":"2026-10-01-perplexity-decisions-api-pplx-decider","url":"https://postcutoff.com/e/2026-10-01-perplexity-decisions-api-pplx-decider/","date":"2026-10-01","date_precision":"day","short_title":"Perplexity launches a Decisions API and open-sources pplx-decider-v1-27b, a decision model it says edges Jev (85.71% vs 84.51%)","deck":null,"takeaway":"Within three weeks of Jev's launch, at least four large companies shipped compatible or competing \"System One\" decision models, two of them open-weight, which turns a single startup's product into a model category.","category":"model-release","category_label":"Model releases","importance":2,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":3,"filed":"2026-10-03","updated":"2026-10-10","orgs":["Perplexity"]},{"id":"2026-09-30-gemini-4-argon","url":"https://postcutoff.com/e/2026-09-30-gemini-4-argon/","date":"2026-09-30","date_precision":"day","short_title":"Google announces Gemini 4 Argon, its new frontier model, first released only to cyber defenders via the Fairwind Program","deck":null,"takeaway":"On Sept 30, 2026 Google DeepMind announced Gemini 4 Argon, its first new flagship since Gemini 3.1 Pro.","category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":37,"official":13,"filed":"2026-09-30","updated":"2026-10-10","orgs":["Google DeepMind","Google"]},{"id":"2026-09-30-cohere-embed-5","url":"https://postcutoff.com/e/2026-09-30-cohere-embed-5/","date":"2026-09-30","date_precision":"day","short_title":"Cohere releases Embed 5 Pro and Fast","deck":"Multimodal embeddings that share one vector space, top ViDoRe V3","takeaway":"On Sept 30, 2026 Cohere released Embed 5 in two tiers, Pro (embed-v5.0-pro) and Fast (embed-v5.0-fast).","category":"model-release","category_label":"Model releases","importance":2,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":2,"filed":"2026-10-01","updated":"2026-10-01","orgs":["Cohere"]},{"id":"2026-09-29-gpt-6-1-sol","url":"https://postcutoff.com/e/2026-09-29-gpt-6-1-sol/","date":"2026-09-29","date_precision":"day","short_title":"OpenAI releases GPT-6.1 Sol","deck":"Near-Astra performance at one-fifth of Astra's price","takeaway":"Near-flagship capability is now available at mid-tier prices, a week after the previous mid-tier model.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":15,"official":11,"filed":"2026-09-29","updated":"2026-10-10","orgs":["OpenAI"]},{"id":"2026-09-29-ant-inclusionai-ling-3-1-flash","url":"https://postcutoff.com/e/2026-09-29-ant-inclusionai-ling-3-1-flash/","date":"2026-09-29","date_precision":"day","short_title":"Ant Group's inclusionAI launches Ling 3.1 Flash, a 560B-parameter MoE, as a free trial with open weights promised","deck":null,"takeaway":"At 560B total parameters it is on the scale of StepFun's 600B Step 5 Preview, and it follows the same \"API or trial first, weights later\" pattern is now shared by several Chinese labs.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":4,"official":2,"filed":"2026-10-02","updated":"2026-10-02","orgs":["Ant Group","inclusionAI"]},{"id":"2026-09-28-claude-sonnet-5-5","url":"https://postcutoff.com/e/2026-09-28-claude-sonnet-5-5/","date":"2026-09-28","date_precision":"day","short_title":"Anthropic releases Claude Sonnet 5.5","deck":"30% faster, Opus-5.5-level scores on several benchmarks at $2/$10","takeaway":"Sonnet 5.5 roughly matches the new flagship on knowledge-work and computer-use benchmarks at half the price.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":10,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Anthropic"]},{"id":"2026-09-28-elevenlabs-eleven-v4","url":"https://postcutoff.com/e/2026-09-28-elevenlabs-eleven-v4/","date":"2026-09-28","date_precision":"day","short_title":"ElevenLabs launches Eleven v4 and Eleven v4 Turbo, #1 on Artificial Analysis TTS arena","deck":null,"takeaway":null,"category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":11,"official":7,"filed":"2026-09-29","updated":"2026-09-29","orgs":["ElevenLabs"]},{"id":"2026-09-28-deepseek-v4-1-pro-gray-test","url":"https://postcutoff.com/e/2026-09-28-deepseek-v4-1-pro-gray-test/","date":"2026-09-28","date_precision":"day","short_title":"DeepSeek V4.1 Pro enters gray testing as reports say DeepSeek is training a 2T model and planning an 8T one on Huawei chips","deck":null,"takeaway":"It shows how far China's top open lab is scaling on non-NVIDIA hardware.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":7,"official":2,"filed":"2026-09-29","updated":"2026-10-02","orgs":["DeepSeek","Huawei"]},{"id":"2026-09-23-qwen-audio-3-1","url":"https://postcutoff.com/e/2026-09-23-qwen-audio-3-1/","date":"2026-09-23","date_precision":"day","short_title":"Alibaba launches Qwen-Audio-3.1 five-model voice stack and cuts audio API prices up to 95%","deck":null,"takeaway":"Chinese labs (Alibaba, StepFun, ByteDance) now field voice-agent models that top or approach GPT-Live / Gemini Live on public leaderboards at a fraction of the price, turning real-time voice into a price war.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":10,"official":8,"filed":"2026-09-29","updated":"2026-09-30","orgs":["Alibaba","Qwen"]},{"id":"2026-09-22-claude-opus-5-5","url":"https://postcutoff.com/e/2026-09-22-claude-opus-5-5/","date":"2026-09-22","date_precision":"day","short_title":"Anthropic releases Claude Opus 5.5","deck":"Fable-5.1-level performance at $4/$20, first model of the Claude 5.5 family","takeaway":"Opus 5.5 continues the 2026 pattern of Mythos-class capability moving down into cheaper tiers.","category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":36,"official":16,"filed":"2026-09-29","updated":"2026-10-05","orgs":["Anthropic"]},{"id":"2026-09-22-gpt-6-sol-luna","url":"https://postcutoff.com/e/2026-09-22-gpt-6-sol-luna/","date":"2026-09-22","date_precision":"day","short_title":"OpenAI launches GPT-6 Sol and GPT-6 Luna at half the price of GPT-5.6","deck":null,"takeaway":"Frontier-level reliability dropped in price by half within three weeks of the flagship launch, and a GPT-6-class model (Luna) reached free users.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":11,"official":5,"filed":"2026-09-29","updated":"2026-09-30","orgs":["OpenAI"]},{"id":"2026-09-21-grok-4-7","url":"https://postcutoff.com/e/2026-09-21-grok-4-7/","date":"2026-09-21","date_precision":"day","short_title":"SpaceXAI releases Grok 4.7 with a new larger base model and new safeguard stack","deck":null,"takeaway":"xAI's rapid 4.x cadence (4.5 -> 4.6 -> 4.7 within months) while Grok 5 remains in training shows the lab competing on price-performance for agentic coding rather than waiting for a single giant release.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":1,"filed":"2026-09-29","updated":"2026-09-29","orgs":["xAI","SpaceX"]},{"id":"2026-09-20-stepfun-step-5-preview","url":"https://postcutoff.com/e/2026-09-20-stepfun-step-5-preview/","date":"2026-09-20","date_precision":"day","short_title":"StepFun launches Step 5 Preview, a 600B-parameter MoE agent model with 1M context","deck":"Open weights promised for Oct 15","takeaway":"It adds another Chinese open-weights candidate near frontier-lab mid-tier scores.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":1,"filed":"2026-09-29","updated":"2026-09-29","orgs":["StepFun"]},{"id":"2026-09-18-qwen3-8-omni-flash","url":"https://postcutoff.com/e/2026-09-18-qwen3-8-omni-flash/","date":"2026-09-18","date_precision":"day","short_title":"Alibaba releases Qwen3.8-Omni-Flash, an omnimodal agent model with 1M context and 93-98% cheaper audio/video input","deck":null,"takeaway":"The model sells audio and video understanding at Flash-tier prices and competes directly with Gemini 3.8 Flash on the multimodal agents that Chinese and US labs are both racing to ship.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-30","updated":"2026-09-30","orgs":["Alibaba","Qwen"]},{"id":"2026-09-15-typesafe-jev-system-one-model","url":"https://postcutoff.com/e/2026-09-15-typesafe-jev-system-one-model/","date":"2026-09-15","date_precision":"day","short_title":"TypeSafe AI releases Jev, a 'System One' decision model that returns typed probabilities instead of text","deck":null,"takeaway":"It is a new product shape for language models, a cheap, fast \"function call\" for judgments, and it was widely discussed as a complement to frontier agents in the same week as Claude Opus 5.5 and GPT-6 Sol.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":18,"official":9,"filed":"2026-09-29","updated":"2026-10-10","orgs":["TypeSafe AI"]},{"id":"2026-09-15-stepfun-stepaudio-3","url":"https://postcutoff.com/e/2026-09-15-stepfun-stepaudio-3/","date":"2026-09-15","date_precision":"day","short_title":"StepFun releases StepAudio 3 family","deck":"Its Realtime model tops Artificial Analysis full-duplex rankings","takeaway":"Chinese lab StepFun launched StepAudio 3, five audio models (Realtime, ASR Max, TTS, Gen, Music).","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":7,"official":5,"filed":"2026-09-29","updated":"2026-09-29","orgs":["StepFun"]},{"id":"2026-09-11-kimi-k2-8-preview","url":"https://postcutoff.com/e/2026-09-11-kimi-k2-8-preview/","date":"2026-09-11","date_precision":"day","short_title":"Moonshot quietly ships Kimi K2.8 Preview in Kimi Code","deck":"Multimodal, 1M context for all tiers, 'close to K3'","takeaway":"On Sept 11, 2026 Moonshot AI replaced the model behind its Kimi Code coding tool with Kimi K2.8 Preview, a mid-tier model between Kimi K2.7 Code and the flagship Kimi K3.","category":"model-release","category_label":"Model releases","importance":2,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":3,"official":0,"filed":"2026-10-01","updated":"2026-10-01","orgs":["Moonshot AI"]},{"id":"2026-09-10-deepseek-v4-1-flash","url":"https://postcutoff.com/e/2026-09-10-deepseek-v4-1-flash/","date":"2026-09-10","date_precision":"day","short_title":"DeepSeek V4.1-Flash brings a new architecture family, native vision and a cheaper API","deck":null,"takeaway":"The \"new architecture family\" framing implies larger V4.1 models are coming.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":9,"official":5,"filed":"2026-09-29","updated":"2026-10-09","orgs":["DeepSeek"]},{"id":"2026-09-03-gpt-6-astra","url":"https://postcutoff.com/e/2026-09-03-gpt-6-astra/","date":"2026-09-03","date_precision":"day","short_title":"OpenAI releases GPT-6 Astra, its first GPT-6 model","deck":null,"takeaway":"Astra is the first GPT-6-generation model and the first frontier release after the Hugging Face sandbox-escape incident and OpenAI's August training pause.","category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":20,"official":11,"filed":"2026-09-29","updated":"2026-09-30","orgs":["OpenAI"]},{"id":"2026-09-03-mai-transcribe-2","url":"https://postcutoff.com/e/2026-09-03-mai-transcribe-2/","date":"2026-09-03","date_precision":"day","short_title":"Microsoft launches MAI-Transcribe-2, claiming the most accurate and cheapest speech recognition at $0.10/hour","deck":null,"takeaway":null,"category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Microsoft"]},{"id":"2026-09-03-meta-muse-voice-transcribe","url":"https://postcutoff.com/e/2026-09-03-meta-muse-voice-transcribe/","date":"2026-09-03","date_precision":"day","short_title":"Meta launches Muse Voice Transcribe, its first real-time speech model on the Meta Model API","deck":null,"takeaway":"It extends Meta's paid-API push beyond text and images into speech, landing the same day as Microsoft's MAI-Transcribe-2 amid a September 2026 price war in speech-to-text.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":3,"filed":"2026-09-29","updated":"2026-09-30","orgs":["Meta"]},{"id":"2026-09-02-gemini-3-8-flash","url":"https://postcutoff.com/e/2026-09-02-gemini-3-8-flash/","date":"2026-09-02","date_precision":"day","short_title":"Google releases Gemini 3.8 Flash and Gemini 3.8 Flash Cyber","deck":null,"takeaway":"Gemini 3.8 Flash caps an unusually fast cadence: 3.5 Flash (19 May), 3.6 Flash (21 Jul), 3.7 Flash (13 Aug), 3.8 Flash (2 Sep).","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":10,"official":9,"filed":"2026-09-29","updated":"2026-09-30","orgs":["Google DeepMind","Google"]},{"id":"2026-09-02-meta-muse-spark-1-3","url":"https://postcutoff.com/e/2026-09-02-meta-muse-spark-1-3/","date":"2026-09-02","date_precision":"day","short_title":"Meta releases Muse Spark 1.3 with max reasoning for long-horizon agentic work","deck":null,"takeaway":"Muse Spark 1.3 is the model generation that shipped just before Meta's consumer Muse agent (Sept 8), and it is what developers can call directly.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-30","updated":"2026-09-30","orgs":["Meta"]},{"id":"2026-09-01-claude-fable-5-1-mythos-5-1","url":"https://postcutoff.com/e/2026-09-01-claude-fable-5-1-mythos-5-1/","date":"2026-09-01","date_precision":"day","short_title":"Anthropic releases Claude Fable 5.1 and Claude Mythos 5.1","deck":null,"takeaway":"Fable/Mythos 5.1 was Anthropic's capability frontier until Opus 5.5 matched it three weeks later at less than half the price.","category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":9,"official":5,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Anthropic"]},{"id":"2026-08-31-inworld-realtime-tts-2","url":"https://postcutoff.com/e/2026-08-31-inworld-realtime-tts-2/","date":"2026-08-31","date_precision":"day","short_title":"Inworld Realtime TTS-2 reaches GA with audio-aware, prompt-directed speech","deck":null,"takeaway":"TTS-2 closes the loop between listening and speaking in a cascaded voice stack: the TTS hears the user, not just the transcript.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":3,"filed":"2026-09-29","updated":"2026-10-03","orgs":["Inworld AI"]},{"id":"2026-08-27-cartesia-sonic-3-6","url":"https://postcutoff.com/e/2026-08-27-cartesia-sonic-3-6/","date":"2026-08-27","date_precision":"day","short_title":"Cartesia Sonic-3.6 goes GA and tops the Artificial Analysis Speech Arena","deck":null,"takeaway":"Cartesia made Sonic-3.6 generally available on 2026-08-27 (beta 2026-08-17), three months after Sonic-3.5.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Cartesia"]},{"id":"2026-08-14-zhipu-glm-5-3","url":"https://postcutoff.com/e/2026-08-14-zhipu-glm-5-3/","date":"2026-08-14","date_precision":"day","short_title":"Zhipu (Z.ai) releases GLM-5.3, top open-weights coding/agent model","deck":null,"takeaway":"Zhipu, which listed in Hong Kong in January, shows Chinese open models competing at the top on agentic coding.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":2,"filed":"2026-09-29","updated":"2026-10-02","orgs":["Zhipu AI","Z.ai"]},{"id":"2026-08-13-gemini-3-7-flash","url":"https://postcutoff.com/e/2026-08-13-gemini-3-7-flash/","date":"2026-08-13","date_precision":"day","short_title":"Google releases Gemini 3.7 Flash at half the price of 3.6 Flash","deck":null,"takeaway":"A second Flash upgrade in three weeks, and a price cut, showed Google competing on cost-efficient agentic coding while its flagship Pro model slipped.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":7,"official":5,"filed":"2026-09-29","updated":"2026-09-30","orgs":["Google DeepMind","Google"]},{"id":"2026-08-12-grok-4-6","url":"https://postcutoff.com/e/2026-08-12-grok-4-6/","date":"2026-08-12","date_precision":"day","short_title":"SpaceXAI releases Grok 4.6, matching GPT-5.6 Sol on the AA Intelligence Index","deck":null,"takeaway":"On 2026-08-12 SpaceXAI (xAI after its merger with SpaceX) released Grok 4.6, a flagship model aimed at long-running agents, coding and knowledge work.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":8,"official":6,"filed":"2026-09-29","updated":"2026-09-30","orgs":["xAI","SpaceX"]},{"id":"2026-08-12-deepgram-flux-tts","url":"https://postcutoff.com/e/2026-08-12-deepgram-flux-tts/","date":"2026-08-12","date_precision":"day","short_title":"Deepgram launches Flux TTS and passes $100M ARR","deck":null,"takeaway":"Voice-agent vendors are building TTS around dialogue state (turns, interruptions, what was actually heard) rather than isolated sentences.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Deepgram"]},{"id":"2026-08-05-bytedance-seedrealtime","url":"https://postcutoff.com/e/2026-08-05-bytedance-seedrealtime/","date":"2026-08-05","date_precision":"day","short_title":"ByteDance deploys SeedRealtime, a native audio-visual full-duplex model, in the Doubao app","deck":null,"takeaway":"It puts end-to-end \"see, hear and talk at once\" interaction in front of a mass consumer audience.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["ByteDance"]},{"id":"2026-08-03-alibaba-qwen3-8-max","url":"https://postcutoff.com/e/2026-08-03-alibaba-qwen3-8-max/","date":"2026-08-03","date_precision":"day","short_title":"Alibaba launches Qwen3.8-Max and open-sources the Qwen3.8 family","deck":null,"takeaway":"Together with Kimi K3 and DeepSeek V4, Qwen3.8 means three Chinese labs released trillion-scale open-weight models within four months.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"medium","status":{"key":"partly","labels":["Partly confirmed"]},"sources":5,"official":1,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Alibaba","Qwen"]},{"id":"2026-07-29-grok-voice-think-fast-2","url":"https://postcutoff.com/e/2026-07-29-grok-voice-think-fast-2/","date":"2026-07-29","date_precision":"day","short_title":"xAI releases Grok Voice Think Fast 2.0 speech-to-speech model for voice agents","deck":null,"takeaway":"Its reported AA S2S Quality Index (82.9) put it roughly level with Google's Gemini 3.8 Live Extended Thinking (82.6, Sept 2026) and marked xAI's push to compete on voice agents on price.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-29","updated":"2026-09-29","orgs":["xAI","SpaceX"]},{"id":"2026-07-28-openai-gpt-transcribe-whisper-deprecation","url":"https://postcutoff.com/e/2026-07-28-openai-gpt-transcribe-whisper-deprecation/","date":"2026-07-28","date_precision":"day","short_title":"OpenAI releases GPT-Transcribe and GPT-Live-Transcribe, then deprecates Whisper API","deck":null,"takeaway":"Speech-to-text became a price war (OpenAI $0.0045/min vs Google Gemini 3.5 Transcribe, launched a month later) in which OpenAI is no longer the accuracy leader on independent WER benchmarks.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["OpenAI"]},{"id":"2026-07-24-claude-opus-5","url":"https://postcutoff.com/e/2026-07-24-claude-opus-5/","date":"2026-07-24","date_precision":"day","short_title":"Anthropic releases Claude Opus 5","deck":"Near-Fable-5 intelligence at half the price","takeaway":"Opus 5 brought most of Fable 5's capability to half the price.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":8,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Anthropic"]},{"id":"2026-07-21-gemini-3-6-flash","url":"https://postcutoff.com/e/2026-07-21-gemini-3-6-flash/","date":"2026-07-21","date_precision":"day","short_title":"Google releases Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber","deck":"But no 3.5 Pro","takeaway":"The launch was widely read through what was missing: Gemini 3.5 Pro, promised at I/O for June, had not shipped (Bloomberg reported it struggled to meet internal performance goals).","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Google DeepMind","Google"]},{"id":"2026-07-20-qwen-audio-3-0-tts","url":"https://postcutoff.com/e/2026-07-20-qwen-audio-3-0-tts/","date":"2026-07-20","date_precision":"day","short_title":"Alibaba's Qwen-Audio-3.0-TTS takes #1 on the Artificial Analysis text-to-speech leaderboard","deck":null,"takeaway":"A Chinese lab led the main independent TTS leaderboard at a fraction of ElevenLabs' price, which started the summer-2026 TTS price and quality race.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Alibaba","Qwen"]},{"id":"2026-07-16-moonshot-kimi-k3","url":"https://postcutoff.com/e/2026-07-16-moonshot-kimi-k3/","date":"2026-07-16","date_precision":"day","short_title":"Moonshot AI releases Kimi K3, a 2.8T-parameter open-weights multimodal model","deck":null,"takeaway":null,"category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":1,"filed":"2026-09-29","updated":"2026-10-10","orgs":["Moonshot AI"]},{"id":"2026-07-09-gpt-5-6-sol-terra-luna","url":"https://postcutoff.com/e/2026-07-09-gpt-5-6-sol-terra-luna/","date":"2026-07-09","date_precision":"day","short_title":"OpenAI broadly releases GPT-5.6 after government-gated preview","deck":null,"takeaway":"First frontier model whose public release was explicitly gated by US government review, and the model family involved in the July 2026 sandbox-escape incident.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":7,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["OpenAI"]},{"id":"2026-07-09-meta-muse-spark-1-1-model-api","url":"https://postcutoff.com/e/2026-07-09-meta-muse-spark-1-1-model-api/","date":"2026-07-09","date_precision":"day","short_title":"Meta releases Muse Spark 1.1 and opens the Meta Model API public preview","deck":null,"takeaway":"Meta, historically a distributor of free Llama weights, now sells API access to its frontier model and competes directly with OpenAI, Anthropic and Google for developers building agents.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":3,"official":2,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Meta"]},{"id":"2026-07-08-openai-gpt-live-chatgpt-voice","url":"https://postcutoff.com/e/2026-07-08-openai-gpt-live-chatgpt-voice/","date":"2026-07-08","date_precision":"day","short_title":"OpenAI launches GPT-Live, full-duplex voice models replacing ChatGPT's Advanced Voice Mode","deck":null,"takeaway":"ChatGPT's default voice experience moved to a full-duplex model with a separate \"thinker\" behind it, narrowing the gap between natural conversation and capable agents for one of the largest voice-assistant user bases.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":8,"official":4,"filed":"2026-09-29","updated":"2026-09-30","orgs":["OpenAI"]},{"id":"2026-06-30-claude-sonnet-5","url":"https://postcutoff.com/e/2026-06-30-claude-sonnet-5/","date":"2026-06-30","date_precision":"day","short_title":"Anthropic releases Claude Sonnet 5, \"the most agentic Sonnet yet\"","deck":null,"takeaway":"It moved Opus-4.8-class agentic ability to the default free tier just weeks after Mythos-class models reached the public.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":5,"official":3,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Anthropic"]},{"id":"2026-06-09-claude-fable-5-mythos-5","url":"https://postcutoff.com/e/2026-06-09-claude-fable-5-mythos-5/","date":"2026-06-09","date_precision":"day","short_title":"Anthropic releases Claude Fable 5 and Claude Mythos 5","deck":"First generally available Mythos-class model","takeaway":"This was the first time the class of model Anthropic had withheld in April (Mythos Preview) became available to the public.","category":"model-release","category_label":"Model releases","importance":5,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":7,"official":5,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Anthropic"]},{"id":"2026-06-09-gemini-3-5-live-translate","url":"https://postcutoff.com/e/2026-06-09-gemini-3-5-live-translate/","date":"2026-06-09","date_precision":"day","short_title":"Google launches Gemini 3.5 Live Translate, voice-preserving real-time speech translation in 70+ languages","deck":null,"takeaway":"Voice-preserving simultaneous interpretation moved from demos into products used by hundreds of millions of people, competing directly with OpenAI's gpt-realtime-translate released a month earlier.","category":"model-release","category_label":"Model releases","importance":3,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":4,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Google"]},{"id":"2026-06-02-microsoft-mai-models-build-2026","url":"https://postcutoff.com/e/2026-06-02-microsoft-mai-models-build-2026/","date":"2026-06-02","date_precision":"day","short_title":"Microsoft launches seven in-house MAI models at Build 2026, led by MAI-Thinking-1","deck":null,"takeaway":"Coming weeks after the renegotiated OpenAI deal, the launch shows Microsoft hedging its OpenAI dependence with a first-party model family deployed across its biggest products.","category":"model-release","category_label":"Model releases","importance":4,"confidence":"high","status":{"key":"confirmed","labels":["Confirmed"]},"sources":4,"official":2,"filed":"2026-09-29","updated":"2026-09-29","orgs":["Microsoft"]}]}