Release tracker

Model, API, and platform releases that matter if you run AI in production, with a one-line note on why. Newest first. Updated October 9, 2026.

RSS feed

Vendor
Type
DateReleaseVendorTypeStatusWhy it matters
Oct 9Microsoft-Decision-1MicrosoftSmall modelGADecision-scoring model post-trained from Qwen3.5-9B; returns probabilities per answer option. $0.042/M input, $0 output on OpenRouter.
Oct 9Dynamic workflows (Claude Managed Agents)AnthropicAPIPreviewAn agent writes a workflow that runs up to 1,000 agents in phases; beta behind managed-agents-2026-04-01.
Oct 8Gemini agent for workGoogleProductAnnouncedSingle "universal agent" across Workspace and Gemini Enterprise, announced at Gemini at Work.
Oct 7GPT-6 Sol / GPT-6 LunaOpenAILLMRolling outChatGPT models with Intelligent UI. Sol for paid tiers, Luna for Free and Go.
Oct 7GPT-6.1 Sol (API) + Decisions APIOpenAIAPIGACoding and computer-use model at $0.10/M cached input; Decisions API in limited preview.
Oct 7Claude Haiku 5.5AnthropicSmall modelGA$0.10 in / $0.50 out per million tokens under 100K context. About 75% cheaper than Haiku 4.5.
Oct 7Claude Opus 5.5AnthropicLLMGAPositioned at Fable 5.1 quality on most work, 40% cheaper to run than Opus 5.
Oct 7GLM 5.3 FastZ.AILLMGALatency-tuned member of the GLM 5.3 family.
Oct 6EmbeddingGemma 2GoogleSmall modelGAGoogle says the open-weight (Apache 2.0) 740M multimodal embedder maps text, images, audio and video into one space, with 768-dim vectors truncatable to 128.
Oct 6Mistral Large 4MistralLLMPreview1.05T-parameter MoE (52B active), 1M context. Open weights promised for October 27.
Oct 6Gemini Nano Banana 2.1GoogleImageGAEfficiency-focused image generation model; GA in the Gemini API on October 8.
Oct 5HY Image 3.5 PreviewTencent CloudImagePreviewImage generation preview.
Oct 2Ling 3.1 FlashinclusionAILLMGAFast-tier open model.
Sep 30Gemini 4 ArgonGoogleLLMPreviewGoogle says access is limited to trusted cyber defenders for now, with paid API customers and AI Ultra next; introductory pricing is $2 input / $10 output per million tokens.