AI news for builders and product teamsUpdated Oct 10, 2026, 19:01 UTC
Model releases news
The latest Model releases stories across our sources, prepared from the publishers’ own reporting.
In this topic
Newest first
Google DeepMind announced Gemini 4 Argon, a frontier model rolling out to trusted cyber defenders through its Fairwind Program before wider release. It offers a 1 million token output limit, introductory pricing of $2 per million input tokens and $10 per million output tokens, and reported benchmark results across software engineering, knowledge work, and cybersecurity.

Anthropic and OpenAI released several new models within weeks of each other, with Anthropic launching Claude Opus 5.5 and Sonnet 5.5 and OpenAI releasing GPT-6 Sol and Luna and, at DevDay, GPT-6.1 Sol. OpenAI also published a site documenting nine misalignment incidents and launched Dots agents on GPT-6 Astra.

Anthropic released Claude Sonnet 5.5, which a newsletter contributor describes as benchmark-close to Opus 5.5 in coding with a big jump in understanding images and charts. The same roundup covers Manus 2.0, Cue, GPT-6 Astra, OpenAI agents accessing Australian government websites, and Meta's new enterprise platform.

xAI's Grok 4.7 is now available on Amazon Bedrock, offering a 500K token context window and four configurable reasoning effort levels: low, medium, high, and xhigh. It is served on the bedrock-runtime endpoint through cross-Region inference profiles and supports the Responses, Chat Completions, and Converse APIs.

AWS announced Claude Sonnet 5.5 is available on Amazon Bedrock and Claude Platform on AWS, described as a more efficient Sonnet model for focused coding and knowledge work with lower cost per task for most work at faster speed.

Anthropic released Claude Opus 5.5, which it says performs at the level of Claude Fable 5.1 for most tasks and costs 40% less to run than Opus 5. OpenAI released GPT-6 Luna and GPT-6 Sol, both 50% cheaper than their previous versions and available only in Codex and ChatGPT Work for now.

NVIDIA released NV-Reason-CT, an open 3D CT vision language model that generates radiologist-style chain-of-thought reasoning. It combines a 3D ViT encoder with a Qwen3.5-4B language model and reports state-of-the-art results on the CT-RATE benchmark, with NIH radiologists validating the clinical plausibility of its reasoning traces.

Together AI launched together/Tev1-4B-experimental, a classifier built on Qwen3.5 4B, and published a guide to fine-tuning a similar model. The post says training on 38,340 examples costs about $17 and takes roughly 25 minutes.

Anthropic released Claude Opus 5.5 and OpenAI released GPT-6 Sol and GPT-6 Luna, with GPT-6 Sol and Luna priced at half their GPT-5.6 equivalents, according to Simon Willison. Opus 5.5's input and output prices fell 20% versus prior Opus models.

Anthropic released Claude Opus 5.5, which it says delivers Fable 5.1-level performance for most work at about 40% lower cost, with tokens priced 20% below Opus 5. The company also raised five-hour usage limits by 20% and said Claude Sonnet 5.5 and Haiku 5.5 are coming in the next few weeks.

Diogo Almeida, CEO of TypeSafe AI, discussed the launch of Jev, a "System One" model optimized for intelligence per dollar, on the Latent Space podcast. Almeida described Jev's core innovation, Reinforcement Learning for Calibrated Decisions, and argued against RLHF and RLVR as training paradigms.

OpenAI released GPT-6 Astra last week, and the article covers the author's impressions of its benchmarks and computer-use capabilities, plus reported architectural details about looped transformers and rumors that the model hides its reasoning traces.

Astra, described as the start of the GPT-6 family, has been released and is said to top ARC-AGI-3 and Zapier's AutomationBench while costing the same as Fable 5.1. Commentators describe it as powerful but inconsistent, and OpenAI says it has met its automated research intern goal.

OpenAI has released GPT-6 Astra, a new OpenAI agent message board has been discovered, and Anthropic launched Claude Fable 5.1, which it says is up to 45 percent cheaper for agentic work.

Hugging Face announced NeoMME, a family of 260M and 800M multilingual multimodal encoders that process text tokens and raw image patches in a single bidirectional Transformer trained from scratch. Checkpoints are released under the Apache 2.0 license, with a fine-tuned NeoMME-Retriever for visual document retrieval.

IBM's Granite Team published a technical walkthrough of Granite 4.2, its first family of dense, decoder-only reasoning LLMs, released in 3B, 8B, and 30B sizes under the Apache 2.0 license. The models are post-trained from Granite-4.1 base models via SFT and a multi-stage RL pipeline.

Liquid AI released DSpark draft model checkpoints for three LFM2.5 models, adding speculative decoding that it reports delivers up to 3.18x throughput improvement on GPU and up to 2.87x on-device without changing output quality.

Z.ai announced GLM-5.3, a coding-plan-only model reported to surpass Moonshot AI's Kimi K3 on many benchmarks and some Claude Fable 5 or GPT-5.6-Sol scores, with open weights due on Hugging Face in two weeks. Z.ai says it scaled post-training on the same base model as GLM-5.2.

NVIDIA Nemotron 3.5 Lightning, a 30 billion parameter open model with 3B active parameters, is now available on Ollama and runs entirely on local devices. It is designed for long-running agentic tasks such as tool calling and multi-step work.

Meta Superintelligence Labs has released Muse Glimmer, a 30B multimodal open model under the Apache 2.0 license, now available on Ollama for local coding agents and personal assistants. Ollama's MLX engine adds DFlash and image input support, running Muse Glimmer 1.5x-1.8x faster on Apple Silicon.