AivexaNewsSearch
AI news for builders and product teamsUpdated Oct 10, 2026, 22:01 UTC

Developer tools news

The latest Developer tools stories across our sources, prepared from the publishers’ own reporting.

In this topic

Newest first

llm 0.36

Simon Willison released llm 0.36, adding support for new OpenAI models gpt-6-sol and gpt-6-luna, a way for model plugins to declare they do not support conversations, and collapsed reasoning traces in llm logs Markdown output.

Read original

Hardware-Agnostic Models in vLLM

vLLM is introducing new "HW agnostic" layers to keep the project portable across diverse hardware as its frontier optimizations move away from fullgraph torch.compile. On NVIDIA H100 GPUs, these layers achieve total token throughput within 3.4% of the native implementation (geometric mean across three recent models).

Read original

Simplifying Model Serving Across Multiple GPUs with NVIDIA TensorRT Multi-Device Integration in NVIDIA Dynamo-Triton

NVIDIA Dynamo-Triton release 26.07 enables the TensorRT backend multi-device capability, letting one KIND_MODEL instance own multiple GPUs and serve distributed inference through a single gRPC endpoint. A Cosmos 3 Nano demonstration cut end-to-end generation latency from 156.6 seconds on one GPU to 34.2 seconds on eight GPUs.

Read original

How to Evaluate AI Agents From Tool Calls to Task Completion

NVIDIA published a developer blog explaining how AI agent evaluation has shifted from scoring single function calls to measuring full task completion in executable environments. It describes step-level and end-to-end scoring on execution traces, a benchmark-trial-task-turn-step metric hierarchy, and reports Nemotron 3.5 Lightning at 86% accuracy on PinchBench while finishing tasks 30% faster than Qwen3.6 35B at comparable accuracy.

Read original

Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2

NVIDIA published a tutorial on AI data assimilation tools in its Earth-2 platform, covering Score-Based Data Assimilation for regional diffusion models and HealDA for global atmospheric state estimation. It reports wind-speed RMSE reductions of 54% in one CorrDiff-COSMO downscaling example and an average of 7.2% across six StormCast-CONUS forecast steps.

Read original