AI news for builders and product teamsUpdated Oct 10, 2026, 19:01 UTC
Developer tools news
The latest Developer tools stories across our sources, prepared from the publishers’ own reporting.
In this topic
Newest first
An AWS Machine Learning Blog post walks through connecting Claude Desktop on Amazon Bedrock to Web Search through Amazon Bedrock AgentCore Gateway, using JWT-based inbound authentication. The setup federates AWS IAM Identity Center with Amazon Cognito as the JWT issuer, and Web Search is available in three AWS Regions.

A GitHub Blog post outlines three skills for developers as AI changes their work: directing AI agents, reviewing AI output rather than trusting the first answer, and using saved implementation time on broader problems.

Cloudflare has launched a Web Search API through its AI Gateway in partnership with Ceramic.ai, Exa and Linkup, letting agents pull live web results into model context instead of guessing URLs. Partners must follow Cloudflare's Verified bots and robots.txt standards, and queries are billed at provider list pricing with no markup.

Docker announced it is bringing the Sandbox Kit Specification, now at v3 under Apache 2.0, to the CNCF, packaging an AI agent, its tools, and requested permissions into an OCI image. Docker announced the move at WeAreDevelopers on September 24.

Pi 1.0 and Pi Durable from Earendil both reached the front page of Hacker News, with Pi Durable porting Pi to TypeScript and externalizing stateful components. The same AINews roundup also covers Gemini 4 Argon, GPT-6.1 Sol, Upstage's Solar Mini 4, and Black Forest Labs' FLUX 3 Image.

An AWS Machine Learning Blog post describes a four-agent pattern built with the Strands Agents SDK on Amazon Bedrock AgentCore for enterprise cloud migrations. The post reports that on a 300+ application program, the pattern cut infrastructure as code development from 3 to 4 weeks per application to minutes, based on internal project tracking data.

Databricks described read restrictions and catalog labels as a way to unify governance across engines and catalogs, continuing its series on open table formats, open APIs and unified governance. No specific version, date, pricing or availability details were stated.

NVIDIA released DOCA AI agent skills on GitHub that provide verified API signatures, hardware capability requirements and build constraints for BlueField DPU development. In a 65-prompt evaluation, agents without skills satisfied 19% of checklist items versus 100% with skills.

NVIDIA's Do Inference Now (DIN) Deploy is an open-source collection of C++ samples combining ONNX Runtime with the TensorRT RTX execution provider for local AI inference on Windows and Linux. It covers speech recognition, image segmentation and image generation, with DGX Spark performance measurements included.

An AWS Machine Learning Blog post shows how to implement Amazon S3 Vectors as a custom persistent memory provider for the NVIDIA NeMo Agent Toolkit (NAT), tested with NAT version 1.6. The walkthrough covers creating S3 Vectors infrastructure, writing a MemoryEditor plugin, configuring an agent workflow, and deploying on Amazon EKS.

Simon Willison released pwasm 0.2a0, a pure Python WebAssembly engine, after 42 commits made with Claude Opus 5.5. The PyPI wheel now bundles working WASM builds of MicroPython, QuickJS and Micro QuickJS, though he warns it is alpha software he would not trust.

LangChain built a model router into Open SWE's harness that cut median cost per coding task by 64% with no measurable change in quality, according to an A/B test across 973 threads. The post describes how to build a similar router.

An AWS Machine Learning Blog post details how to build ambient agents using Amazon Bedrock AgentCore, where event streams like Amazon S3 uploads trigger agent jobs that pause for human input via a single ask_human tool. The reference sample combines AgentCore Runtime, Lambda, DynamoDB, SQS, and a React frontend into a serverless platform.

An AWS Machine Learning Blog post details a step-by-step guide for giving three environment types access to Claude Platform on AWS inference: AWS workload accounts via cross-account SigV4, developer laptops via a workspace-scoped API key, and external workloads via OIDC federation, all sharing one subscription with workspace isolation.

NVIDIA detailed an end-to-end HSTU generative recommender inference workflow in its recsys-examples repository, served through NVIDIA Dynamo-Triton with PyTorch AOTI and FlexKV-backed KV caching. On an RTX PRO 6000 Blackwell Workstation Edition GPU, the eight-layer HSTU model reached up to 5.93x lower latency at batch size 8 with a 100% GPU KV-cache hit rate versus the same AOTI configuration without caching.

NVIDIA announced general availability of its cuObject client and server libraries and expanded the xio-sig consortium to include cuObject alongside cuFile, with Google Cloud evaluating participation and Microsoft planning to join the board. A new SCADA Server SDK lets storage providers build servers for GPU-initiated requests, and IBM demonstrated a prototype integrating SCADA with IBM Storage Scale.

InfoQ is running two five-week online certification cohorts in October 2026: AI Security & Privacy Engineering starting October 26 and AI-Assisted Engineering starting October 19. The cohorts cover protecting sensitive data in AI products and verification of coding agents working in existing codebases.

PyTorch Conference North America 2026 in San Jose will feature a series of Ray-focused sessions covering Ray's integration with Kubernetes and production use cases at LinkedIn, Uber, Pinterest, and Anyscale.

NVIDIA published a tutorial on using NeMo Relay execution traces to inspect Hermes Agent behavior, including two runnable examples and a ToolPerf case study comparing baseline and fixed harness revisions across 108 runs.

Torch Spyre reached CRCR L2 integration by testing the PyTorch backend for IBM's Spyre accelerator against PyTorch's test suite, using an agentic test-selection pipeline and a declarative YAML framework that adapts tests without patching upstream. The design targets any generic privateuse1 device, and the authors intend to work with the PyTorch team to upstream reusable parts.