AivexaNewsSearch
AI news for builders and product teamsUpdated Oct 11, 2026, 00:01 UTC

Developer tools news

The latest Developer tools stories across our sources, prepared from the publishers’ own reporting.

In this topic

Newest first

Spaces: A CLI Built for Humans and Agents

Mistral AI published an engineering blog post describing Spaces, its internal CLI for scaffolding projects, running dev environments, and deploying to staging. The post explains how the tool was adapted for use by coding agents, through flags, explicit state, introspectable plugins, and generated context.json and AGENTS.md files.

Read original

OpenClaw

OpenClaw is a personal AI assistant that links messaging platforms to local AI coding agents through a centralized gateway running on the user's own devices. Ollama published installation and launch instructions, including an ollama launch openclaw command and a recommended context length of at least 64k tokens.

Terminally online Mistral Vibe.

Mistral AI released Mistral Vibe 2.0, an upgrade to its terminal-native coding agent, powered by the Devstral 2 model family. It adds custom subagents, multi-choice clarifications, slash-command skills, unified agent modes, and automatic updates, and is available on Le Chat Pro and Team plans.

Read original

ollama launch

Ollama released a new command, ollama launch, that sets up and runs coding tools such as Claude Code, OpenCode and Codex with local or cloud models without environment variables or config files. It requires Ollama v0.15+ and recommends at least 64000 tokens of context length.

Read original

OpenAI Codex with Ollama

Open models can now be used with OpenAI's Codex CLI through Ollama, according to an Ollama post dated January 15, 2026. Codex can read, modify, and execute code in the working directory using models such as gpt-oss:20b, gpt-oss:120b, or other open-weight alternatives.

Read original

Web search

Ollama released a web search API on September 24, 2025, alongside a web fetch API. A free tier is available for individuals, with higher rate limits via Ollama's cloud, plus REST support and Python and JavaScript library integrations.

Read original

New model scheduling

Ollama released a new model scheduling system that measures exact memory requirements before running a model, aiming to cut out-of-memory crashes and improve GPU utilization, including on multi-GPU and mismatched GPU systems. It is enabled by default for models on Ollama's new engine.

Read original

Cloud models

Ollama announced on September 19, 2025 that cloud models are now in preview, allowing users to run larger models on datacenter-grade hardware. The cloud offering integrates with existing local tools and Ollama's OpenAI-compatible API, and Ollama says its cloud does not retain user data.

Read original

Thinking

Ollama added the ability to enable or disable model thinking, separating reasoning from output when enabled. DeepSeek R1 and Qwen 3 support the feature, available through CLI flags, interactive commands, a scripting option, and a new think parameter in the generate and chat APIs.

Read original

Building A Generative AI Platform

Chip Huyen outlines the common architecture of generative AI platforms, starting from a minimal query-to-model setup and progressively adding context construction, RAG retrieval approaches, guardrails for input and output, query rewriting, and agentic actions. The post describes components and considerations, not model evaluation, prompt engineering, finetuning, or RAG chunking strategies.

Read original