AivexaNewsSearch
AI news for builders and product teamsUpdated Oct 10, 2026, 19:01 UTC

Ollama

First-party releases and research from Ollama. Headlines and excerpts link to the original articles.

Latest stories

Newest first
OllamaFirst partyIndustry

Ollama's transparent pricing

Ollama announced that its Pro, Max, and Team plans now use per-token pricing with monthly included usage credits, and introduced its Team plan at $500/month introductory pricing. Existing subscribers keep their current plans and can upgrade in account settings.

Read original

Claude Desktop support with Ollama

Ollama announced that Claude Desktop can now be configured to work with Ollama as a third-party gateway provider, allowing developers to run open models inside Claude. The integration supports both local and cloud models, with telemetry disabled by default and a Zero Data Retention policy.

Read original

NVIDIA Nemotron 3.5 Lightning

NVIDIA Nemotron 3.5 Lightning, a 30 billion parameter open model with 3B active parameters, is now available on Ollama and runs entirely on local devices. It is designed for long-running agentic tasks such as tool calling and multi-step work.

Read original

NVIDIA Nemotron 3 Ultra

NVIDIA Nemotron 3 Ultra, a 550 billion parameter open model with 55B active, is now available on Ollama's cloud for long-running agentic workflows. It offers 1M token context and is optimized for NVIDIA's NVFP4 format, with Ollama citing up to 30% cost savings versus other leading open models.

OpenClaw

OpenClaw is a personal AI assistant that links messaging platforms to local AI coding agents through a centralized gateway running on the user's own devices. Ollama published installation and launch instructions, including an ollama launch openclaw command and a recommended context length of at least 64k tokens.

ollama launch

Ollama released a new command, ollama launch, that sets up and runs coding tools such as Claude Code, OpenCode and Codex with local or cloud models without environment variables or config files. It requires Ollama v0.15+ and recommends at least 64000 tokens of context length.

Read original

OpenAI Codex with Ollama

Open models can now be used with OpenAI's Codex CLI through Ollama, according to an Ollama post dated January 15, 2026. Codex can read, modify, and execute code in the working directory using models such as gpt-oss:20b, gpt-oss:120b, or other open-weight alternatives.

Read original