Ollama's transparent pricing
Ollama's Pro, Max, and Team plans now use industry-standard per-token pricing with usage included on every plan.
Source profile
AI News Hub tracks Ollama Blog AI updates with visible source status, reuse boundaries, collection method, and published articles.
Official local AI model runtime blog; confirm reuse terms before full body display.
Ollama's Pro, Max, and Team plans now use industry-standard per-token pricing with usage included on every plan.
Claude Desktop can now be configured to work with Ollama as a third-party gateway provider, making it possible to use open models in Claude.
NVIDIA Nemotron 3.5 Lightning is now available on Ollama. It's a 30 billion parameter (3B active) open model built for agents that stay running, gathering context, calling tools, and working through multi-step tasks on your own hardware.
Meta's Muse Glimmer, the first open model released by Meta Superintelligence Labs, is now available. Muse Glimmer is a 30B multimodal model released under the Apache 2.0 license, designed for local coding agents, and accelerated by Ollama's MLX engine with new native DFlash and image input support.
Serving 8.9 million developers, Ollama has raised $88M from Benchmark, Theory Ventures, 8VC, Y Combinator, and many incredible angel investors.
Gemma 4 is now significantly faster in Ollama 0.31 on Apple Silicon via multi-token prediction (MTP), powered by MLX. Performance improves up to 90% on coding-agent benchmarks.
Ollama's MLX engine has been updated to deliver its highest performance on Apple Silicon yet. By leaning more heavily on Apple's unified memory and the Metal-backed MLX framework, models output higher quality responses, respond faster, and use less memory. The update includes support for NVFP4 format, up to 20% faster output, and a snapshot system for agent workflows.
Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp, augmenting MLX on Apple silicon and supporting more models on wider hardware.
NVIDIA Nemotron 3 Ultra is a 550 billion parameter (55B active) open model designed for long-running agentic workflows, with 1M token context and NVFP4 optimization, leading in agentic benchmarks and cost efficiency.
OpenJarvis v1.0 is now available: an open-source framework for building personal AI agents that run on your own hardware, with Ollama support built-in.
Ollama announces a preview release powered by Apple's MLX framework, delivering significant performance improvements on Apple Silicon, including NVFP4 support and enhanced caching.
Setup OpenClaw in under two minutes with a single Ollama command.
Ollama now supports subagents and web search in Claude Code. No MCP servers or API keys required. Subagents can run tasks in parallel, keeping context clean. Web search is built-in via the Anthropic compatibility layer.
OpenClaw is a personal AI assistant that connects your messaging apps to local AI coding agents, all running on your own device for privacy.
Ollama introduces `ollama launch`, a new command that sets up and runs coding tools like Claude Code, OpenCode, and Codex with local or cloud models, without needing environment variables or config files.
Ollama v0.14.0+ now supports the Anthropic Messages API, enabling tools like Claude Code to work with open-source models. Run locally or connect to cloud models via ollama.com.
Open models can be used with OpenAI's Codex CLI through Ollama. Codex can read, modify, and execute code in your working directory using models such as gpt-oss:20b, gpt-oss:120b, or other open-weight alternatives.
Ollama partners with OpenAI and ROOST to launch the gpt-oss-safeguard reasoning models for safety classification. Available in 20B and 120B sizes under Apache 2.0 license, these models support custom policies, interpretable reasoning, and configurable effort.
MiniMax M2 is now available on Ollama's cloud. It is a model built for coding and agentic workflows, with 10 billion activated parameters (230B total). It ranks #1 among open-source models in composite intelligence benchmarks and excels at multi-file edits, agentic tool use, and long-horizon task execution.