/

Ollama Blog

25 stories

Ollama Blog

Ollama BlogToolsOllama's transparent pricing Ollama's Pro, Max, and Team plans now use industry-standard per-token pricing with usage included on every plan.

August 31
Ollama Blog

Ollama BlogToolsClaude Desktop support with Ollama Claude Desktop can now be configured to work with Ollama as a third-party gateway provider, making it possible to use open models in Claude.

August 25
Ollama Blog

Ollama BlogToolsNVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is now available on Ollama. It's a 30 billion parameter (3B active) open model built for agents that stay running, gathering context, calling tools, and working through multi-step tasks on y

August 11
Ollama Blog

Ollama BlogToolsMuse Glimmer from Meta Superintelligence Labs is now available Meta's Muse Glimmer, the first open model released by Meta Superintelligence Labs, is now available. Muse Glimmer is a 30B multimodal model released under the Apache 2.0 license, designed for local coding agents, and acc

August 10
Ollama Blog

Ollama BlogToolsOllama: all aboard open models Serving 8.9 million developers, Ollama has raised $88M from Benchmark, Theory Ventures, 8VC, Y Combinator, and many incredible angel investors.

July 9
Ollama Blog

Ollama BlogToolsFaster Gemma 4 on MLX with multi-token prediction Gemma 4 is now significantly faster in Ollama 0.31 on Apple Silicon via multi-token prediction (MTP), powered by MLX. Performance is now up to 90% faster when used with coding agents, as measured using the Aider polyglot

June 29
Ollama Blog

Ollama BlogToolsOllama's highest performance on Apple Silicon yet with MLX Ollama's MLX engine has been updated to deliver its highest performance on Apple Silicon yet. Models output higher quality responses, respond faster, and use less memory.

June 11
Ollama Blog

Ollama BlogToolsImproved performance and model support with GGUF Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp. This augments Ollama's MLX engine on Apple silicon, bringing support to more models on a wider range of hardware.

June 5
Ollama Blog

Ollama BlogToolsNVIDIA Nemotron 3 Ultra NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.

June 4
Ollama Blog

Ollama BlogToolsOpenJarvis: a local-first personal AI is now available to run with Ollama OpenJarvis v1.0 is now available: an open-source framework for building personal AI agents that run on your own hardware, with Ollama support built-in.

May 28
Ollama Blog

Ollama BlogToolsOllama is now powered by MLX on Apple Silicon in preview Today, we're previewing the fastest way to run Ollama on Apple silicon, powered by MLX, Apple's machine learning framework.

March 30
Ollama Blog

Ollama BlogToolsThe simplest and fastest way to setup OpenClaw Setup OpenClaw in under two minutes with a single Ollama command.

February 23
Ollama Blog

Ollama BlogToolsSubagents and web search in Claude Code Ollama now supports subagents and web search in Claude Code.

February 16
Ollama Blog

Ollama BlogToolsOpenClaw OpenClaw is a personal AI assistant that connects your messaging apps to local AI coding agents, all running on your own device.

February 1
Ollama Blog

Ollama BlogToolsollama launch ollama launch is a new command which sets up and runs coding tools like Claude Code, OpenCode, and Codex with local or cloud models. No environment variables or config files needed.

January 23
Ollama Blog

Ollama BlogToolsImage generation (experimental) Generate images locally with Ollama on macOS. Windows and Linux support coming soon.

January 20
Ollama Blog

Ollama BlogToolsClaude Code with Anthropic API compatibility Ollama is now compatible with the Anthropic Messages API, making it possible to use tools like Claude Code with open models.

January 16
Ollama Blog

Ollama BlogToolsOpenAI Codex with Ollama Open models can be used with OpenAI's Codex CLI through Ollama. Codex can read, modify, and execute code in your working directory using models such as gpt-oss:20b, gpt-oss:120b, or other open-weight alternatives.

January 15
Ollama Blog

Ollama BlogToolsOpenAI gpt-oss-safeguard Ollama is partnering with OpenAI and ROOST (Robust Open Online Safety Tools) to bring the latest gpt-oss-safeguard reasoning models to users for safety classification tasks. gpt-oss-safeguard models are available in two

October 29
Ollama Blog

Ollama BlogToolsMiniMax M2 MiniMax M2 is now available on Ollama's cloud. It's a model built for coding and agentic workflows.

October 28
Ollama Blog

Ollama BlogToolsNVIDIA DGX Spark performance We ran performance tests on release day firmware and an updated Ollama version to see how Ollama performs.

October 23
Ollama Blog

Ollama BlogToolsNew coding models & integrations GLM-4.6 and Qwen3-coder-480B are available on Ollama’s cloud service with easy integrations to the tools you are familiar with. Qwen3-Coder-30B has been updated for faster, more reliable tool calling in Ollama’s new engi

October 16
Ollama Blog

Ollama BlogToolsQwen3-VL Ollama now supports Alibaba's Qwen3-VL.

October 14
Ollama Blog

Ollama BlogToolsNVIDIA DGX Spark The latest NVIDIA DGX Spark is here! Ollama has partnered with NVIDIA to ensure it runs fast and efficiently out-of-the-box.

October 13