Ollama Blog
25 stories
Ollama Blog
Ollama BlogToolsOllama's transparent pricing Ollama's Pro, Max, and Team plans now use industry-standard per-token pricing with usage included on every plan.
Ollama Blog
Ollama BlogToolsClaude Desktop support with Ollama Claude Desktop can now be configured to work with Ollama as a third-party gateway provider, making it possible to use open models in Claude.
Ollama Blog
Ollama BlogToolsNVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is now available on Ollama. It's a 30 billion parameter (3B active) open model built for agents that stay running, gathering context, calling tools, and working through multi-step tasks on y
Ollama Blog
Ollama BlogToolsMuse Glimmer from Meta Superintelligence Labs is now available Meta's Muse Glimmer, the first open model released by Meta Superintelligence Labs, is now available. Muse Glimmer is a 30B multimodal model released under the Apache 2.0 license, designed for local coding agents, and acc
Ollama Blog
Ollama BlogToolsOllama: all aboard open models Serving 8.9 million developers, Ollama has raised $88M from Benchmark, Theory Ventures, 8VC, Y Combinator, and many incredible angel investors.
Ollama Blog
Ollama BlogToolsFaster Gemma 4 on MLX with multi-token prediction Gemma 4 is now significantly faster in Ollama 0.31 on Apple Silicon via multi-token prediction (MTP), powered by MLX. Performance is now up to 90% faster when used with coding agents, as measured using the Aider polyglot
Ollama Blog
Ollama BlogToolsOllama's highest performance on Apple Silicon yet with MLX Ollama's MLX engine has been updated to deliver its highest performance on Apple Silicon yet. Models output higher quality responses, respond faster, and use less memory.
Ollama Blog
Ollama BlogToolsImproved performance and model support with GGUF Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp. This augments Ollama's MLX engine on Apple silicon, bringing support to more models on a wider range of hardware.
Ollama Blog
Ollama BlogToolsNVIDIA Nemotron 3 Ultra NVIDIA Nemotron 3 Ultra is built for high-throughput reasoning and long-running agent workflows.
Ollama Blog
Ollama BlogToolsOpenJarvis: a local-first personal AI is now available to run with Ollama OpenJarvis v1.0 is now available: an open-source framework for building personal AI agents that run on your own hardware, with Ollama support built-in.
Ollama Blog
Ollama BlogToolsOllama is now powered by MLX on Apple Silicon in preview Today, we're previewing the fastest way to run Ollama on Apple silicon, powered by MLX, Apple's machine learning framework.
Ollama Blog
Ollama BlogToolsThe simplest and fastest way to setup OpenClaw Setup OpenClaw in under two minutes with a single Ollama command.
Ollama Blog
Ollama BlogToolsSubagents and web search in Claude Code Ollama now supports subagents and web search in Claude Code.
Ollama Blog
Ollama BlogToolsOpenClaw OpenClaw is a personal AI assistant that connects your messaging apps to local AI coding agents, all running on your own device.
Ollama Blog
Ollama BlogToolsollama launch ollama launch is a new command which sets up and runs coding tools like Claude Code, OpenCode, and Codex with local or cloud models. No environment variables or config files needed.
Ollama Blog
Ollama BlogToolsImage generation (experimental) Generate images locally with Ollama on macOS. Windows and Linux support coming soon.
Ollama Blog
Ollama BlogToolsClaude Code with Anthropic API compatibility Ollama is now compatible with the Anthropic Messages API, making it possible to use tools like Claude Code with open models.
Ollama Blog
Ollama BlogToolsOpenAI Codex with Ollama Open models can be used with OpenAI's Codex CLI through Ollama. Codex can read, modify, and execute code in your working directory using models such as gpt-oss:20b, gpt-oss:120b, or other open-weight alternatives.
Ollama Blog
Ollama BlogToolsOpenAI gpt-oss-safeguard Ollama is partnering with OpenAI and ROOST (Robust Open Online Safety Tools) to bring the latest gpt-oss-safeguard reasoning models to users for safety classification tasks. gpt-oss-safeguard models are available in two
Ollama Blog
Ollama BlogToolsMiniMax M2 MiniMax M2 is now available on Ollama's cloud. It's a model built for coding and agentic workflows.
Ollama Blog
Ollama BlogToolsNVIDIA DGX Spark performance We ran performance tests on release day firmware and an updated Ollama version to see how Ollama performs.
Ollama Blog
Ollama BlogToolsNew coding models & integrations GLM-4.6 and Qwen3-coder-480B are available on Ollama’s cloud service with easy integrations to the tools you are familiar with. Qwen3-Coder-30B has been updated for faster, more reliable tool calling in Ollama’s new engi
Ollama Blog
Ollama BlogToolsQwen3-VL Ollama now supports Alibaba's Qwen3-VL.
Ollama Blog
Ollama BlogToolsNVIDIA DGX Spark The latest NVIDIA DGX Spark is here! Ollama has partnered with NVIDIA to ensure it runs fast and efficiently out-of-the-box.
Nothing matches this filter yet.