AI Linkbase
Ollama logo
Freemiumcoding

Ollama

Open models locally and in the cloud through one runtime and API

Open-model runtime for local and cloud inference with CLI, API, desktop apps and individual, team and enterprise plans.

4.3
Try Ollama Free →

Opens official website

local AIOllamaopen modelsCLIAPIcloud inferenceagentsself hosted

Live change feed

Recent verified changes

Shared with the homepage and tool pages

ollama

Ollama 0.32.15 adds model metadata caching

The release adds a model metadata cache designed to reduce per-request overhead for local model serving.

Verify official source →
github-copilot

GitHub Copilot for JetBrains adds persistent memory and Ollama support

GitHub is rolling out persistent memory, access to local models through Ollama, additional enterprise controls, improved chat workflows, and MCP reliability fixes in its JetBrains integration.

Verify official source →
ollama

Ollama 0.32.5 fixes MLX output quality for NVFP4 models

Ollama 0.32.5 fixes an MLX Metal issue that could reduce output quality for NVFP4 models, with the release notes specifically highlighting Laguna.

Verify official source →
ollama

Ollama 0.32.4 adds Apple MLX support and inference fixes

Ollama 0.32.4 adds Laguna support through the MLX engine for Apple GPUs and includes fixes for quantized-model decoding and speculative decoding.

Verify official source →

About Ollama

Ollama runs open models locally and through optional Ollama Cloud. Local models remain unlimited on a user's own hardware. Free includes starter cloud credits; Pro is $20 monthly or $200 annually with $60 monthly usage credits, larger pro models and multiple concurrent requests. Max is $100 monthly with $300 monthly usage credits, early access and ten concurrent requests. Team is Early Access at $500 monthly for unlimited users with $1,000 shared monthly credits. Enterprise is custom.

✅ Pros

  • +Unlimited local models on owned hardware
  • +Simple CLI and API
  • +Free cloud evaluation
  • +Individual and team plans
  • +No training on prompt or response data

❌ Cons

  • −Cloud limits and token rates are model-dependent
  • −Extra usage can add variable cost
  • −Team is Early Access
  • −Local speed depends on hardware
  • −Model licenses still require review

Pricing Plans

Free$0

Run models locally, receive starter cloud usage credits and add credits to unlock all cloud models. No service fee.

Pro Monthly$20 / month

$60 of monthly usage credits, access to larger pro models and multiple concurrent cloud requests.

Pro Annual$200 / year

Billed annually; includes the Pro cloud capabilities and monthly usage credits.

Max$100 / month

$300 of monthly usage credits, early access to newest models and ten concurrent requests.

Team Early Access$500 / month

Unlimited users, $1,000 shared monthly usage credits, centralized administration and priority support.

EnterpriseCustom

Volume pricing, custom security review, support and cost controls.

Prices are public list prices before tax and may vary by country, billing term, and workspace. Last verified 2026-09-01. Official pricing source →

✅ Best For

DevelopersCoding agentsLocal API servingOpen-model experimentationTeams needing shared cloud usage

⚠️ Not Recommended For

Nontechnical users wanting a guided document workspaceSmall teams needing a low per-seat planUsers expecting fixed token quotas across every model

Quick Info

Pricing
freemium
Category
coding
Rating
4.3 / 5.0
Last Reviewed
2026-09-01T00:00:00+00:00

💰 Pricing Breakdown

Official pricing verified 2026-09-01. Free $0; Pro $20/month or $200/year with $60 monthly usage credits; Max $100/month with $300 monthly usage credits; Team Early Access $500/month for unlimited users with $1,000 shared monthly credits; Enterprise custom. Local models on your hardware are unlimited. Extra usage can be added and model token rates vary.

View current pricing →
More coding tools

Compare before you choose

Related buying guides

Use these comparisons and guides to decide whether Ollama fits your workflow, budget, and output needs.