Ollama

The most popular way to run open models locally on your own machine.

Visit website

Free, open-source runtime for running open LLMs locally on macOS, Windows, and Linux — CLI, OpenAI-compatible API, GGUF import. Local runs unlimited.

Ollama homepage — run open models locally

Ollama homepage — run open models locally

Overview

What is Ollama?

The most popular way to run open models locally on your own machine.

Ollama is a free, open-source runtime for running open large language models on your own machine (macOS, Windows, Linux). It bundles a command-line interface, a REST API compatible with OpenAI and Anthropic endpoints, and a model library you pull with a single command — plus GGUF import via Modelfiles.

Local runs are always free and unlimited on any plan; prompts and data never leave your machine. Ollama Cloud offers hosted inference for models too large to run locally, with paid plans (Pro $20/month, Team $500/month) covering cloud usage credits.

Platforms and languages

Ollama Availability

Platforms

WindowsMacosLinux

Languages

English

Capabilities

Ollama Key Features

One-command local models

Pull and run open models locally with `ollama pull` and `ollama run`.

GGUF import

Bring your own quantized models with a Modelfile — FROM a .gguf file, then `ollama create`.

Compatible APIs

OpenAI-compatible /v1/chat/completions and Anthropic-compatible /v1/messages endpoints for drop-in integration.

Privacy-first

Local runs never send prompts or data to ollama.com; cloud prompts are never stored or trained on.

Ollama Cloud

Run larger models on Ollama's servers via :cloud model names when local hardware isn't enough.

Best for

Who uses Ollama?

Developers and privacy-minded users

Run open LLMs locally with a CLI and OpenAI-compatible API

Plans and access

Ollama Pricing

Freemium

Local runs always free and unlimited. Pro $20/mo, Team $500/mo, Enterprise custom. Cloud credits extra.

Free trial available

Common questions

Ollama FAQs

Does Ollama send my prompts back to ollama.com?

No for local runs. Cloud-hosted prompts are processed to provide the service but never stored, logged, or trained on.

Where are models stored?

macOS ~/.ollama/models, Linux /usr/share/ollama/.ollama/models, Windows C:\Users\%username%\.ollama\models — changeable via the OLLAMA_MODELS environment variable.

Are local runs limited by plan?

No — running models on your own hardware is always unlimited; paid plans only cover cloud usage credits.

How do I import a GGUF model?

Write a Modelfile with a FROM line pointing at the .gguf file, then `ollama create my-model` and `ollama run my-model`.

Reviews

See what the community thinks and share your experience.

—

Based on 0 ratings

Rating distribution

0
0
0
0
0

Leave a review

Sign in to rate

Community reviews

No reviews

No written reviews yet.

Explore by category

AI Development Infrastructure

View all AI Development Infrastructure websites
vLLM preview
vLLM

Open-source (Apache 2.0) engine for serving LLMs — PagedAttention, continuous batching, OpenAI-compatible API on your own GPUs. Free.

AI Development Infrastructure

Explore by category

AI Development Platforms

View all AI Development Platforms websites
Lloyal preview
Lloyal

An open-source TypeScript platform for turning open-weight models into downloadable AI apps with built-in inference and a multi-agent runtime.

AI Development Platforms
Moative preview
Moative

AI Transformation, Consulting, Forward-deployed engineers for mid-markets in regulated, asset-heavy industries. Skin-in-the-game models. Build with us. Roll out

AI Development Platforms

Explore by category

Llms Foundation Models

View all Llms Foundation Models websites
Hugging Face preview
Hugging Face

The Hub for open machine learning: 2M+ models, 1.5M datasets and Spaces apps, plus inference APIs. Free for public use; paid plans from $9/month.

Llms Foundation Models
vLLM preview
vLLM

Open-source (Apache 2.0) engine for serving LLMs — PagedAttention, continuous batching, OpenAI-compatible API on your own GPUs. Free.

Llms Foundation Models
LM Studio preview
LM Studio

Desktop app to run open LLMs locally on macOS, Windows, and Linux — chat UI, OpenAI-compatible local server, offline transcription. Free plan.

Llms Foundation Models
Clef preview
Clef

Cloudflare's open-source 27B multimodal decision model: state plus typed questions in, probabilities for every option out.

Llms Foundation Models