Ollama
Ollama is an open-source program that downloads open AI models and runs them on your own computer, with a simple command and a local API.
Ollama is an open-source program for running open AI models, the kind anyone can download, on your own computer. You can install it on Linux, macOS or Windows. Its code is under the MIT License, so anyone may use and change it.
You type one command, such as ollama run with a model’s name. Ollama fetches that model from its online library and saves it in a folder on your machine. Then it loads the model into memory, on the main processor, the graphics card or a mix of both, and you can chat in the terminal, the text window where you type commands. A loaded model stays in memory for five minutes by default.
Behind the scenes, Ollama runs a small local server. It listens only on your own computer, at port 11434, a numbered doorway other programs knock on. Apps can send questions there through Ollama’s REST API, a standard way for programs to talk over the web. They can also use routes that copy part of the OpenAI API. Local requests need no key, and Ollama says it does not see your prompts when you run locally. It also offers optional cloud models, which you can switch off.
Trying an open model on your own computer has three snags.
Follow one question from your keyboard to a local reply.
- 1 · pickYou choose a model from the Ollama library by name, for example with ollama run.
- 2 · pullOllama downloads the model files and keeps them in a folder on your computer.
- 3 · loadThe Ollama server loads the model into memory on the processor, the graphics card or both.
- 4 · answerYour terminal or any app sends a question to the local API and gets the reply back.
Once a model is downloaded, local requests need no API key and your prompts stay on your machine.
| Who | What they ask | What it works with |
|---|---|---|
| Student | “Can I chat with an open model on my laptop without paying for an API?” | A model pulled from the Ollama library and run in the terminal |
| App developer | “Can my code talk to a local model the same way it talks to OpenAI?” | The OpenAI-compatible routes on localhost port 11434 |
| Hobbyist | “Can I make a chatbot with its own personality and settings?” | A Modelfile with a SYSTEM message and parameters |
| Privacy-minded team | “Can we keep our prompts off outside servers?” | Local-only mode with cloud features turned off |
- It gets an open model running on your computer with one command.
- It stores downloaded models locally so they are ready next time.
- It offers a REST API and OpenAI-compatible routes for other apps.
- It lets you package a custom model with a Modelfile.
- Big models need lots of graphics memory; with too little, replies can be slower.
- Cloud models send your prompts to Ollama's servers, so they are not fully local.
Sources used
This explainer is written in original language. The links below support its factual claims.
- repoollama/ollama: Start building with open models, Ollama · read 28 Sept 2026
- repoOllama LICENSE, Ollama · read 28 Sept 2026
- docsAPI introduction, Ollama · read 28 Sept 2026
- docsOpenAI compatibility, Ollama · read 28 Sept 2026
- docsModelfile Reference, Ollama · read 28 Sept 2026
- docsFAQ, Ollama · read 28 Sept 2026
- docsQuickstart, Ollama · read 28 Sept 2026