Ollama logo
AI Models

Ollama

Run Llama and other large language models locally

Visit Official Website
PRODUCT PREVIEW

A look at Ollama

Open source page
Screenshot of Ollama on its official website
Captured from the official website · 2026-10-05Product pages can change over time.
OVERVIEW

About Ollama

Ollama is a command line tool for running large language models on your local computer. It allows users to download and run models like Llama 2, Code Llama and others locally, and supports customization and creation of their own models. This free and open source project currently supports macOS and Linux operating systems, and will also support Windows systems in the future.

In addition, Ollama also provides an official Docker image, making it easier to deploy large language models using Docker containers, ensuring that all interactions with these models occur locally without the need to send private data to third-party services. Ollama supports GPU acceleration on macOS and Linux and provides a simple command line interface (CLI) as well as a REST API for interacting with applications.

This tool is particularly useful for developers or researchers who need to run and experiment with large language models on their local machine, without relying on external cloud services.

Ollama has now launched the Ollama desktop version for macOS and Windows, with new file processing and multi-modal interaction functions, allowing users to interact with local large language models more intuitively and conveniently.

Obtaining the Ollama installation package

Obtain the Ollama installation package, scan the QR code and follow the reply: Ollama

Ollama supported models

Ollma provides a model library. Users can choose to install the model they want to run. Currently, it supports 40+ models and is still increasing. The following are examples of open source models that can be downloaded:

model Parameter size File size Download and run commands
DeepSeek-R1 1.5B, 7B, 14B, 32B, etc. 12-320GB ollama run deepseek-r1
Neural Chat 7B 4.1GB ollama run neural-chat
Starling 7B 4.1GB ollama run starling-lm
Mistral 7B 4.1GB ollama run mistral
Llama 2 7B 3.8GB ollama run llama2
Code Llama 7B 3.8GB ollama run codellama
Llama 2 Uncensored 7B 3.8GB ollama run llama2-uncensored
Llama 2 13B 13B 7.3GB ollama run llama2:13b
Llama 2 70B 70B 39GB ollama run llama2:70b
Orca Mini 3B 1.9GB ollama run orca-mini
Vicuna 7B 3.8GB ollama run vicuna