Local by default
Weights and inference remain on your own hardware.

Ollama manages model downloads and exposes a local HTTP service. Your scripts, notebooks, and editors all talk to the same endpoint.
What it is
Ollama manages model downloads and exposes a local HTTP service. Your scripts, notebooks, and editors all talk to the same endpoint.
Weights and inference remain on your own hardware.
CLI, scripts, and editors share the service on port 11434.
OpenAI-style clients can switch to a local base URL.
Capabilities
Model management, chat, embeddings, and tool calling work out of the box.
Pull and cache open models such as Qwen, Gemma, and DeepSeek.
Generate text, chat, create embeddings, and manage models over HTTP.
Typed helpers make streaming chat and embeddings a few lines of code.
How to use
The first run downloads model weights; later runs reuse the local copy.
Install the desktop service for your platform.
curl -fsSL https://ollama.com/install.sh | shDownload a model once and begin an offline conversation.
ollama run gemma4
# list locally cached models
ollama listKeep exploring
BioinformaticsPython tools for computational biology — read sequence files, query NCBI, parse protein structures, and build phylogenetic trees.
Scientific computingThe foundation of scientific computing in Python — N-dimensional arrays, broadcasting, and fast vectorized math.
Data appsTurn a normal Python script into an interactive web app in minutes — dashboards, reports, and chat interfaces without a frontend build.