Ollama
Get up and running with Llama 3.3, DeepSeek-R1, Phi-4, Gemma 3, Mistral Small 3.1 and other large language models.
Ollama is a tool for running large language models locally on your own hardware. It provides a simple command-line interface and a REST API compatible with the OpenAI API format, making it straightforward to pull, run, and switch between models from a growing library. Models run entirely on your machine, with no data sent to external services.
The platform integrates with thousands of applications and developer tools — including coding assistants, RAG pipelines, document processing workflows, and AI chat interfaces — through its API layer. Custom models can be created and shared using a Modelfile format similar to a Dockerfile, and multi-modal models that handle text and images are supported.
Ollama is aimed at developers, researchers, and organisations who want to run AI inference locally for privacy, cost control, or offline use, and who need a straightforward way to manage and serve multiple models without complex infrastructure setup.
Repository details
Updated 8/12/2026, 2:00:23 PM
View RepositoryRepository activity
Compare Ollama with
Similar open source alternatives
Lobe Chat
Open-source AI chat and agent platform with multi-provider support, RAG, plugins, and an agent marketplace — self-hostable ChatGPT alternative.
Onyx
Onyx is an open-source AI chat and enterprise search platform with RAG that works with every LLM. Bring your own model and unify all your team's knowledge.
Glean
Open WebUI
User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
Pi
Open-source AI agent toolkit with an interactive coding agent CLI, unified multi-provider LLM API, and TUI libraries.
