Infrastructure & Serving

Ollama

Ollama · #9 most active of 17 in Infrastructure & Serving

Compare →

Local model runner that made desktop inference trivial.

Current version

0.x

Entry cost

Free

Changes / 30d

0

The short answers

What is the latest version of Ollama?
The current shipped version of Ollama is 0.x.
How much does Ollama cost?
Ollama has no published flat entry price. Free
What AI model does Ollama use?
Ollama runs primarily on Gemma.

Source: Ollama (no public source page recorded) · Reusable under CC BY 4.0 — cite Tomorrow

Capabilities

  • Local models
  • OpenAI-compatible API
  • GGUF

Pricing

Free

Entry price over time

Not enough pricing history yet — we start charting from the second observation.

Geek mode

Models under the hood

  • Gemma

    Google

    disclosed
  • Llama

    Meta

    disclosed
  • Qwen

    Alibaba

    disclosed

Context window

Public API

no

Multi-model routing

no

Stack signals

  • llama.cpp
  • Go

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

green disclosed · amber inferred · grey unattributed. Aliases are folded into one model; provider concentration counts only models with an established vendor.

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Infrastructure & Serving

All Infrastructure & Serving
Baseten

Baseten

Production inference platform with dedicated deployments.

Version
2025
Cost
Usage-based, enterprise tiers
Model
open weights
Braintrust

Braintrust

Eval-first platform for shipping LLM features safely.

Version
2025
Cost
Free tier, usage-based
Model
model-agnostic
Chroma

Chroma

Embedded-first vector store popular in prototypes and agents.

Version
1.x
Cost
Open source; cloud usage-based
Model
embeddings-agnostic
Fireworks AI

Fireworks AI

Low-latency serving for open models and custom fine-tunes.

Version
2025
Cost
Per-token usage
Model
Llama
Hugging Face

Hugging Face

The model and dataset registry the open ecosystem runs on.

Version
Hub
Cost
Free tier · from $9/mo
Model
hosts most open weights
LangSmith

LangChain

Tracing, evals and monitoring for LLM applications.

Version
2025
Cost
Free tier · from $39/mo
Model
model-agnostic

Elsewhere on the same models

Products in other segments that name one of Ollama's models in their stack.

Qwen-Image

AlibabaChina

1

Open-weight image generation and editing from the Qwen line.

Version
Qwen-Image
Cost
Pricing not specified in the provided documentation.
Model
Qwen

Change history

No changes recorded yet.

Subscribe to Ollama changes

Ollama compared

Straight head-to-head pages against the busiest products in Infrastructure & Serving.

How to cite this page

Free to cite and reuse under CC BY 4.0. Permalink: https://tomorrow.aliensquad.ai/tools/ollama

APA
Tomorrow. (2026). Ollama — version, pricing and model stack [Data set entry]. AlienSquad. Retrieved 2026-09-17, from https://tomorrow.aliensquad.ai/tools/ollama
BibTeX
@misc{tomorrow-tools-ollama,
  author       = {{Tomorrow}},
  title        = {Ollama — version, pricing and model stack},
  year         = {2026},
  publisher    = {AlienSquad},
  howpublished = {\url{https://tomorrow.aliensquad.ai/tools/ollama}},
  note         = {Accessed: 2026-09-17}
}