Baseten
Production inference platform with dedicated deployments.
- Version
- 2025
- Cost
- Usage-based, enterprise tiers
- Model
- open weights
Replicate · #12 most active of 17 in Infrastructure & Serving
Run and ship any open model behind a simple API.
Current version
2025
Entry cost
Per-second GPU billing
Changes / 30d
0
Per-second GPU billing
Entry price over time
Not enough pricing history yet — we start charting from the second observation.
FLUX
SDXL
Llama
Context window
—
Public API
no
Multi-model routing
no
Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.
green disclosed · amber inferred · grey unknown
Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.
No public usage figure on record for this product.
Baseten
Production inference platform with dedicated deployments.
Braintrust
Eval-first platform for shipping LLM features safely.
Chroma
Embedded-first vector store popular in prototypes and agents.
Fireworks AI
Low-latency serving for open models and custom fine-tunes.
Hugging Face
The model and dataset registry the open ecosystem runs on.
LangChain
Tracing, evals and monitoring for LLM applications.
Products in other segments that name one of Replicate's models in their stack.
Databricks
Ask your lakehouse questions
Freepik
Stock library fused with multi-model generation.
Oracle
Embedded AI agents across Fusion ERP, HCM and supply chain.
No changes recorded yet.