Baseten
Production inference platform with dedicated deployments.
- Version
- 2025
- Cost
- Usage-based, enterprise tiers
- Model
- open weights
Model registries, inference platforms, vector stores and the eval/observability layer everything else is built on.
17 tools · 0 changes / 30d
Baseten
Production inference platform with dedicated deployments.
Braintrust
Eval-first platform for shipping LLM features safely.
Chroma
Embedded-first vector store popular in prototypes and agents.
Fireworks AI
Low-latency serving for open models and custom fine-tunes.
Hugging Face
The model and dataset registry the open ecosystem runs on.
LangChain
Tracing, evals and monitoring for LLM applications.
LM Studio
Desktop app for running and serving local models.
Modal Labs
Serverless GPU compute for Python-native AI workloads.
Ollama
Local model runner that made desktop inference trivial.
OpenRouter
One API in front of 400+ models with live price and latency data.
Pinecone
Managed vector database built for retrieval at scale.
Replicate
Run and ship any open model behind a simple API.
Scale AI
Data engine and evaluation for frontier model builders.
Together AI
Open-model inference and fine-tuning at frontier speed.
vLLM (PyTorch Foundation)
The de facto open-source inference engine for LLM serving.
Weaviate
Open-source vector database with built-in hybrid search.
CoreWeave
Experiment tracking and Weave evals for AI teams.
No changes recorded yet.