Cerebras
Wafer-scale inference and training
- Version
- WSE-3 / Inference
- Cost
- Free tier, then from $0.06 / M tokens
- Model
- Llama / Qwen / GPT-OSS
Qualcomm · #16 most active of 18 in Chips & Compute
Rack-scale inference accelerators entering the datacentre.
Current version
AI200
Entry cost
System sales
Changes / 30d
0
System sales
Entry price over time
Not enough pricing history yet — we start charting from the second observation.
n/a
Context window
—
Public API
no
Multi-model routing
no
Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.
green disclosed · amber inferred · grey unknown
Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.
No public usage figure on record for this product.
Cerebras
Wafer-scale inference and training
Inference-first TPU pods
Groq
Deterministic low-latency inference
High-memory GPU alternative
On-device inference across Mac, iPhone and iPad.
Cloud-native training silicon
Products in other segments that name one of Qualcomm AI200's models in their stack.
Fivetran
Managed pipelines feeding the AI data layer.
Vertical AerospaceEU
UK eVTOL programme in piloted flight testing.
No changes recorded yet.