Tenstorrent
Open RISC-V AI hardware
- Version
- Blackhole
- Cost
- from $999/unit
- Model
- Tensix cores
Cerebras · #3 most active of 18 in Chips & Compute
Single-wafer accelerators delivering the fastest published open-model token rates.
Trust score · Mixed
3 data points checked · 2026-09-12
Current version
WSE-3
Entry cost
Free tier / developer tier pricing available (contact/partner integrations)
Changes / 30d
9
The short answers · verified September 15, 2026
Source: Cerebras — www.cerebras.ai · Reusable under CC BY 4.0 — cite Tomorrow
Developer Tier
Free
Entry price over time
11 of 49 points reconstructed from Internet Archive captures (hollow dots) — dated by capture, not by when we first saw the page. View a capture
Codex-Spark
Cerebras
Gemma
Gemma-4-31B
GLM-4.7
Zhipu
GPT 5.6 Sol
OpenAI
GPT-5.6-Sol-Ultrafast
OpenAI
Kimi K2.6
Moonshot
Llama
Meta
Llama / Qwen / GPT-OSS
open weights
Qwen
Alibaba
WSE-3
Cerebras
Context window
—
Public API
yes
Multi-model routing
yes
Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.
green disclosed · amber inferred · grey unattributed. Aliases are folded into one model; provider concentration counts only models with an established vendor.
Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.
No public usage figure on record for this product.
Tenstorrent
Open RISC-V AI hardware
Inference-first TPU pods
NVIDIA
GB200/B200 rack-scale AI systems
Cloud-native training silicon
Intel
Cost-focused training and inference accelerator.
SambaNova
Dataflow chips for model serving
Products in other segments that name one of Cerebras WSE-3's models in their stack.
Fireworks AI
Low-latency serving for open models and custom fine-tunes.
Moonshot AIChina
Long-context assistant and agentic K-series models.
Ollama
Local model runner that made desktop inference trivial.
AlibabaChina
Open-weight image generation and editing from the Qwen line.
Together AI
Open-model inference and fine-tuning at frontier speed.
New capabilities: On-Premise Deployment, Cloud API
96 tracked98 tracked · +On-Premise Deployment, Cloud API
sourceNew capabilities: OpenAI API compatible, Custom kernels
94 tracked96 tracked · +OpenAI API compatible, Custom kernels
sourceNew capabilities: PyTorch SDK, Low-latency token generation
92 tracked94 tracked · +PyTorch SDK, Low-latency token generation
sourceNew capabilities: Up to 30x faster inference than GPU, Pre-training
90 tracked92 tracked · +Up to 30x faster inference than GPU, Pre-training
sourceNew capabilities: Low latency inference, Dedicated support
88 tracked90 tracked · +Low latency inference, Dedicated support
sourceNew capabilities: Developer Tier Pricing, Enterprise SLA
86 tracked88 tracked · +Developer Tier Pricing, Enterprise SLA
sourceNew capabilities: GPU-free Wafer-Scale Engine, OpenAI-compatible API, Fine-tuning and training services
83 tracked86 tracked · +GPU-free Wafer-Scale Engine, OpenAI-compatible API, Fine-tuning and training services
sourceNew capabilities: Developer SDK, Low-latency queue priority
81 tracked83 tracked · +Developer SDK, Low-latency queue priority
sourceNew capabilities: inference, training, fine-tuning
76 tracked81 tracked · +inference, training, fine-tuning, api-access, on-premise
sourceNew capabilities: Up to 30x faster than GPUs, Dedicated on-premise deployment
73 tracked75 tracked · +Up to 30x faster than GPUs, Dedicated on-premise deployment
sourceNew capabilities: AI Inference API, Model training services
68 tracked70 tracked · +AI Inference API, Model training services
sourceNew capabilities: AI Accelerator, AWS Marketplace
66 tracked68 tracked · +AI Accelerator, AWS Marketplace
sourceNew capabilities: Fastest AI inference, Developer API, AWS Marketplace integration
60 tracked66 tracked · +Fastest AI inference, Developer API, AWS Marketplace integration, OpenRouter integration, HuggingFace integration, Vercel integration
sourceCerebras WSE-3 added Codex-Spark to its model stack
GLM-4.7, GPT 5.6 Sol, GPT-5.6-Sol-Ultrafast, Gemma, Gemma-4-31B, Kimi K2.6, Llama, Llama / Qwen / GPT-OSS, Qwen, WSE-3Codex-Spark, GLM-4.7, GPT 5.6 Sol, GPT-5.6-Sol-Ultrafast, Gemma, Gemma-4-31B, Kimi K2.6, Llama, Llama / Qwen / GPT-OSS, Qwen, WSE-3
sourceNew capabilities: AWS Marketplace Integration, OpenRouter Integration, HuggingFace Integration
55 tracked59 tracked · +AWS Marketplace Integration, OpenRouter Integration, HuggingFace Integration, Vercel Integration
sourceNew capabilities: AI Hardware Accelerator, Inference Cloud API, World-record speeds
52 tracked55 tracked · +AI Hardware Accelerator, Inference Cloud API, World-record speeds
sourceNew capabilities: AI Training, PyTorch Support
50 tracked52 tracked · +AI Training, PyTorch Support
sourceCerebras WSE-3 added GPT 5.6 Sol, GPT-5.6-Sol-Ultrafast to its model stack
GLM-4.7, Gemma, Gemma-4-31B, Kimi K2.6, Llama, Llama / Qwen / GPT-OSS, Qwen, WSE-3GLM-4.7, GPT 5.6 Sol, GPT-5.6-Sol-Ultrafast, Gemma, Gemma-4-31B, Kimi K2.6, Llama, Llama / Qwen / GPT-OSS, Qwen, WSE-3
sourceNew capabilities: Model training, SDK, Developer support
47 tracked50 tracked · +Model training, SDK, Developer support
sourceNew capabilities: Wafer-scale hardware acceleration, LLM Inference API
45 tracked47 tracked · +Wafer-scale hardware acceleration, LLM Inference API
sourceNew capabilities: AI Hardware / Chip, AI Training & Fine-tuning
43 tracked45 tracked · +AI Hardware / Chip, AI Training & Fine-tuning
sourceNew capabilities: Training and fine-tuning services, Dedicated queue priority, Custom model weights
40 tracked43 tracked · +Training and fine-tuning services, Dedicated queue priority, Custom model weights
sourceNew capabilities: AI Inference, Model Fine-tuning, Model Training
34 tracked38 tracked · +AI Inference, Model Fine-tuning, Model Training, OpenAI API Compatibility
sourceNew capabilities: AI training, AI inference, Low-latency programming
31 tracked34 tracked · +AI training, AI inference, Low-latency programming
sourceStraight head-to-head pages against the busiest products in Chips & Compute.
Free to cite and reuse under CC BY 4.0. Permalink: https://tomorrow.aliensquad.ai/tools/cerebras
Tomorrow. (2026). Cerebras WSE-3 — version, pricing and model stack [Data set entry]. AlienSquad. Retrieved 2026-09-17, from https://tomorrow.aliensquad.ai/tools/cerebras
@misc{tomorrow-tools-cerebras,
author = {{Tomorrow}},
title = {Cerebras WSE-3 — version, pricing and model stack},
year = {2026},
publisher = {AlienSquad},
howpublished = {\url{https://tomorrow.aliensquad.ai/tools/cerebras}},
note = {Accessed: 2026-09-17}
}