Groq LPU
SRAM-based language processing units sold as a token-priced inference cloud.
Full Groq LPU record →Head to head
Groq LPU (Groq) and AWS Trainium2 (AWS) both sit in Chips & Compute. AWS Trainium2 shipped more tracked changes in the last 30 days (5 vs 4). Every value below comes from the latest crawl of the vendors' own pages.
| Attribute | Groq LPU Groq | AWS Trainium2 AWS |
|---|---|---|
| Segment | Chips & Compute | Chips & Compute |
| Version | LPU | Trainium2 |
| Entry cost | Pricing not published. Contact sales or check account console for usage-based rates. | Pricing not published; available via AWS EC2 instance pricing. |
| Pricing tiers | GroqCloud Console Free Tier / Pay-As-You-Go free | — |
| Model stack | canopylabs/orpheus-arabic-saudi · canopylabs/orpheus-v1-english · kimi-k2-instruct-0905 · Llama / Kimi / GPT-OSS · Llama 3.1 8B Instant · Llama 3.3 70B Versatile · LPU v2 · meta-llama/llama-4-maverick-17b-128e-instruct · meta-llama/llama-4-scout-17b-16e-instruct · minimax-m2.5 · minimax/minimax-m2.5 · minimaxai/minimax-m2.5 · moonshotai/kimi-k2-instruct-0905 · openai/gpt-oss-120b · openai/gpt-oss-20b · openai/gpt-oss-safeguard-20b · Orpheus English · orpheus-arabic-saudi · orpheus-v1-english · Qwen 3.6 27B · qwen/qwen3-vl-32b-instruct · qwen3-vl-32b-instruct · Whisper V3 Large · whisper-large-v3 | Neuron compiler · Trainium2 |
| Context window | 131k | — |
| Public API | ||
| Routes models | ||
| Changes / 30d | 4 | 5 |
| Origin | Global | Global |
| Capabilities | deterministic latency, OpenAI-compatible API, open-model catalogue, batch API, Fast LLM inference, OpenAI Compatibility, Prompt Caching, Speech to Text | Trn2 UltraServer, 64-chip NeuronLink domain, Neuron SDK, Rainier cluster, Generative AI training, Generative AI inference, Up to 20.8 FP8 petaflops, 1.5 TB HBM3 memory |
SRAM-based language processing units sold as a token-priced inference cloud.
Full Groq LPU record →Trn2 UltraServers and Project Rainier capacity, programmed through the Neuron SDK.
Full AWS Trainium2 record →