Chips & Compute

Groq LPU

Groq · #3 most active of 18 in Chips & Compute

Compare →

SRAM-based language processing units sold as a token-priced inference cloud.

Current version

GroqCloud

Entry cost

Free tier, then from $0.05 / M tokens

Changes / 30d

1

Capabilities

  • deterministic latency
  • OpenAI-compatible API
  • open-model catalogue
  • batch API

Pricing

  • Free

    Free

  • Developer

    per M input tokens

    $0.05

  • Enterprise

    contract

    Free

Entry price over time

Not enough pricing history yet — we start charting from the second observation.

Geek mode

Models under the hood

  • Llama / Kimi / GPT-OSS

    open weights

    disclosed
  • LPU v2

    Groq

    disclosed

Context window

Public API

yes

Multi-model routing

yes

Stack signals

  • SRAM-only memory
  • compiler-scheduled execution
  • OpenAI-compatible

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

  • Llama / Kimi / GPT-OSSopen weights · 1 product
  • LPU v2Groq · 1 product

green disclosed · amber inferred · grey unknown

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Chips & Compute

All Chips & Compute

Wafer-scale inference and training

Version
WSE-3 / Inference
Cost
Free tier, then from $0.06 / M tokens
Model
Llama / Qwen / GPT-OSS

Inference-first TPU pods

Version
v7 Ironwood
Cost
from ~$1.2 per chip-hour
Model
TPU v7

High-memory GPU alternative

Version
MI355X
Cost
~$2-3 per GPU-hour on clouds
Model
CDNA 4

On-device inference across Mac, iPhone and iPad.

Version
M5
Cost
Bundled with hardware
Model
Apple Foundation Models

Cloud-native training silicon

Version
Trainium2
Cost
from ~$1.3 per accelerator-hour
Model
Trainium2

Co-designed XPUs behind most hyperscaler silicon.

Version
2025
Cost
Enterprise supply agreements
Model
powers Google TPU

Change history

  • pricing

    Token pricing cut across the open-model catalogue

    0.080.05

    source