Chips & Compute

Intel Gaudi 3

Intel · #13 most active of 18 in Chips & Compute

Compare →

Gaudi 3 targets price-performance against Nvidia parts with open Ethernet scale-out and an oneAPI/PyTorch software path.

Current version

Gaudi 3

Entry cost

OEM systems and cloud instances

Changes / 30d

0

Capabilities

  • 128GB HBM2e
  • Open Ethernet scale-out
  • PyTorch native
  • oneAPI toolchain
  • Price-performance focus

Pricing

  • OEM/cloud

    quoted

    Free

Entry price over time

Not enough pricing history yet — we start charting from the second observation.

Geek mode

Models under the hood

  • Accelerator

    Intel

    disclosed

Context window

n/a

Public API

yes

Multi-model routing

no

Stack signals

  • HBM2e
  • RoCE Ethernet
  • SynapseAI
  • PyTorch

Shared model stack

Other tracked products running on the same foundation models — a quick read on how much of the catalog moves when one of these models changes.

green disclosed · amber inferred · grey unknown

Reported scale

Reference data. Each figure is whatever the source actually said — weekly users, downloads, revenue run-rate — with its own definition and date. These are not comparable between tools and are never used to rank anything.

No public usage figure on record for this product.

Also in Chips & Compute

All Chips & Compute

Wafer-scale inference and training

Version
WSE-3 / Inference
Cost
Free tier, then from $0.06 / M tokens
Model
Llama / Qwen / GPT-OSS

Inference-first TPU pods

Version
v7 Ironwood
Cost
from ~$1.2 per chip-hour
Model
TPU v7

Deterministic low-latency inference

Version
GroqCloud
Cost
Free tier, then from $0.05 / M tokens
Model
Llama / Kimi / GPT-OSS

High-memory GPU alternative

Version
MI355X
Cost
~$2-3 per GPU-hour on clouds
Model
CDNA 4

On-device inference across Mac, iPhone and iPad.

Version
M5
Cost
Bundled with hardware
Model
Apple Foundation Models

Cloud-native training silicon

Version
Trainium2
Cost
from ~$1.3 per accelerator-hour
Model
Trainium2

Change history

  • capability

    Gaudi 3 instances land in additional cloud regions

    Broader cloud availability

    source