GroqPlatform

Build your perfect stack. GroqMetal provides infrastructure, GroqCore adds inference, and GroqAssured adds enterprise controls. Each includes the layers below.

  • Infrastructure

    GroqMetal

    Dedicated bare-metal infrastructure, tuned for speed and reliability, with full control in your hands.

  • Inference

    GroqCore

    A tested inference stack that turns dedicated capacity into production-ready performance, no infrastructure expertise required.

  • Control

    GroqAssured

    Enterprise-grade governance, auditability, and control layered on top, so scale never comes at the cost of trust.

  • 256 LPUs per rack
  • 40 PB/s SRAM bandwidth
  • 1,000 tokens/sec/user
  • 128 GB of on-chip SRAM per rack
  • 315 PFLOPS of FP8 inference compute

Groq operates LPX globally to deliver fast, reliable inference at massive scale with fine-grained control.

Each NVIDIA Groq 3 LPX rack connects 256 next-generation LPU accelerators to NVIDIA’s Vera Rubin to deliver low-latency, large-context inference for production AI.

Globally online

Now operating 13 data centers across four continents

  • US-1Liberty Lake, USA
  • US-2St Paul, USA
  • US-3Dallas, USA
  • US-4Houston, USA
  • CA-1Kamloops, Canada
  • CA-2Vaudreuil-Dorion, Canada
  • CA-3Calgary, Canada
  • EU-1Vantaa, Finland
  • AU-1Sydney, Australia
  • EU-2London, UK
  • SA-1Riyadh, Saudi Arabia

Put inference
to work