NVIDIA Groq 3 LPX

NVIDIA Groq 3 LPX

The NVIDIA Groq 3 LPX is an inference accelerator designed for NVIDIA Vera Rubin, optimized for low-latency, large-context agent systems and large-scale content generation. Each rack supports 256 LPX units, delivering 128GB of SRAM, 40PB/s of memory bandwidth, and 640TB/s of rack-level scaling bandwidth.

In Stock
Inventory:99item

Core Specifications

GPU Architecture
Vera Rubin
LPX Quantity
256 per rack
SRAM
128GB
Memory bandwidth
40 PB/s
Expand Bandwidth
640 TB/s
Usage
Low-latency inference acceleration

CompleteSpecifications

GPU ArchitectureVera Rubin
LPX Quantity256 per rack
SRAM128GB
Memory Bandwidth40 PB/s
Expand Bandwidth640 TB/s
UsageLow-latency inference acceleration
ShapeRack-level
Beijing ICP License No. 2026026760