Samsung-built Groq 3 LPX racks enter full production

Nvidia's Samsung-made inference chip hits full production with a record 3,400 tokens per second.

ChipNews Staff
1 Min Read

Samsung’s foundry unit has cleared a milestone that Nvidia hopes changes the economics of AI inference: the Groq 3 LPX accelerator it builds on 4nm has moved into full production.

The August 24 announcement at Hot Chips marks the payoff of a relationship Nvidia CEO Jensen Huang made public at GTC in March, when he named Samsung as the manufacturer for the chip Nvidia acquired through its $20B purchase of Groq.

Groq 3 LPX is built for the token-heavy end of AI, where agents chain hundreds of inference steps. Nvidia says it reached 3,400 output tokens per second in Artificial Analysis benchmarks, roughly 4x faster response for latency-sensitive agents, and compresses coding work that spanned hours into minutes.

The accelerator slots beside the Vera Rubin NVL72 platform, Nvidia’s training and general-purpose workhorse, with the LPX handling the interactive layer of AI factories.

Huang framed inference as the growth engine of AI, and the ramp gives Samsung foundry a marquee win as it fights for AI-era customers against TSMC.

Racks built around the chip should come online this year, with Nebius among the first to deploy them, according to reports.

Share This Article