Skip to content
VibekollenBETAVibekollen
BlogNVIDIA

With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents

Article image or reusable cover for NVIDIA

NVIDIA has launched Groq 3 LPX, a system now in full production designed to run agent-based AI systems that can make independent decisions and use tools.

The system can generate 3,400 output tokens per second—approximately four times faster than the nearest competitor's solution. NVIDIA combines three components: Vera Rubin processors for reading and processing large amounts of text, Spectrum-X networks for fast data communication, and Groq 3 LPX for generating responses. Companies including SpaceXAI, CoreWeave, and Nebius are already adopting the platform.

The next era of AI inference won't be defined by a single breakthrough chip, network or system. It'll be defined by how every layer of the AI factory works together.
Verbatim from the article at NVIDIA
Read the full story at NVIDIA →

Vibekollen prepared this summary with AI from the original publication. The content belongs to NVIDIA.

More to read