With Groq 3 LPX in Full Production, NVIDIA Extends Vera Rubin Inference for Agents
Article image or reusable cover for NVIDIA
NVIDIA has launched Groq 3 LPX, a system now in full production designed to run agent-based AI systems that can make independent decisions and use tools.
The system can generate 3,400 output tokens per second—approximately four times faster than the nearest competitor's solution. NVIDIA combines three components: Vera Rubin processors for reading and processing large amounts of text, Spectrum-X networks for fast data communication, and Groq 3 LPX for generating responses. Companies including SpaceXAI, CoreWeave, and Nebius are already adopting the platform.
The next era of AI inference won't be defined by a single breakthrough chip, network or system. It'll be defined by how every layer of the AI factory works together.
Vibekollen prepared this summary with AI from the original publication. The content belongs to NVIDIA.
More to read
Server-Side Code Execution Tools for AI Agents, Compared
OpenRouter 10 h ago
v0.40.0
Ollama 11 h ago
Google froze its open source bug bounty program due to a ‘significant rise’ in AI submissions
TechCrunch AI 14 h ago
Can ‘super intelligence’ and a non-binding safety pact solve AI’s image problem?
TechCrunch AI 14 h ago