Trending

    Nvidia Launches Groq 3 LPX Inference Accelerator Amid Benchmark Controversy

    Section editor: ·Moderate4 articles covering this·3 news sources·Updated 2 hours ago·World
    Share:
    Infographic comparing Nvidia Groq 3 LPX performance metrics with Cerebras.

    Here's what it means for you.

    The launch of Nvidia's Groq 3 LPX could redefine performance benchmarks in AI inference, impacting your tech strategy.

    What happened

    On August 25, 2026, Nvidia announced the full production of its Groq 3 LPX inference accelerator at the Hot Chips 2026 conference.

    The Context

    • Acquisition Impact: Nvidia's $20 billion acquisition of Groq assets in December 2025 aims to meet the rising demand for low-latency inference in AI workloads.
    • Performance Claims: Nvidia claims the Groq 3 LPX achieves 3,400 tokens per second, reportedly four times faster than Cerebras, though critics question the validity of this comparison.
    • Market Dynamics: The competition heats up as Cerebras launches its CS-4 hardware, intensifying the race for specialized inference solutions.

    The Number

    3,400

    — This is the number of tokens per second the Groq 3 LPX can generate, a significant metric for professionals focused on AI performance and efficiency.

    Takeaway

    As Nvidia ramps up production, expect increased competition and innovation in the AI inference space, which could influence your tech investments.

    4 Articles
    THE DECODER

    Nvidia says its Groq 3 LPX is four times faster than Cerebras, but the math is more complicated

    Nvidia has announced that its Groq 3 LPX inference chip has entered full production, achieving a performance of 3,400 tokens per second on the Gemma 4 31B model, which is reported to be four times faster than Cerebras. However, this performance requi...

    16 hours ago
    Read Full Article
    Techmeme

    Nvidia says its Groq 3 LPX racks delivered 3,400 tokens per second in an Artificial Analysis benchmark running Gemma 4 31B with a 100,000-token input sequence (The Register)

    Nvidia announced that its Groq 3 LPX racks achieved a performance of 3,400 tokens per second in an Artificial Analysis benchmark using the Gemma 4 31B model with a 100,000-token input sequence. This milestone highlights the capabilities of Nvidia's l...

    Techmeme

    Nvidia says its inference accelerator Groq 3 LPX has entered full production and Nebius has signed on as the first customer; SpaceXAI will adopt Vera CPUs (Mike Wheatley/SiliconANGLE)

    Nvidia has announced that its Groq 3 LPX inference accelerator has entered full production, with Nebius as its first customer. This development signifies a significant step in Nvidia's efforts to enhance its AI capabilities and market presence.

    Investing.com

    Nvidia launches Groq 3 LPX AI inference accelerator

    Nvidia has launched the Groq 3 LPX AI inference accelerator, marking a significant advancement in its AI technology offerings. This new product aims to enhance the performance and efficiency of AI applications, catering to the growing demand in vario...