Nvidia launches Groq 3 AI chip and CPU server aimed at Intel during GTC 2026

Nvidia launches Groq 3 AI chip and CPU server aimed at Intel during GTC 2026

Nvidia (NVDA) kicked off its GTC occasion in San Jose, Calif., on Monday, debuting various chips and platforms starting from its all-new Nvidia Groq 3 language processing unit (LPU) to its large Vera central processing unit (CPU) rack, designed to go head-to-head with choices from Intel (INTC) and AMD (AMD).

All totaled, Nvidia mentioned it’s rolling out 5 large server racks, every serving totally different functions inside AI information facilities.

The largest announcement of the lot, although, is the Nvidia Groq 3 chip. Nvidia introduced it had entered into an settlement to license know-how from Groq and employed founder Jonathan Ross, president Sunny Madra, and different members of the Groq crew as a part of a $20 billion deal in December.

Groq’s processors concentrate on AI inferencing, or operating AI fashions. It’s what occurs if you sort one thing into OpenAI’s (OPAI.PVT) ChatGPT, Anthropic’s (ANTH.PVT) Claude, or Google’s (GOOG, GOOGL) Gemini and get a response.

Nvidia’s graphics processing items (GPUs) are multipurpose and can each prepare and run AI fashions, however because the AI market strikes towards operating fashions, making certain the corporate has a devoted inferencing chip has turn into paramount.

That’s the place Groq 3 is available in.

FILE PHOTO: Nvidia CEO Jensen Huang speaks at a conference hosted by chip design software firm Synopsys in Santa Clara, California, U.S., March 11, 2026. REUTERS/Stephen Nellis/File Photo
Nvidia CEO Jensen Huang speaks at a convention hosted by chip design software program agency Synopsys in Santa Clara, Calif., on March 11, 2026. (Reuters/Stephen Nellis) · Reuters / REUTERS

According to Nvidia vp of hyperscale and high-performance computing Ian Buck, whereas Nvidia’s GPUs assist way more reminiscence than Groq 3, the LPU’s reminiscence is quicker. So the corporate is combining the efficiency advantages of each chips.

To try this, Nvidia is launching its Groq 3 LPX platform, a server rack powered by 128 particular person Groq 3 LPUs. When used along with Nvidia’s Vera Rubin NVL72 rack the corporate says clients might see 35x increased throughput per megawatt of energy and 10x extra income alternative.

“Optimized for trillion-parameter models and million-token context, the codesigned LPX architecture pairs with Vera Rubin to maximize efficiency across power, memory and compute. The additional throughput per watt and token performance unlocks a new tier of ultra-premium, trillion-parameter, million-context inference, expanding revenue opportunity for all AI providers,” the corporate mentioned in an announcement.

The LPX rack ought to assist handle issues that Nvidia might ultimately lose its edge within the AI race to upstart corporations designing inference-focused processors.

In addition to the LPX, Nvidia revealed its Vera CPU rack. When Nvidia talks about its Vera Rubin superchip, it’s referring to 3 processors in a single: a Vera CPU and two Rubin GPUs.

Now the corporate is breaking off Vera into its personal standalone chip, which it’s going to slot into devoted Vera server racks that mix 256 liquid-cooled Vera chips into one system.

(*3*) as agentic AI — the place autonomous and semiautonomous bots carry out duties on customers’ behalf — begins to take off. While GPUs and LPUs are necessary for serving to to energy AI fashions, when AI brokers go off to, say, browse an internet site or pull info from a spreadsheet, they’re counting on CPU efficiency.

The chips additionally play an integral half relating to information mining, personalization, and the evaluation that gives context to a GPU and in the end an AI mannequin.

“Vera is the best CPU for agentic AI workloads,” Buck mentioned. “We’ve designed a new kind of CPU, the Olympus core, engineered by NVIDIA for AI execution. Vera enables faster agentic responses under the extreme conditions for all of the agentic AI use cases and reinforcement learning.”

This isn’t the primary time Nvidia has talked about its personal CPU server. Last month, the company announced a deal with Meta (META) that may see Nvidia present the social media big with the largest-ever deployment of its prior-generation Grace CPUs.

But the Vera announcement serves as Nvidia’s effort to cement its standing as not only a GPU firm but in addition a CPU firm, establishing a showdown with rival and collaborator Intel, in addition to AMD, within the information heart.

In addition to the Vera Rubin NVL72, Groq LPX, and Grace rack, Nvidia confirmed off its new storage rack system referred to as the Bluefield-4 STX, which the corporate mentioned improves efficiency in comparison with conventional storage racks. Nvidia additionally talked up its Spectrum-6 SPX networking rack.

The new choices in Nvidia’s lineup ought to assist the corporate preserve increasing its information heart income as demand for AI platforms continues to develop. The firm reported information heart income of $193.5 billion in its fiscal 2026, up from $ 116.2 billion in fiscal 2025.

And with hyperscalers like Amazon (AMZN), Google, Meta, and Microsoft (MSFT) set to spend $650 billion this year on AI capabilities, Nvidia is certain to see a superb chunk of that.

Sign up for Yahoo Finance's Week in Tech newsletter.
Sign up for Yahoo Finance’s Week in Tech e-newsletter. · Yahoo Finance

Email Daniel Howley at dhowley@yahoofinance.com. Follow him on Twitter at @DanielHowley.

Click here for the latest technology news that will impact the stock market

Read the latest financial and business news from Yahoo Finance

Leave a Reply

Your email address will not be published. Required fields are marked *