Invest2 distinct publishers3 min readUpdated
Hardware from the largest acquisition in Nvidia's history is in production and headed to a neocloud this year. The market sold it anyway, two days before earnings.
The Investor · Invest desk

Compiled by The InvestorSomething wrong?How this is made
Roughly five points of intraday reversal on a company valued near $5.2 trillion is about $260 billion of market capitalisation moving on one product announcement [21], and it came in a month when the shares had been running nearly 7% against 2.5% for the S&P 500 [20].
What the buyer of that stock is underwriting here is smaller than the headline suggests. Nvidia paid $20 billion for Groq's assets in December [3]. Set against the $1 trillion in cumulative Blackwell and Vera Rubin sales Jensen Huang projected through 2027 when the products were unveiled in March, the deal is about 2% of the revenue Nvidia has already pointed the market at [23]. Groq does not need to be a large business. It needs to stop someone else from owning the fast end of serving.
The hardware is narrower than a GPU by design. Each Groq 3 chip carries 500 megabytes of SRAM on the die itself to cut memory bottlenecks [7], and Nvidia packages 256 of them into an LPX rack [6], so a full rack holds roughly 128 gigabytes of on-chip memory [22]. That is a decode-phase machine, which is exactly what Nvidia says it is: low-latency parts handle the decode portion of serving, while GPUs still do training and adapt to new models [11]. Senior director Dion Harris put it as "using the right price, right processor for the right part of the workload" rather than replacing GPUs [10]. Huang's own allocation was more specific about the ceiling: a quarter of the data centre space intended for coding goes to Groq, and "the rest of my data center is all 100% Vera Rubin" [15].
The 3,400 tokens per second Nvidia cites is a full-rack number from an Artificial Analysis benchmark [6]. OpenAI's Ultrafast mode is quoted at 750 tokens per second on Cerebras silicon [13]. The division gives 4.5x [26], but the sources do not say the two figures are measured on the same basis, and one is a rack while the other is a service tier, so it is a marketing comparison rather than a competitive one. AMD has already said it will pair its rack-scale systems with Cerebras chips for the same low-latency work [12].
The quieter point is the fab. Samsung builds Groq's chips while TSMC builds Nvidia's GPUs [8]. Wedbush's Matt Bryson, who rates the stock Outperform, argues that "component and material access, not end demand, is defining shipments" [17], which makes a Samsung-fabbed product line a second source of units as much as a latency play.
The sell side shows how wide the range still is. Rosenblatt keeps a Buy and a $325 target against a $214.72 quote, about 51% above the tape [16][24], while InvestingPro's fair value of $259.96 sits 21% above the price, with Rosenblatt's target a quarter above that [25]. All of it rests on a consensus that already has quarterly revenue and earnings roughly doubling from a year ago [19]. And the premium-tier logic, that cloud providers can charge more for tokens carrying stricter latency terms [9], is still the vendor's own assertion about what buyers will pay.
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
Nvidia announced on Monday that its Groq 3 LPX rack is in full production, marking the commercialization of technology from the company's largest acquisition on record.
The Groq rack will be deployed alongside Vera central processors and Rubin graphics processors at neocloud Nebius and will be online later this year, Nvidia senior director Dion Harris told reporters.
In December, Nvidia bought assets from chip startup Groq for $20 billion, the company's largest purchase.
Nvidia fell about 2% Monday after opening roughly 3% higher; the stock jumped about 3% soon after the opening bell, then dropped roughly 2% after the Groq news came out.
Nvidia packages 256 individual Groq 3 chips into its LPX racks and says the rack can deliver 3,400 tokens per second, citing a benchmark from Artificial Analysis.
The Groq architecture includes 500 megabytes of fast SRAM on the chip's die itself to reduce memory-related bottlenecks.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Named sourcing, single-origin technical claims
Core facts are on the record: an Nvidia senior director briefing reporters, an attributed CEO projection from March, and financial data attributed to FactSet, LSEG and named brokers. But the performance number is vendor-cited rather than independently reproduced, the second publisher largely restates the first, and the causal link between the announcement and the share move rests on one outlet's sequencing.
Production announced, one named customer, nothing live
Adoption evidence is a production declaration plus a single named neocloud (Nebius) with a 'later this year' go-live. There is no disclosed capacity, customer count, revenue, price tier or usage figure, and the vendor's own capacity model caps Groq at a quarter of coding-oriented data-centre space. Rival Cerebras silicon is already behind a user-facing OpenAI mode, which is adoption for the category but not for this product.
Milestone language runs ahead of deployed reality
'Full production' and a 4.5x-looking throughput headline sit against zero live capacity, one named customer, an unverified benchmark and no attributable revenue from a $20 billion purchase that equals roughly 2% of the vendor's own $1 trillion cumulative projection. The overstatement is modest rather than severe because the primary source itself supplies the limiting caveats — decode-phase only, not a GPU replacement, 25% of coding capacity — and the market's negative reaction ran counter to the promotional framing.
Vendor briefing plus long-side sell-side, ahead of earnings
Every substantive technical claim originates in an Nvidia-controlled press call two days before the company's own earnings print, and the performance figure is one the vendor selected. The financial layer is supplied by a Buy-rated broker with a $325 target, an Outperform-rated analyst forecasting a beat, and a subscription valuation tool, with the second publisher also promoting its own live earnings coverage and newsletter. These interests are disclosed in the sources but not interrogated by either outlet.
Facts firm, significance unresolved
The announcement, deployment target, acquisition price, specifications and market data are consistently reported across two sources and are unlikely to be wrong. What remains uncertain is what any of it is worth: no independent benchmark, no live deployment, no revenue attribution, and only one publisher's word that the Groq news caused the intraday reversal. Confidence in the ledger is therefore well above confidence in the story's implied significance.
product
Cerebras's CS-4 is three old wafers in a new rack: price the packaging, not the silicon2 distinct publishers
build
NVIDIA ships Groq 3 LPX and starts quoting inference in tokens per user, not per rack1 distinct publisher
invest
Nvidia's August 26 print: 92% of the quarter rides on one segment1 distinct publisher
build
Solar Pro 4 turns model routing into a procurement decision, not a research one1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 24, 2026
1 article · August 24, 2026