AI Chips · Brand guide

Groq: The Complete Guide

Groq built the LPU, a chip made only for running AI models fast. In December 2025 it licensed that technology to NVIDIA and its founder moved there, while Groq kept running its cloud. What the LPU is, how it works, who makes it, what GroqCloud costs, and the facts in plain sight. Three reading levels, every number sourced.

Stuck at any point? Ask an AI about this page →
Groq 16 sections · 3 levels 35 linked sources Checked September 2026
How to read this page. Every brand in our AI Chips section uses the same 16 sections in the same order. Each section starts Simple, then goes Deeper, then Expert: stop wherever you have what you need. The small numbers are sources: click one to open the original document. Where a company has not published something, we say so rather than estimate.

1.At a glance

SimpleStart here

Groq designed the LPU (Language Processing Unit), a chip built only to run AI models quickly, not to train them 1. Instead of separate memory chips, it keeps everything in fast memory on the chip itself 1. Groq rents LPUs through its cloud service, GroqCloud, which it says millions of developers use 2 3.

What it isAn AI chip for inference (answering), with memory on the chip 1
Big changeLicensed its technology to NVIDIA in December 2025; founder moved to NVIDIA 4
Who builds itFirst chip on a 14 nm process; Samsung makes NVIDIA's Groq 3 chip (press) 1 5
Can you use one?Yes, through GroqCloud, paid per token 6
DeeperThe detail

On 24 December 2025 Groq announced a "non-exclusive licensing agreement" giving NVIDIA its inference technology; founder Jonathan Ross, president Sunny Madra and other staff joined NVIDIA, while Groq said it "will continue to operate as an independent company" and "GroqCloud will continue to operate without interruption" 4. Neither company gave a price; press reports put it at about $20 billion 7.

ExpertFor specialists

NVIDIA now sells a rack called the NVIDIA Groq 3 LPX, in full production from August 2026; "Groq and LPU are used under license from Groq, Inc." 8. Groq itself became an NVIDIA Cloud Partner and plans to be among the first to bring those racks to market 9 10. In August 2026 Groq raised $350 million at a $3.5 billion valuation, about half its September 2025 value 3 11.

2.Company card

SimpleStart here

Groq was founded in 2016 by Jonathan Ross 11. After he left for NVIDIA, Groq has been led by Adam Winter as chief executive 12 9.

Legal nameGroq, Inc. 8
HeadquartersSan Francisco, California (2026 releases); earlier Mountain View 3 11
FounderJonathan Ross, now at NVIDIA 4
Chief ExecutiveAdam Winter 12 9
Executive ChairmanAlex Davis, also CEO of lead investor Disruptive 3
Chief Technology OfficerSinclair Schuller 10
StatusPrivate company 12

Groq does not publish a street address or staff numbers in the releases we read.

DeeperThe detail

Leadership changed quickly. Simon Edwards became chief executive at the NVIDIA deal 4; a Bloom Energy filing shows he was Groq's finance chief from September 2025 and CEO only from December 2025 to March 2026, before becoming Bloom's finance chief 13. By June 2026 Groq's releases named Adam Winter as CEO 12.

ExpertFor specialists

Funding: $640 million at a $2.8 billion valuation (August 2024, led by BlackRock) 14; $750 million at $6.9 billion (September 2025, led by Disruptive) 11; $650 million in June 2026 12; and $350 million called a "Series A" at $3.5 billion in August 2026, with planned participation from NVIDIA 3.

3.The story

SimpleStart here
  • 2016 Groq founded 11.
  • 2023 Picks Samsung's Texas plant for a next-generation chip 15.
  • 2024 $640 million funding; Saudi data centre built in eight days 14 16.
  • 2025 $1.5 billion Saudi commitment, Meta Llama API, Bell Canada, NVIDIA licence 16 17 18 4.
  • 2026 NVIDIA Cloud Partner; $1 billion of new funding 9 3.
DeeperThe detail

Groq grew a developer following fast: 360,000 in August 2024, 2 million by September 2025 and more than 6 million by August 2026, by its own count 14 11 3. In 2024 it aimed to deploy more than 108,000 LPUs by the end of March 2025 14.

ExpertFor specialists

The 2023 Samsung deal was for a 4 nm chip at Samsung's new Taylor, Texas plant, which Groq said would scale "from 85,000 to over 600,000 chips" 15. No Groq-branded 4 nm chip has shipped; the chip that emerged is NVIDIA's Groq 3, which press reports say Samsung makes on 4 nm 19 5.

4.The lineup

SimpleStart here
ProductWhat it isStatus
LPU / GroqChip (gen 1)The original 14 nm chipRuns GroqCloud 1
GroqCardOne chip on a PCIe cardDatasheet published 20
GroqNode / GroqRack8 cards per server; racks of serversUsed in labs 21
GroqCloudCloud service, pay per token13 data centres 3
NVIDIA Groq 3 LPXNVIDIA's rack of 256 LPUs, under licenceFull production, Aug 2026 22 8
DeeperThe detail

GroqCloud now has three tiers: GroqMetal (infrastructure), GroqCore (inference) and GroqAssured (enterprise controls) 23. Groq's home page says: "We pioneered the LPU. Now, with LPX, it works alongside NVIDIA's next-generation GPUs" 2.

ExpertFor specialists

NVIDIA's Groq 3 chip has 500 MB of SRAM and 150 TB/s of memory bandwidth; each LPX rack holds 256 of them, 128 GB of SRAM plus 12 TB of DDR5 memory 22. Groq quotes 315 petaflops of FP8 compute and 1,000 tokens a second per user per rack 23.

5.How it’s made

SimpleStart here

Groq designs the chip; others make it.

  1. First chip: 14 nm process 1; press reports name GlobalFoundries as the maker 24.
  2. Groq 3 (NVIDIA): Samsung, 4 nm, press reports say 5.
  3. Cards: the GroqCard datasheet is published by BittWare, a Molex company 20.
  4. Racks and servers: Groq is deploying NVIDIA racks with Dell 10.
DeeperThe detail

In 2024 Groq's public-sector partner said "all Groq systems are designed, fabricated, and assembled in North America" 25. Groq's 2023 Samsung release said the deal "strengthens Groq's already completely North American-based operations for engineering and manufacturing" 15.

ExpertFor specialists

The first chip measures about 725 mm² with about 28.6 billion transistors, according to press reports citing analysts 24. An analyst estimate puts its wafer cost at "likely less than $6,000", well below leading-edge chips 26. At NVIDIA's GTC in March 2026, press reports say, Jensen Huang said Samsung "manufactures the Groq 3 LPU chip" 5.

6.Where it’s made: factories and addresses

SimpleStart here

Groq does not name a chip factory address. Its GroqCloud data centres, from official releases:

SiteLocationWhat we know
Saudi ArabiaDammam (2025 release); Riyadh (current list)"Brought online in just eight days" (Dec 2024) 16 23
KamloopsBritish Columbia, Canada7 MW, for Bell Canada 18
Helsinki / VantaaFinland (with Equinix)First European site, 2025 27 23
SydneyAustralia (Equinix)4.5 MW, Nov 2025 28
US sitesLiberty Lake, St Paul, Dallas, HoustonListed on Groq's platform page 23

Groq publishes towns but not street addresses 23.

DeeperThe detail

Groq's current list also names Vaudreuil-Dorion and Calgary in Canada and London in the UK 23. In August 2026 it said it runs 54 MW across 13 data centres and expects more than 200 MW in 2027 3.

ExpertFor specialists

Bell Canada plans 500 MW across six sites, with Groq as its exclusive inference provider 18. The Saudi site moved from "Dammam" in the 2025 release to "Riyadh" on the current list; Groq has not explained the change 16 23.

7.How it works

SimpleStart here

A GPU fetches the AI model from separate memory chips for every word it writes. Groq's LPU keeps the model in fast memory on the chip (SRAM), which Groq says has "memory bandwidth upwards of 80 terabytes/second, while GPU off-chip HBM clocks in at about eight terabytes/second" 1.

DeeperThe detail

The LPU is deterministic: the compiler plans every step in advance, so data "executes the same way every time" 1. Chips connect directly, with "no need for routers or controllers", so many chips act like one long conveyor belt 1. Groq calls this a Tensor Streaming architecture 15 29.

ExpertFor specialists

The catch is memory size: 230 MB per first-generation chip 20. An analyst counted 576 chips, in 8 racks of 9 servers with 8 chips each, to serve one Mixtral model 26. Groq has not published an official chip count for a given model size.

8.Specs explained

SimpleStart here

Official figures for the first-generation chip (GroqCard datasheet and Argonne lab documentation):

ItemValue
On-chip memory (SRAM)230 MB 20 21
Memory bandwidthUp to 80 TB/s 20
Compute750 TOPS (INT8); 188 TFLOPS (FP16) at 900 MHz 21
Power (card)240 W typical, 275 W rated, 375 W maximum 20
Process14 nm 1
DeeperThe detail
ChipSRAMBandwidthPer rack
Groq LPU (gen 1)230 MB80 TB/sRacks of 9 servers × 8 cards 21
NVIDIA Groq 3 (LPX)500 MB150 TB/s256 chips, 128 GB SRAM 22
ExpertFor specialists

Press reports give Groq 3 at 98 billion transistors on Samsung's SF4X process 5; NVIDIA's page gives 500 MB of SRAM where one press report says 512 MB, so we use NVIDIA's figure 22 19. The GroqCard has up to nine chip-to-chip connectors 20.

9.Official pricing

SimpleStart here

GroqCloud's official prices, from Groq's model documentation:

ModelInputOutputSpeed Groq lists
GPT OSS 120B$0.15 per million tokens$0.60 per million tokens500 tokens/s 6
GPT OSS 20B$0.075 per million tokens$0.30 per million tokens1,000 tokens/s 6
Llama 3.3 70B"Contact Sales"280 tokens/s 6
Whisper Large V3 (speech)$0.111 per hour of audio6
Chips or racksNot published by the company20
DeeperThe detail

The independent tester Artificial Analysis measured Groq at about 477 tokens a second on GPT OSS 120B and 270 on Llama 3.3 70B (third-party measurements) 30. For the same GPT OSS 120B model, Cerebras charges $0.35 in and $0.75 out; see our Cerebras guide.

ExpertFor specialists

Groq does not publish hardware prices, and neither NVIDIA's LPX page nor the GroqCard datasheet gives one 22 20. Groq's price list changes often; we show only models whose prices appear on the official page.

10.The software lock-in

SimpleStart here

Groq's edge was speed and predictability: a chip planned step by step by software, with memory on board 1. Groq claims "up to a 10X speed advantage" over GPUs and similar energy savings (Groq's own claim) 1.

DeeperThe detail

The technology was valuable enough for NVIDIA to license it and hire its founder, and NVIDIA now sells LPU racks alongside its GPUs 4 8. Because the licence is non-exclusive, Groq keeps the right to use its own designs 4.

ExpertFor specialists

Groq's remaining advantage is its cloud: 13 data centres, millions of developers and exclusive deals such as Bell Canada 3 18. But it now runs NVIDIA hardware too, so its edge is increasingly the service rather than the chip 9.

11.Who buys it and why

SimpleStart here

Official partners and customers include Meta (Llama API), Bell Canada, Saudi Arabia's HUMAIN, McLaren's Formula 1 team, Paytm in India and Canva 17 18 31 32 33 28. US government agencies can buy through Carahsoft 25.

DeeperThe detail

Saudi Arabia committed $1.5 billion to expand Groq's inference infrastructure in February 2025 16. Meta's Llama API launched with Groq at speeds of up to 625 tokens a second 17.

ExpertFor specialists

Groq has not published revenue in any release we read 12. Press reports, citing The Information, said in July 2025 that Groq cut its 2025 revenue forecast from about $2 billion to about $500 million after delays to the Saudi deal 34.

12.Hidden in plain sight

SimpleStart here

Public facts in plain sight. Each links to the exact document.

  1. NVIDIA sells a product with Groq's name on it, under licence, and Groq plans to deploy it 8 10.
  2. A "Series A" after earlier rounds up to Series D, at about half the September 2025 valuation 3 14 11.
  3. The post-deal CEO lasted about three months; it shows up in a Bloom Energy filing, not a Groq release 13.
  4. The Saudi data centre went up in eight days 16.
  5. The executive chairman runs the lead investor 3.
DeeperThe detail
  1. The headquarters moved in releases: Mountain View, Palo Alto once, then San Francisco 11 31 3.
  2. The Samsung 4 nm chip announced in 2023 never shipped as a Groq product 15 19.
  3. Groq was listed on a US Army software contract through Carahsoft 25.
  4. Llama 3.3 70B has no public price, only "Contact Sales" 6.
  5. Developers grew about 17 times in two years, from 360,000 to more than 6 million 14 3.
ExpertFor specialists
  1. No price for the NVIDIA deal appears in either company's release 4 8.
  2. The US Justice Department is reported to be examining the deal; no official statement yet 35.
  3. Groq's own speeds and outside tests differ: 500 against about 477 tokens a second for GPT OSS 120B 6 30.
  4. Yann LeCun was named a technical adviser in 2024 14.
  5. The Saudi site name changed from Dammam to Riyadh 16 23.

13.Weak spots

SimpleStart here

Limits, from the public record:

  • Tiny memory per chip: 230 MB, so big models need hundreds of chips 20 26.
  • Older process: 14 nm for the first chip 1.
  • Lost its founder, president and other staff to NVIDIA 4.
  • Three chief executives in about six months 4 13 12.
DeeperThe detail

The valuation fell from $6.9 billion to $3.5 billion between September 2025 and August 2026 11 3. Press reports say the Justice Department is looking at whether the NVIDIA deal was structured to avoid a merger review; neither company nor the department has commented publicly 35.

ExpertFor specialists

Groq's dependence on NVIDIA is growing: it is certified to "design, deploy, and operate NVIDIA accelerated computing" and plans NVIDIA Groq 3 LPX and Vera Rubin racks in its data centres 9 10. The 2025 revenue forecast cut reported by the press shows how much one customer region mattered 34.

14.India angle

SimpleStart here

Groq's India link is Paytm: in November 2025 Paytm chose GroqCloud for real-time AI in payments, including fraud and risk checks 33. Groq has not announced an India office or data centre.

DeeperThe detail

The Paytm release was datelined Noida and Mountain View, and Groq's Asia-Pacific general manager, Scott Albin, covers the region 33. Groq's nearest data centre to India is in Saudi Arabia 23.

ExpertFor specialists

For India's own chip plants and plans, see our India's semiconductor push guide.

15.What’s next

SimpleStart here

Only what Groq and NVIDIA have announced:

  • More capacity: over 200 MW in 2027 3.
  • NVIDIA hardware: Groq 3 LPX and Vera Rubin racks in GroqCloud, with Dell 10.
  • LPX: in full production at NVIDIA since August 2026 8.
DeeperThe detail

Groq now describes itself as an inference cloud rather than a chip company 12. Whether it builds another chip of its own has not been announced.

ExpertFor specialists

The open questions are regulatory, the reported US review of the NVIDIA deal 35, and commercial, whether GroqCloud can grow against far larger clouds 3.

16.Sources

35 sources, all checked September 2026. Official = the company or organisation’s own page; Filing = a document filed with a regulator; Press = news coverage, used only where no official source exists and labelled in the text.

  1. Groq Blog: The Groq LPU explainedOfficial
  2. Groq: Groq home pageOfficial
  3. Groq Newsroom: Groq closes $350 million Series A (17 Aug 2026)Official
  4. Groq Newsroom: Groq and NVIDIA enter non-exclusive inference technology licensing agreement (24 Dec 2025)Official
  5. SamMobile (press): Samsung makes Groq 3 LPU chips for NVIDIA (Mar 2026)Press
  6. GroqCloud Docs: Supported models: prices and speedsOfficial
  7. DatacenterDynamics (press): NVIDIA to license tech from Groq and hire its leadership (Dec 2025)Press
  8. NVIDIA Newsroom: NVIDIA Groq 3 LPX now in full production (24 Aug 2026)Official
  9. Groq Newsroom: Groq becomes an NVIDIA Cloud Partner (12 Aug 2026)Official
  10. Groq Blog: Groq among the first to bring NVIDIA Groq 3 LPX and Vera Rubin NVL72 to market (24 Aug 2026)Official
  11. Groq (PR Newswire): Groq raises $750 million as inference demand surges (Sep 2025)Official
  12. Groq Newsroom: Groq raises $650M to scale its AI inference cloud business (22 Jun 2026)Official
  13. Bloom Energy (SEC): Form 8-K: appointment of Simon Edwards as CFO (26 Mar 2026)Filing
  14. Groq (PR Newswire): Groq raises $640M to meet soaring demand for fast AI inference (5 Aug 2024)Official
  15. Groq (PR Newswire): Groq selects Samsung Foundry to bring next-gen LPU to market (Aug 2023)Official
  16. Groq (PR Newswire): Saudi Arabia announces $1.5 billion expansion with Groq (10 Feb 2025)Official
  17. Groq (PR Newswire): Meta and Groq collaborate on the official Llama API (29 Apr 2025)Official
  18. Groq Newsroom: Groq becomes exclusive inference provider for Bell Canada's sovereign AI network (28 May 2025)Official
  19. Tom's Hardware (press): NVIDIA's Groq deal produces its first chipPress
  20. BittWare, a Molex company (datasheet via Mouser): GroqCard accelerator product briefOfficial
  21. Argonne Leadership Computing Facility: Groq system in the ALCF AI TestbedGovernment
  22. NVIDIA: Groq 3 LPX product pageOfficial
  23. Groq: Platform page: data centres and service tiersOfficial
  24. The Register (press): Groq AI chip dev kit: 14 nm, fabbed by GlobalFoundries (29 Sep 2020)Press
  25. Carahsoft: Groq and Carahsoft partner to bring AI inference to the public sector (May 2024)Official
  26. SemiAnalysis (analyst): Groq inference tokenomics: speed, but at what cost? (Feb 2024)Research
  27. Groq Newsroom: Groq launches European data centre footprint in Helsinki (Jul 2025)Official
  28. Groq Newsroom: Groq expands to Asia-Pacific with Sydney data centre (17 Nov 2025)Official
  29. IEEE Computer Society: Think Fast: A Tensor Streaming Processor for accelerating deep learning workloads (ISCA 2020)Paper
  30. Artificial Analysis (third-party): Groq provider benchmarksResearch
  31. Groq (PR Newswire): Groq and HUMAIN launch OpenAI's new open models day zero (5 Aug 2025)Official
  32. Groq (PR Newswire): McLaren Racing announces Groq as an official partner (26 Sep 2025)Official
  33. Groq Newsroom: Groq partners with Paytm for real-time AI in India (5 Nov 2025)Official
  34. TrendForce (press, citing The Information): Groq cuts 2025 revenue projection by $1.5B (31 Jul 2025)Press
  35. SDxCentral (press): NVIDIA's Groq deal facing DOJ probe, report says (Sep 2026)Press

In the news

Newest stories mentioning Groq, straight from the companies’ own newsrooms and the wider press. Links open the original.

All of today’s AI news →

Ask an AI about this page

Opens your assistant with this page as the source, and a question rather than a summary. It will ask what you are building before it answers.

ChatGPTClaudeGeminiPerplexityGrok

Nothing is sent from here. The link carries only this page’s title and address.