We are thrilled to announce that Groq will be among the first adopters of NVIDIA Groq 3 LPX, deploying it alongside NVIDIA Vera Rubin NVL72 in our purpose-built AI inference Cloud. Groq is working with Dell Technologies to deploy NVIDIA Groq 3 LPX. When Groq brings NVIDIA Groq 3 LPX capacity online, it arrives on infrastructure already optimized for high-demand inference workloads. For enterprises and AI companies building the next generation of agents, Groq will provide one of the earliest paths to put NVIDIA Groq 3 LPX to work on real production workloads. Read more here: https://lnkd.in/gNpJvM8F
Groq
Technology, Information and Internet
Mountain View, California 200,399 followers
Groq is fast, low cost inference. The Groq LPU delivers inference with the speed and cost developers need.
About us
Inference is the engine that powers AI, and Groq was built from the silicon up to deliver the world's fastest inference at scale. We pioneered the LPU—the first processor designed specifically for AI inference—and are transforming that innovation into a global cloud platform powering production AI workloads. With the capital, infrastructure, and team to execute, we're uniquely positioned to define the next era of AI infrastructure. The opportunity is massive, and it's still wide open. Now let's go build it!
- Website
-
https://groq.com/
External link for Groq
- Industry
- Technology, Information and Internet
- Company size
- 51-200 employees
- Headquarters
- Mountain View, California
- Type
- Privately Held
- Founded
- 2016
- Specialties
- ai, ml, artificial intelligence, machine learning, engineering, hiring, compute, innovation, semiconductor, llm, large language model, gen ai, systems solution, generative ai, inference, LPU, Language Processing Unit, neocloud, and data ceter
Employees at Groq
Locations
-
Primary
Get directions
Mountain View, California, US
Updates
-
The hits keeping coming 🎶 congrats to the team on this amazing news!
Today we announced a $350M Series A, led by Disruptive with planned participation from NVIDIA, valuing Groq at $3.5B. Together with June's $650M, that's $1 billion raised in two months. Inference is becoming the largest and most critical layer of AI infrastructure and it's what we do better than anyone. The injection of capital will support those seeking usage of medium and larger sized clusters of NVIDIA accelerated computing for training and inference. Groq expects to scale from 54 megawatts to 200+ megawatts in 2027. Today, more than 6 million developers, Fortune 500 enterprises and thousands of AI-native companies build on Groq, generating trillions of tokens every week. Full announcement: https://lnkd.in/gsMg3kwU
-
Today we announced a $350M Series A, led by Disruptive with planned participation from NVIDIA, valuing Groq at $3.5B. Together with June's $650M, that's $1 billion raised in two months. Inference is becoming the largest and most critical layer of AI infrastructure and it's what we do better than anyone. The injection of capital will support those seeking usage of medium and larger sized clusters of NVIDIA accelerated computing for training and inference. Groq expects to scale from 54 megawatts to 200+ megawatts in 2027. Today, more than 6 million developers, Fortune 500 enterprises and thousands of AI-native companies build on Groq, generating trillions of tokens every week. Full announcement: https://lnkd.in/gsMg3kwU
-
Groq is now an NVIDIA Cloud Partner. A milestone for the team and validation of what our customers already experience: Groq runs AI infrastructure to the highest standard. "Inference is becoming the largest and most critical layer of AI, and we intend to run it better than anyone." — Adam Winter, CEO of Groq https://lnkd.in/gzt4CGwi
-
We've successfully raised $650 million in new capital to scale our inference cloud. We have strong conviction that inference will scale to be the largest and most consequential part of AI, and we've built the system that delivers the best value-to-performance for the most demanding AI applications. Moreover, industry heavyweights are joining our team. Alan Rice is joining us as COO, bringing 25+ years in hyperscale data-center operations from xAI and Meta. In addition, cloud experts Sinclair Schuller (CTO) and Rakesh Malhotra (CPO) recently joined as well. The pair are longtime partners who worked together at Apprenda, the enterprise cloud platform Schuller founded and later sold to Atos. They then co-founded Nuvalence, a software-engineering and digital-transformation firm acquired by EY in 2024. To read the full story, please read our announcement here: https://lnkd.in/gSH-XaNp
-
What does it look like when a fintech company gives everyone access to a personal financial advisor? For Stash, it looks like the Money Coach—an AI-powered agent network delivering real-time, personalized financial guidance to 1.4 million users. And with Groq powering the engine, it has never worked better. ⚡ ~37% decrease in average request time 📈 ~10% increase in total requests processed 💰 Significant cost reduction from ~$40K/month on OpenAI Stash is now faster, smarter, and more compliant than ever, and they are just getting started. Read the full story 👇 https://lnkd.in/gJymeuhX
-
-
Congrats to NVIDIA on their newly integrated NVIDIA Groq 3 LPU!
The NVIDIA Vera Rubin platform is opening the next frontier of AI. At #NVIDIAGTC, we announced that the NVIDIA Vera Rubin platform's seven new chips are in full production to scale the world’s largest AI factories. The platform brings together the Vera CPU, Rubin GPU, NVLink 6 Switch, ConnectX-9 SuperNIC, BlueField-4 DPU, and Spectrum-6 Ethernet switch, as well as the newly integrated NVIDIA Groq 3 LPU, to operate together as one incredible AI supercomputer to power every phase of #AI. Learn more: https://nvda.ws/4sLSysW #NVIDIAVeraRubin
-
-
How will AI transform the future of law? Groq's COO and CLO, Claire Hart, shares her unique perspective sitting at the intersection of both worlds.
Claire Hart told me she recently asked a group of law firm associates how they were using AI, and they said they’re not allowed to. What??? So of course, I had to have her on Pearls On, Gloves Off: The Legal Innovation Podcast. Claire is COO, CLO, and Board Member at Groq, operating right at the center of AI infrastructure and innovation. She’s spent her career at companies like Google, Blizzard Entertainment, and Genies, so she knows exactly what real technology shifts look like from the inside. And in this conversation, she does not hold back. We get into: - Why she says she’d be “horrified” if her law firms aren’t using AI - The growing disconnect between client expectations and law firm behavior - How to build great teams - What young lawyers need to do right now to stay relevant - And how legal teams should be thinking about talent, ops, and technology in this moment But what I loved most about this conversation is Claire herself. She’s thoughtful, practical, and genuinely funny… while also being incredibly clear-eyed about where things are going. Find this episode anywhere you listen to your podcasts. On Apple Podcasts here: https://lnkd.in/gjbukpJt On Spotify here: https://lnkd.in/gcRY3bkP
-
-
What if you could ship code faster than ever and still catch every bug before it reaches your customers? That's exactly what Autonoma AI built. Founded by four ex-Google engineers, Autonoma AI uses AI agents to automate software testing end-to-end, simulating real users, analyzing UIs with computer vision, and validating every step of your application, automatically. The results? Staggering: ✅ Regression testing slashed from 3 days → single-digit minutes ✅ Hundreds of thousands of tests run every week ✅ 20+ enterprise clients across fintech, retail, and tech ✅ 5,000+ installs — roughly 10x their nearest competitor But speed only works if the entire experience feels instant. That's where Groq came in. When Autonoma moved their inference workloads to GroqCloud, time-to-first-token dropped from seconds to milliseconds. Users can now describe a test in plain English and watch it come to life in real time—no lag, no friction, no waiting. Learn more about how speed and quality are no longer a trade-offs for Autonoma. They're the same thing. 👇