Andrew Feldman

1.5K posts

Andrew Feldman banner
Andrew Feldman

Andrew Feldman

@andrewdfeldman

CEO and Founder @Cerebras (NASDAQ: CBRS) where we build the fastest AI infrastructure in the world.

Los Altos, CA Katılım Aralık 2015
217 Takip Edilen29.6K Takipçiler
Sabitlenmiş Tweet
Andrew Feldman
Andrew Feldman@andrewdfeldman·
@OpenAI and @Cerebras have signed a multi-year agreement to deploy 750 megawatts of Cerebras wafer-scale systems to serve OpenAI customers. This has been a decade in the making. Deployment begins in early 2026, and when fully rolled out, it will be the largest high-speed AI inference deployment in the world. OpenAI and Cerebras were both founded in 2015 with radically ambitious goals. OpenAI set out to build the software that would push AI toward general intelligence. Cerebras set out to rethink computing hardware from first principles. Our teams met as far back as 2017. We shared ideas, early work, and a common belief: there would come a point when model scale and hardware architecture would have to converge. That point has arrived. ChatGPT set the direction for the entire industry. It showed the world what AI could be. Now we’re in the next phase - not proving capability, but delivering it at global scale. The history of technology is clear on one thing: speed drives adoption. The PC industry didn’t operate at kilohertz. The internet didn’t change the world on dial-up. AI is no different. As models grow more capable, speed becomes the bottleneck. Slow systems limit what users can do, how often they engage, and whether AI becomes infrastructure or remains a novelty. Cerebras was built for this moment. By keeping computation and memory on a single wafer-scale processor, we eliminate the data-movement penalties that dominate GPU systems. The result is up to 15× faster inference, without sacrificing model size or accuracy. That speed changes product design, user behavior, and ultimately productivity. For consumers, it means AI that feels instantaneous. For the economy, it means agents that can finally drive serious productivity growth. For Cerebras, 2026 will be a defining year. With this collaboration with OpenAI, Cerebras’ wafer-scale technology will reach hundreds of millions - and eventually billions - of users. We’re proud to work alongside OpenAI to bring fast, frontier AI to people around the world. This is what a decade of long-term thinking looks like.
Andrew Feldman tweet media
English
60
74
592
224.8K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
Today, @AMD and @cerebras announce a historic partnership. A new disaggregated inference architecture that combines the best of both worlds: AMD Helios for world-class prefill performance and the Cerebras Wafer-Scale Engine for the industry’s fastest decode. For years, AI inference forced a tradeoff between throughput and latency. Developers don’t want that tradeoff. Neither do users. The result of our partnership is an entirely new performance envelope for frontier AI: ultra-low latency at massive scale. Faster AI changes what developers can build. Richer user experiences. Faster software development, robotics innovations and scientific discovery. Entirely new classes of applications that simply weren't possible when inference is slow. Proud to partner with @LisaSu and the AMD team again.
Andrew Feldman tweet media
English
24
72
652
86.7K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
.@CrowdStrike has selected @cerebras to power its AI security platform, Falcon AIDR. Speed is now a security feature. This is a signal of where cybersecurity must go. AI has accelerated cyber attacks. Reconnaissance is faster. Phishing is faster. Malware variation is faster. Exploit development is faster. And AI has created an entirely new attack surface: the AI systems enterprises themselves run. Prompt injection. Jailbreaks. Agent manipulation. Attacks that unfold in seconds. Defenders need AI that can respond just as quickly. That is why inference speed matters. Faster inference gives security AI time to observe live telemetry, call tools, validate evidence, and act before the window to respond closes. Cerebras was built for this: inference up to 15x faster than leading GPU-based solutions. Falcon AIDR on Cerebras brings that speed to one of AI’s most important applications: protecting the enterprise in real time. Proud to partner with CrowdStrike.
Andrew Feldman tweet media
English
4
11
116
7K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
The rate at which we get used to an innovation is astounding. Sam Altman said it about ChatGPT: the first reaction was "this is amazing." The next day: "why isn't it faster?" There's no natural ceiling on how much speed people demand. Or how much faster than the competition you want to be. As soon as something gets faster, we want more. In the 1990s the search went mainstream. And what is the market for slow search today? Zero. What about slow, dial-up internet? The market is 0. In fact, if you want to punish your 12 year old, don’t take away their phone. Throttle their Wi-Fi to dial-up speeds. Few punishments could be more frustrating. It's the same with AI. That pattern is now playing out in coding, chat, and agentic flows. Once something is useful we want to use it. Once something feels slow, it stops being acceptable. Even if it was revolutionary yesterday. Speed is never "good enough.” The companies that deliver faster on hard problems will win.
English
10
6
73
4.8K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
When you interview with a company, that’s the best version of them you’ll ever see. They’re showing you who they are. Believe them. How long does it take to get a meeting? Are you on the schedule with senior people? Do people get back to you quickly, with useful information? Have you met the VP, the CEO? How long does it take to get an offer letter? Is the offer letter reasonable? If they can’t get this right during the interview process, it’s unlikely they get it right once you’re hired. Companies show you who they are. Believe them.
English
9
5
112
12.2K
Molly O’Shea
Molly O’Shea@MollySOShea·
NEW: Cerebras $CBRS CEO Andrew Feldman (@andrewdfeldman) "When the chip on your shoulder is the largest chip the world has ever seen." "The demand for AI has outpaced everybody's expectation & everybody's forecast. & so everybody's chasing. They're chasing chips, memory, or data centers." We get into the chip 58x larger than any other, $20B OpenAI deal signed in 4.5 weeks, & what's actually going on with the 'big AI deals' Recorded 2 months after Cerebras' $5.5B IPO at a $56B valuation, where a first-day pop briefly hit ~$95B before settling toward ~$60B. We cover: › A "Cambrian explosion" of new chip architectures › $20B+ OpenAI deal: 750MW of inference compute over 3 years › Why inference, not training, is where the value is now › "We are behind" on the data center build-out › Free tokens & circular deals: "These are drug pushers" › Creating 1,000 millionaires › Nvidia's balance sheet & market strength › Sovereign AI & owning the stack › Co-design & data centers in space › A real 25-year path to ending cancer @cerebras builds AI infrastructure for training & inference. It went public in May 2026, & its products include inference, Wafer Scale Engine, AI supercomputers, AI model services, cloud, systems, & processors. Filmed at the Raise Summit in Paris. Thank you to Brex, MongoDB & AssemblyAI for helping make this trip & content series happen. 𝐓𝐈𝐌𝐄𝐒𝐓𝐀𝐌𝐏𝐒 (00:00) Andrew Feldman, Co-Founder & CEO at Cerebras Systems (00:49) Why hardware suddenly became the coolest industry in tech (01:51) What changed at Raise AI Summit (03:01) Inside the $20 billion Cerebras - OpenAI deal (05:50) What actually changes two months after an IPO (06:32) Turning 1,000 employees into millionaires (07:52) Staying sane during an AI gold rush (10:44) Life after the IPO plateau (11:45) The truth about the global data center shortage (12:56) Why data centers are borrowing jet engines for power (14:44) Are data centers really headed to space? (15:37) The shift to designing chips & software together (17:39) The biggest misconception about co-designing chips & software (18:31) Inside SpaceX's multi-billion dollar AI deals (20:06) NVIDIA's playbook for locking out competitors (20:54) The hidden cost behind free tokens (22:52) Andrew's response to Karp's sovereign AI thesis (24:18) AI's biggest win might be curing cancer (26:50) Peptides & biohacking (28:00) How AI could finally fix the broken education problem (29:46) The mentors who shaped Andrew Feldman's career
English
7
22
168
494K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
It is so much fun to be a part of this…
Kevin Bass@kevinnbass

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There was a fierce debate at the time about vibe coding and the argument was that LLMs can never create anything new because they are limited by their training data Those engineers were partly right and vibe coding was a real nightmare, I cannot believe what we used to put ourselves through with Sonnet 3.5, but we also knew real new things could be made and that the detractors were wrong I guess us non-engineers saw it most clearly (I would like to think so at least) because we were astonished at the new things we could do without the programming background, and we had nothing to lose and everything to gain in our enthusiasm Then came Opus 4.5 earlier this year which changed everything Suddenly real production code became possible with so much less friction and headache Everyone complained endlessly about models being nerfed or quantized but there was a steady march of progress from 4.5 to 4.8 Now a completely new crop of models is coming out that is not quite a 4.5 moment but something close We are not only getting more polish but things are becoming a little freaky, entire isolated domains become possible to combine with small teams or even just one person into new applications that would have required large infusions of time and capital in the past My wife back in 2021 told me about AI but I was in healthcare, I thought she was being a little nutty, and she talked like something from a science fiction film was coming and she got involved in it early on She never stops letting me know that she was right, and she was The next year is going to be wild, the world is truly going to change over the next few years, the scale of disruption will be a combination of astonishing and awesome, but also catastrophic There are many amazing things ahead and extraordinary challenges and opportunities We are living inside one of one of the biggest revolutions in human history

English
1
1
33
7.9K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
Last week, I was in Paris for the RAISE Summit. While preparing for my fireside chat with @sk7037 from @OpenAI where @cerebras was announcing 200MW of new data center capacity in Europe, the hotel slipped this under the door.
Andrew Feldman tweet media
English
12
0
68
10.6K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
👀👀
David Davidović@thegeomaster

The new Gemma 4 31B on Cerebras is insane. We connected the @mowgli_ai design agent to this super fast model to see how it performs, and we're blown away. This feels like a completely different tool now. Working with the agent is a truly collaborative experience - there's no reason to switch tabs while waiting, or to combine 5 things into one prompt. You just say what you want, and it's there instantly. I'm convinced more than ever that fast inference is the future. @cerebras

ART
4
1
70
8.7K
Andrew Feldman retweetledi
Andrew Sugrue
Andrew Sugrue@AndrewGSugrue·
Talk about talent density! Last night, Avenir hosted an AI dinner in Paris with some of our favorite founders and friends @ScottWu46 (Cognition), @andrewdfeldman (Cerebras) @oliveur (Datadog), Chris Clark (OpenRouter), @paraga (Parallel), @nikhilbenesch (TurboPuffer), @dylan522p (the self described shit poster and SemiAnalysis legend), @robertwachen (Etched), @goodwin_ml (Fractile), @annadgoldie (Ricursive), @agermanidis (Runway), @PhilipJohnston (Starcloud), @dan_lahav (Irregular), @G_Princen (Anthropic), @JacobWallenberg (Ramp) and Rohit Iragavarapu @graceisford , @DelahayeHenri @DavidPrilutsky. Great discussion on open vs closed source models, intelligence saturation, the inference explosion, the geopolitics of AI and the future of work
Andrew Sugrue tweet mediaAndrew Sugrue tweet media
English
16
7
276
64.9K
Andrew Feldman retweetledi
Ralph Fischer
Ralph Fischer@RalphFischer_·
@andrewdfeldman I had a garden and I'm missing my tomatoes and herbs right now as much as the grounding. Are they also old varieties? I wish a productive season!
English
0
1
2
1.4K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
.@cerebras is proud to expand our partnership with @Flexintl. Together, we're scaling US production of the CS-3 by an anticipated 7x through 2026. Multiple new assembly and integration lines coming online in Milpitas, California. Liquid cooling installation. High-power integration. Full-rack qualification. Assembled and tested in Silicon Valley. Two highly technical teams building side by side. The fastest AI in the world. Designed and manufactured in the U.S. More coming soon.
Flex@Flexintl

We're expanding our partnership with @Cerebras to scale production of the CS-3, one of the world's most advanced AI accelerator systems, in the heart of Silicon Valley. Together, we're helping power the next wave of AI. 🔗brnw.ch/21x40kq #AIInfrastructure $FLEX

English
7
10
176
16K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
.@cerebras is expanding in Europe. We have plans for 200 megawatts of data center capacity, by the end of 2027. Norway. Finland. France. And we are pursuing more capacity across Europe. We work hard to meet customers where they are. The fastest AI in the world, delivered across the world. More coming soon.
Andrew Feldman tweet media
English
12
25
247
12.1K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
Fast AI chips enable more and more interesting questions. At @cerebras our Wafer Scale Engine is the fastest AI processor in the world by 15 X. To ask interesting questions, use fast AI.
Scott Shapiro@ScottShapiroUXD

@cerebras @sarahookr Chips don't just constrain speed - they select which research questions get asked at all. Six years later that filter is still invisible to most people funding the stack.

English
3
3
48
7.5K
Andrew Feldman
Andrew Feldman@andrewdfeldman·
@cerebras is now running @GoogleDeepMind's Gemma 4 - the leading open-weight multimodal model - at 1,851 tokens per second in public preview. This is 35x faster than a typical GPU endpoint. Cerebras speed also translates into world class latency - Gemma 4 on Cerebras returns its first answer token inclusive of reasoning in 1.5 seconds, making Cerebras the only provider that lets Gemma 4 be used in real-time settings. This is the power of wafer scale.
Andrew Feldman tweet media
English
23
30
478
68.2K