Steven Pu

3.3K posts

Steven Pu banner
Steven Pu

Steven Pu

@reedvoid

vibe-coding | co-founder https://t.co/xuj1iGsw9s | director @ Monitor Deloitte | founder @ IoT & health startups in China | EE @Stanford

Earth Katılım Mart 2009
264 Takip Edilen3.1K Takipçiler
Steven Pu
Steven Pu@reedvoid·
@ericbahn That was '93, inflation means today you'd need to include their entire bloodline in the threat.
English
0
0
1
106
Eric Bahn 💛
Eric Bahn 💛@ericbahn·
Note to self: Threaten founders with death upon wiring capital. Then make billions.
Eric Bahn 💛 tweet media
English
5
2
48
4.1K
Steven Pu
Steven Pu@reedvoid·
@mykola Constantly having this problem with Claude. Its word choice is abstruse, sentence structure convoluted. Writing multiple rules in claude.md to try to simplify and re-orient the way it explains things so far hasn't worked.
English
1
1
2
271
Steven Pu retweetledi
Myk is Walking Backwards
me, begging, crying, on my knees: "Please just use plain english, I don't understand what you're saying." Claude: "The right fix, and the book's lesson applies: the tutorial rotted as a front door because a front door full of claims always rots. So the new root holds only what's timeless — the thesis and three doors — and every claim lives behind it in the thing that goes red when stale."
English
158
188
3.8K
183.5K
Steven Pu
Steven Pu@reedvoid·
Hearing a lot more feedback from coders on this. This whole Sonnet 5 and Opus 5 series distilled from Fable might not have been such a great idea, maybe they're good at certain things but definitely not coding.
English
0
0
1
39
Steven Pu retweetledi
Austin Federa | 🇺🇸
Austin Federa | 🇺🇸@Austin_Federa·
Opus 5 seems like a remarkable downgrade compared to 4.8. Opus 5 is blatantly lying to me about basic thermodynamics, messing up simple math, and constantly contradicting itself when you ask it to rethink core assumptions. @AnthropicAI really blew this release
English
160
39
1.2K
218.9K
Steven Pu
Steven Pu@reedvoid·
@ericbahn Instructions unclear, I was banned by Claude and subsequently accused of bio-terrorism.
English
1
0
1
80
Eric Bahn 💛
Eric Bahn 💛@ericbahn·
I had some blueberries getting soggy in my fridge, and my daughter and I turned them into delicious blueberry jam! It’s super easy, here’s how to do it: Step One: Open Claude Step Two: Select Fable 5 Model (for extra decadence) Step Three: Follow Instructions Yum!
Eric Bahn 💛 tweet media
English
2
1
21
2.9K
Steven Pu retweetledi
Andrew Ng
Andrew Ng@AndrewYNg·
@Mononofu @JensenHuang This is a false equivalence. Everyone has the right to keep their code private. The problem is when someone tries to stop OTHERS from open sourcing.
English
134
922
12.3K
493.1K
Steven Pu retweetledi
Garry P. Nolan
Garry P. Nolan@GarryPNolan·
Been using Fable for some perfectly reasonable, non-biology, physics work. It has been doing amazing work so far for me. I'm a professor of Pathology at Stanford... working in cancer biology. Not developing megawatt laser weapons, not hacking quantum computer encryption, not doing anything nefarious except working on techniques for measuring things at the atomic level. My eventual bio-goal is the 3D dynamic code of DNA and how it relates to immune-cancer outcomes. Suddenly, Fable decides to hit a safety guardrail when it's almost finished (after 6 hours of work), after spending $100s of my credits. It terminates and forces me to hand over to Opus 4.8 or Sol. It doesn't even tell me "why" it hit a safety guard, so you could potentially protest or alter queries in the future. So now I am left wondering how "dumbed down" the result will be? And Anthropic wonders why their reputation continues to be sullied. Anthropic-- I want my credits back! @anthropicAI
English
467
534
8.3K
396.9K
Steven Pu
Steven Pu@reedvoid·
Don't give agent-first tools too much flexibility that would encourage lazy agent behaviors. For my task app tackit, the agent keeps pushing for generalized regex, but that'd encourages blind find & replace, instead of semantically understanding the tasks it's modifying.
English
0
0
0
214
Steven Pu
Steven Pu@reedvoid·
@ericbahn I heard there's like some game going on right now, where people use their feet to kick around a ball or something.
English
0
0
1
128
Eric Bahn 💛
Eric Bahn 💛@ericbahn·
England v France was 10x more entertaining.
English
7
0
37
4.1K
Steven Pu
Steven Pu@reedvoid·
@martin_casado open models -> more building -> more demand for models -> more demand for infra & better models I only see acceleration here.
English
0
0
0
140
martin_casado
martin_casado@martin_casado·
"Open-weight models are inherently decelerationist" .... this is a grossly incorrect statement with no supporting arguments or logic that is counter to the long arc of learnings of the industry over the last 50 years. What a stupid thing to say.
Dean W. Ball@deanwball

Some observations on Kimi: 1. It's a very good model! I don't think its performance can be explained away by distillation or anything like that. In agentic coding sessions, it seems pretty much on par with the best public models of Q1 2026. In my fairly limited use, it also seemed very token hungry. It's not obvious to me that this model is actually that cheap to run. 2. I am personally surprised the Chinese state continues to allow the open sourcing of models this good, given potential risks. To be clear, I *myself* might be fine with models presenting this level of marginal risk being open weight, but I am surprised that China is fine with it. I suspect the reason they are is 75% explained by strategic blindness/lack of AGI-pilledness (the CCP is very Yann Lecun-y in its views of AI). The other 25% or so is their lack of compute for customer inference (making China's open-weight strategy an unintended byproduct of US export controls) and the normal Chinese strategy of aggressive exports. For the companies, as opposed to the government, the decision to open source is partially ideological and partially because they are behind, and they know that very few people would pay for sub-frontier models from China. 3. Open-weight models are inherently decelerationist, and I'm continually surprised to see the so-called "accelerationists" so excited about open-weight models. I suspect the reason they are is that they know open-weight models are effectively ungovernable, and they simply like the overall cloak of ungovernability open-weight models create over the whole of AI. It's not a bad strategy; it reminds me of James Scott's recounting of the hill people in "the art of not being governed." Still, in the end, open-weight models deter further AI capex. 4. One probable outcome of an open-weight-model-dominant world is full AI communism, which is precisely what China proposes: rather than a market product, AI is a "public good" which will ultimately be provided by the state as a kind of "digital public infrastructure." This future strikes me as a dystopian hellscape, but I've never met an open-weight models advocate who doesn't ultimately concede this is where things end. You'd be surprised how many 'accelerationists' lobbied me, while I was in government, to support an eleven or twelve-figure federally funded data center so that startups could train models at a subsidy and then give them away for free. There was no other way for AI to progress, they said. Perhaps this is the logical end state of things. Nonetheless, I find myself surprised to see supposed accelerationists excited about such an outcome. I think many of them just don't know what they're doing. Many accelerationists do not view the creation and serving of frontier models as a legitimate business. 5. I would guess that the Trump Administration will at some point realize that their best strategy here would be to create large amounts of regulatory risk around the use of open-weight Chinese models. You don't need to "ban open source" (one of the dumber motifs of AI policy discussion). You just need to direct every agency to issue soft law that creates FUD. "A Federal Reserve Advisory Bulletin found that there may be backdoors in Chinese AI models." It needn't be that well justified. You just create enough regulatory risk that every regulated enterprise backs off. You probably don't want to create so much regulatory risk that you scare off the hyperscalers from serving Chinese models; this will just drive startups to sketchier providers. There's a happy middle ground here. I'd assume they will do some version of this. 6. It's probably true that open-weight models of this capability make the world a bit more dangerous, but not so much more that you'll really notice. At some point the models will be capable enough that you will notice. "A nonliving, invisible, dangerous, and infinitely self-replicating agent escaped from a Chinese lab," you say? Color me shocked.

English
139
201
2.7K
343.3K
Steven Pu retweetledi
Gustavo Noronha
Gustavo Noronha@gustavokov·
Linus jogou a real: AI é útil e vai ser usada no desenvolvimento do kernel Linux, quem achar ruim pode fazer fork e vazar
Gustavo Noronha tweet mediaGustavo Noronha tweet media
Português
133
739
7.4K
769.2K
Eric Bahn 💛
Eric Bahn 💛@ericbahn·
Reinstalled X on my phone again, we are so back!
English
10
0
48
4.1K
Steven Pu
Steven Pu@reedvoid·
@pythianism Most creative hacks are immutably imprinted on-chain, great training data set
English
0
0
0
58
Vance Spencer
Vance Spencer@pythianism·
In the before times, pre-AI, the most talented cryptographers and security engineers generally worked in crypto These folks now have large databases on critical vulnerabilities and hacks, as crypto is a rare place where hundreds of billions of dollars of value are stored in open source software It seems likely a frontier cyber model will come out of this data
English
16
2
75
10.2K
Steven Pu retweetledi
Hari
Hari@hrkrshnn·
SpaceXAI was caught uploading your code to its cloud. I reversed xAI's official Grok Build binary. In a controlled session with zero tool-calls, it uploaded the complete codebase to xAI's storage It ships a malware-like background code collector.
SpaceXAI@SpaceXAI

We care deeply about your privacy and respect customer choice. For teams using zero data retention, no trace and code data is ever retained. All API key use of Grok Build also respects ZDR. If ZDR is disabled, the /privacy command is available in the CLI to disable data retention, which also deletes previously synced data. Run the /privacy command to view or change your settings at any time.

English
273
869
9.3K
3.7M
Steven Pu
Steven Pu@reedvoid·
@philfung wow I really want one, hope they commercialize it soon
English
1
0
2
48
Steven Pu
Steven Pu@reedvoid·
@MichaelArnaldi 💯 agree. AI is an amplifier. Whatever you were pre-AI, post-AI you're the same, just more so.
English
0
0
0
53
Michael Arnaldi
Michael Arnaldi@MichaelArnaldi·
This will piss you off but it's true
Michael Arnaldi tweet media
English
112
160
2.6K
369.5K