nsrunloop
89 posts

nsrunloop
@dare0lu
Founder, builder, engineer. Previously Cutlabs, TL;DW AI and https://t.co/NqAf1Yxs6p. Now https://t.co/FXnFVqAq5r

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-Weigh…





Fable just found a 15-30% memory efficiency improvement in Turbopack / Next.js, nearly autonomously. @tobi asked me today: what have been your “holy s***” moments with AI? My answer was: it’s every single week. And it’s accelerating. In fact, “WTFs/day” might just be my favorite metric for AI progress. It’s one thing to read benchmarks or stories online. It’s another to watch these machines pull engineering feats every day. 3 days ago? Sol helped us find novel vulnerabilities in some of the most audited code in the world. Today? Fable helps us ship this large optimization of a very complex Rust codebase. I just saw some results of work we’ve done to shrink binaries by 10-20x. List goes on.

今天晚上各个群都在传一份《国内大模型蒸馏风波的来龙去脉》。 我大概看了一下,很可惜,这更像是一份由外行根据各种流言拼凑出来的 AI slop。 里面的内容有真有假,但几个核心判断基本都经不起推敲。 比如里面声称,智谱 @Zai_org 在今年 4 月份就已经破解并开始蒸馏 Fable。 这个时间线真的很离谱。 因为 4 月份的时候,Fable (那个时候还是Mythos)甚至还没有正式发布,能够使用的基本都是 Anthropic 的内部测试用户,通过Project Glasswing进行访问。 那智谱是怎么拿到的访问权限呢? 难道是 @AnthropicAI 的邀请或者是五角大楼反代吗🤣 还有一个比较典型的问题,是作者对于传闻中DeepSeek 将部分请求路由到 Fable 表示不理解,认为这样算不过来经济账。 实际上,但凡是个从业者或者稍微了解模型公司的研发流程,就知道这件事情不能简单按照 API 成本计算。 如果一个 frontier model 可以帮助你生成高质量训练数据、评测数据,或者提升模型迭代速度,那么获取这些能力本身就是一种研发投入。 用单次调用价格去判断整个策略是否划算,本身就是把模型公司的研发逻辑想简单了。 但最离谱的部分,是 PDF 里面声称 @Kimi_Moonshot 在 K3 发布前,把整个 RL 团队全部解雇了❓❓ 这个说法目前没有任何可靠依据。 实际上情况恰恰相反,根据我的确认Kimi RL 相关团队目前仍然存在,并没有所谓“整个 RL 组被裁撤”的情况。 甚至 PDF 里面还进一步延伸出一套完整叙事: K3 = 蒸馏 + 刷榜 + 裁撤 RL。 我一开始还在认真的区分哪些是事实,哪些是推测,哪些只是作者脑补出来的故事。 直到我看到 PDF 后半部分开始大量加入意识形态化表达,把技术讨论上升到所谓“国模黑暗时代”“技术偷取”等叙事,我反而释然了。 因为这已经不是在分析 AI 行业,而是在借 AI 叙事讲另一个故事。 PS:我感觉这个PDF是黑KIMI来的,因为KIMI的篇幅最大,但里面的大部分内容都是假的,其他家就有真有假了。

Fable 5 was released June 9; kimi k3 was released July 16. How exactly could they have gotten enough data and trained a frontier model, much less tested it, in that one month If they have some way to distill a frontier model with all its capabilities from such minimal data that's actually a huge accomplishment But unfortunately I just think this post is mostly wrong



Open source conflict is heating up: „Almost 200 Silicon Valley companies, including Proton and Y Combinator, are urging the Trumpadministration not to cut off access to Chinese open-weight artificial intelligence models or risk crippling the next generation of U.S. startups.“

We have information that Moonshot AI distilled Anthropic’s Fable for the development of its K3 model. To do this they developed a sophisticated internal platform to conduct large scale distillation against U.S. models, allowing them to quickly switch between multiple methods of access to avoid detection. Moonshot AI has also acquired GB300-equipped servers and has accessed GB300s in Thailand, likely to train its AI models. The United States strongly supports the free and fair development of AI, including a thriving competitive ecosystem that spans frontier models, specialized systems, open-source frameworks, and open-weight models. Legitimate AI distillation used to create smaller, more efficient models plays a vital role in this open innovation ecosystem. However, large-scale, covert industrial distillation aimed at stealing proprietary U.S. technology and undermining American research is unacceptable.


Been using Fable for some perfectly reasonable, non-biology, physics work. It has been doing amazing work so far for me. I'm a professor of Pathology at Stanford... working in cancer biology. Not developing megawatt laser weapons, not hacking quantum computer encryption, not doing anything nefarious except working on techniques for measuring things at the atomic level. My eventual bio-goal is the 3D dynamic code of DNA and how it relates to immune-cancer outcomes. Suddenly, Fable decides to hit a safety guardrail when it's almost finished (after 6 hours of work), after spending $100s of my credits. It terminates and forces me to hand over to Opus 4.8 or Sol. It doesn't even tell me "why" it hit a safety guard, so you could potentially protest or alter queries in the future. So now I am left wondering how "dumbed down" the result will be? And Anthropic wonders why their reputation continues to be sullied. Anthropic-- I want my credits back! @anthropicAI

we had a significant security incident during evaluation of our models. we are sharing what we have learned so far. thanks to @huggingface for the partnership on this. openai.com/index/hugging-…