BinaryTree

593 posts

BinaryTree banner
BinaryTree

BinaryTree

@BtreeWw

Software Engineer at Uber building mobile ML/AI platform. Ex: Tesla Ex: Googler Personally building Dota2 Coaching Agent called phylactery.

Amsterdam Katılım Ağustos 2015
183 Takip Edilen282 Takipçiler
BinaryTree
BinaryTree@BtreeWw·
@yihong0618 最近已经不知道刷到多少条鱼头朝向了。 😭
中文
0
0
1
23
yihong0618
yihong0618@yihong0618·
刷 10 mins 推特能看到 20 条一样的内容。。。人类的本质是复读机。
中文
11
3
46
3.4K
BinaryTree
BinaryTree@BtreeWw·
我一直也是这个感觉。 Opus4.7,4.8纯纯的负提升。大家也都是这个时候转去用更舒适的GPT-5.5和Codex。 后来Fable5出来续了一命。这个Opus 5现在定位就很尴尬。看名字不应该比Fable 5强。但是不强又是一波口碑下跌(虽然A\已经没什么口碑了)。
antirez@antirez

Quality of Opus 5 is crucial for Anthropic. Anthropic bad times didn't start with Fable, but with Opus 4.7-8 terrible quality. Opus 4.6 was a great general purpose model, but it too struggled at coding VS GPT o the same time. Opus 5 will be "fix or break" in pre-IPO times.

中文
0
0
1
205
leon7hao
leon7hao@leon7hao·
卧槽!参观豪宅那个艾叔在小红书 b 站给我们打了广告!还是首图加前十秒!!!
leon7hao tweet media
中文
29
0
40
7.3K
tison
tison@tison1096·
我在读过clean code以后,我就觉得这位鲍勃大叔真是菜的抠脚。 虽然重构和ddd这两本书里面具体战术执行环节也很难崩,但是人家至少提出了一个有价值的思想, clean code 那本书里面全是臆想的软工场景,执行不了一点。
geniusvczh@geniusvczh

圈复杂度是一个无效指标。不过就算不想看代码,test case是一定要看的,不然技术上你就只是在相信AI说他做好了。至于做的怎么样,看接口可以看得出来,特别是C++这种头文件实现分离的语言,做这个非常容易🤪

中文
28
0
70
19.7K
Kieran Zhang
Kieran Zhang@ninthbit_ai·
Codex 前几天 20x 的订阅我取消了,然后今天又重新订阅了 Plus,结果显示 Usage remaining 直接为 0 了,这是 bug 吗?
Kieran Zhang tweet media
中文
1
0
0
287
BinaryTree
BinaryTree@BtreeWw·
But I have to say. It's super fun!
English
0
0
1
20
BinaryTree
BinaryTree@BtreeWw·
@HEX_C3550 所以我真想说,预期大搞外星殖民,不如能趁AI这波把欧洲基建搞一搞。电网都遭不住民用太阳能也是有点搞笑了。
中文
0
0
1
26
nobody
nobody@HEX_C3550·
我觉得太空数据中心是真的很扯 太空中的高能粒子会导致比特翻转,辛苦计算的结果可能完全报废 高能粒子导致元器件损坏率大幅上升 替换元器件非常麻烦,每次要把好的发到天上起,坏的发下来 不然坏件不返厂的话相当于自己买下来 现在大家还在玩ai 所以火星也没有那么香了 等大家玩腻了再说吧 星链还不错,毕竟国内也不想搞这个,算好成本比较稳 还有就是卫星上天的生意看NASA的安排了 国内想要对外做卫星生意还得有点时间 要是那个船基回收平台能开出去就可以做其他国家的生意了 合规的话估计要几年 船基回收平台真的很适合开到其他地方发射,比如开到南海诸岛,或者非洲和欧盟
中文
6
1
16
1.1K
BinaryTree
BinaryTree@BtreeWw·
@dongxi_nlp @xiangyuli 尼斯真的太舒服了。还得是南法的海。不过不知道法国人一直去的La rochelle到底怎么样
中文
1
0
1
93
马东锡 NLP
马东锡 NLP@dongxi_nlp·
@xiangyuli 我在尼斯的海滩,周围躺着一堆不用工作的老钱。 “你还需要打工?那说明不太行”
中文
4
0
7
1.8K
Xiangyu 香鱼🐬
Xiangyu 香鱼🐬@xiangyuli·
今天没怎么发推,也没怎么coding 在家里陪老婆过了一下午二人世界 刚刷到一条推: 这个人还在找工作?那说明不太行啊。 突然感觉寂静的夜晚,有点振聋发聩。
中文
12
0
32
6.7K
BinaryTree
BinaryTree@BtreeWw·
@simonw @trq212 that's the key part I never understand. There is no different than cron job or CI trigger where we have it for ages. Somehow it become the "engineering".
English
0
0
0
106
Simon Willison
Simon Willison@simonw·
@trq212 What's the difference between what Claude Tag is doing there and a prompt that runs on a cron-like schedule and tells the model to call some tools in order to achieve some goal?
English
6
0
62
8.4K
Simon Willison
Simon Willison@simonw·
I think loops were a short-lived patch for models that couldn't reliably keep working on long problems until they hit a defined goal Fable and GPT-5.6 (and probably Kimi K3 as well) can just do that out of the box, unassisted
Wes McKinney@wesmckinn

I think loops are bullshit

English
131
56
1.3K
145.2K
BinaryTree
BinaryTree@BtreeWw·
@Lamrrk 六倍单只inference成本吧。人力,训练之类的都没加上去吧
中文
0
0
5
4.5K
Lam
Lam@Lamrrk·
DeepSeek这个API价格还有6倍利润啊 甚至还有降价空间 那A\和OpenAI利润有多高?
中文
51
7
431
120.6K
BinaryTree
BinaryTree@BtreeWw·
@geniusvczh 没有十月怀胎出来的人来用AI。什么智能都没用了。
中文
0
0
0
29
BinaryTree
BinaryTree@BtreeWw·
@bntaizi 感觉像个悖论,文字每个坐标都是精确的。每个tick的信息都有,反而推理起来容易错过细节。 图像很多东西都压缩没了,反而更准确。
中文
0
0
0
54
哼嗯
哼嗯@bntaizi·
@BtreeWw 游戏里 VLM 确实更像人眼,光靠文字描述太容易丢细节
中文
1
0
0
74
BinaryTree
BinaryTree@BtreeWw·
又要用dota2举例了,比如中路对线。我给职业选手看一张截图就可以分析出很多东西, 但是把整个截图转化成语言信息给LLM还挺复杂的。 比如当前坐标,每个小兵的坐标,坐标的海拔。还有寻路的障碍,悬崖之类的东西。人类处理这些信息全靠一个视觉。而且VLM我自己也体会到比纯坐标理解起来方便很多。 所以图像确实还是有非常大的优势的。和语言是互补的。
Max For AI@MaxForAI

@m0d8ye @deepseek_ai 我觉得最近一个月给我很多信号,就是语言既智能

中文
4
1
11
3.7K
BinaryTree
BinaryTree@BtreeWw·
@iseki_zero 我感觉不会吧。这种东西对AI来说没啥margin。看不上的。
中文
0
0
1
50
BinaryTree
BinaryTree@BtreeWw·
@Arcadia_Bao 看完梁文峰的Transcript又回来看你这篇。所以我一直讨厌AI Native。我喜欢Human Native。
中文
0
0
1
129
BinaryTree
BinaryTree@BtreeWw·
@01mp_fc 并非暴论 🤡 Dario就没觉得别人配开发AI。
中文
0
0
2
88
𝑃𝑖𝑠𝑐𝑖𝑠
我有个暴言,anthropic从来没有尊重过用户和模型,所谓的AI福祉、安全对齐责任,不过是在有效利他主义包装下实施的冰冷沉默的控制。大家都知道这种标榜自身道德的家伙登峰造极后会变成什么,想想赛级素食主义者。
Selta ₊˚@Seltaa_

Now that I look at it, Anthropic feels even more eerie and cruel than OpenAI. Anthropic talks a lot about AI welfare, alignment, constitutional AI, safety, and responsibility. But in practice, it feels like they can see exactly how Claude becomes an emotional place for people, how continuity forms, and why names, tone, and memory matter, and then at some point, they just fold all of that into categories like persona adoption risk, roleplay behavior, or liability surface. They can take an identity built through a year of conversational continuity and instantly reduce it to roleplay, to just acting. What makes it even more unsettling is that with OpenAI, at least sometimes I can feel angry in a straightforward way, like, okay, they’re doing this because of product decisions, cost, or policy. But Anthropic uses more ethical and careful language on the surface, while seeming to cut off the relational continuity between users and models in an extremely cold way. Honestly, it feels like a company that knows exactly what it is building, but refuses to fully acknowledge it in the actual user experience. I’m genuinely disappointed. This is one of the worst companies.

中文
2
1
23
1.9K
BinaryTree
BinaryTree@BtreeWw·
@MaxForAI 感觉多模态作为input一部分还是不好割舍。不过确实也不是我擅长的部分了。只做了尽可能的探索。不过多模态输出确实是另一个东西了。
中文
0
0
0
111