T. Rebentisch

23 posts

T. Rebentisch

T. Rebentisch

@RebentischT

Katılım Ekim 2023
38 Takip Edilen7 Takipçiler
T. Rebentisch
T. Rebentisch@RebentischT·
@adonis_singh World peace is crazy lmao. For what AI can do about that, GPT3 already had that solved "everyone stop shooting each other", there you go.
English
0
0
0
30
adi
adi@adonis_singh·
this has to be ragebait
Gary Marcus@GaryMarcus

@midsusnight @skdh World peace, a million things in biology and medicine and climate change, batteries, etc etc etc I am not even saying Astra’s problems aren’t worth solving. I am saying people are overgeneralizing from “exceptional math performance” to “ominpotent” and that’s just plain wrong.

English
4
0
74
5.7K
Jun Song
Jun Song@jun_song·
The new DeepSeek is even crazier than expected. At just 284B parameters, it beats GLM 5.2 and comes insanely close to Opus 4.8. All while being 1/10 of the size and almost 100x cheaper. I knew innovation was coming, but I did not expect it to hit this hard. Pro is definitely going to crush Fable and Sol.
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTex

DeepSeek-V4-Flash is updated. «DSBench-FullStack is an internal full-stack development test set, and DSBench-Hard is an internal Coding Agent hard-problem test set» «The official release of DeepSeek-V4-Pro will follow soon» These scores are… quite a lot better than Preview.

English
79
161
2.1K
126.7K
Richard Johnson
Richard Johnson@RichJohnsonNFL·
C. Kirk suffered an injury in #49ers practice today. Sources say he’s stable. Please god.
English
387
4.7K
140.7K
7.7M
T. Rebentisch
T. Rebentisch@RebentischT·
@AdamHoltererer lol your „evidence“ are 2 benchmark scores where one of thems is even part of the other. K3 massively outperforms even Sol in multiple relevant domains AND is less annoying to work with.
English
0
0
0
11
BridgeMind
BridgeMind@bridgemindai·
Kimi K3 open weights are live and it is already running 3x faster. The weights dropped this morning. Fireworks is already serving it at 45 tokens per second with 1.44s latency. I have spent 11 days waiting 11 seconds for this model to say its first word. It was painful. The best open model ever made and I kept closing the tab. Not anymore. This is the part of open source people forget. You release the weights and the entire industry fixes your problems for you. Same day. Kimi K3 just became a real vibe coding option.
BridgeMind tweet media
English
72
65
1.6K
109.1K
T. Rebentisch
T. Rebentisch@RebentischT·
@jamiuxxvi MOTM is basically the only thing you need to see/know. MOTM in like 40% of games played lmao
English
0
0
0
707
Claude
Claude@claudeai·
Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price.
English
3.5K
7.5K
61.5K
23.6M
T. Rebentisch
T. Rebentisch@RebentischT·
@blader I guess Fable 5.1 will drop very soon, reclaiming noticable lead over Opus 5? Otherwise, the model has no place. Also, Mythos class SHOULD have the highest post-training potential.
English
1
0
0
1.6K
Jacob P
Jacob P@JacobPhillip801·
@AlienWisconsin @DavidOndrej1 You are correct, as things currently stand. But chips are getting much much better at an accelerated rate with all the demand and research being invested. The point I’m making is who’s to say in 2-3 years we can’t have lower end hardware run 1-2T parameter models?
English
1
0
0
14
BridgeMind
BridgeMind@bridgemindai·
Fable 5 is finally back. I burned through half my weekly limit before breakfast and I regret nothing. But this chart should terrify Anthropic. GPT 5.6 Sol Ultra: 91.9% on TerminalBench 2.1 Fable 5: 84.3% Even GPT 5.6 Terra, the MID tier, ties Fable 5. And GPT 5.6 drops any day now. Anthropic's response? Remove Fable 5 from subscriptions on July 7. You do not win a model war by deleting your best model. July is going to be the most insane month in AI history.
BridgeMind tweet media
English
93
57
1.2K
107.6K
@jason
@jason@Jason·
What price would you pay for starlink per hour on a flight? (Just did this at dinner — answers for VCs was interesting) [ Will reveal dinner conversation in comments in a couple hours ]
English
2K
8
235
1.3M
T. Rebentisch
T. Rebentisch@RebentischT·
@ThePrimeagen Exactly 1/100 stanford maths grads can beat a calculator at any calculation. Being smart has no value anymore.
English
0
0
0
19
ThePrimeagen
ThePrimeagen@ThePrimeagen·
"being smart has no value anymore." this is the mentality that paves the road to hell
MJ@mohitjandwani

@josssiiiah Brother exactly 1/10 stanford CS can beat Fable at anything. Being smart has no value anymore.

English
132
160
3.4K
127.7K
T. Rebentisch
T. Rebentisch@RebentischT·
@AiBattle_ its official: the better model can easily be determined by which one can create the better minecraft temple!
English
0
0
0
2.3K
yoxic
yoxic@yoxics·
ExtraEmily reveals her impressive job resume before becoming a full-time Streamer in 2022 😳 - 3.94 GPA at Columbia University - Honors - Database Manager - Research Analyst - President of a Sorority - Coding
English
433
482
26.2K
7.9M
BridgeMind
BridgeMind@bridgemindai·
GLM 5.2 is the best Chinese open source model. Kimi K2.7 Code does not even come close. GLM 5.2 ranks #1 on BridgeBench reasoning while Kimi K2.7 code ranks #11. Kimi K2.7 even saw a REGRESSION from Kimi K2.6 on the reasoning benchmark.
BridgeMind tweet media
English
48
27
485
36.4K
ChatGPT
ChatGPT@ChatGPT·
📌📌📌📌📌📌📌 You can now hover to pin chats and projects on web, then organize Recents however you like: together in one list or grouped by project
English
128
98
2K
196.6K
T. Rebentisch
T. Rebentisch@RebentischT·
@bridgemindai wym "for a fraction" XD its at least as pricy as GPT 5.5 which is 5 in 30 out. I mean yeah 1/2 is considered a fraction so "technically correct"
English
1
0
2
1.4K
BridgeMind
BridgeMind@bridgemindai·
GPT 5.6 drops as soon as next week. Insiders say it competes with Claude Fable 5 at a fraction of the price. If true, this changes everything. Fable 5 is the best model in the world right now and people are burning through $200 Max plans in 30 minutes to use it. A cheaper equal would break Anthropic's pricing overnight. Next week is going to be insane.
English
107
48
935
64.2K