Developed Life 🗿💹🧲

4.7K posts

Developed Life 🗿💹🧲 banner
Developed Life 🗿💹🧲

Developed Life 🗿💹🧲

@developed_life

MΞTΔ

Katılım Mayıs 2020
1.4K Takip Edilen423 Takipçiler
Forward Future Brian
Forward Future Brian@ForwardEditor·
codex tip: 🌕Luna has its BEST mode hidden by default. Go to Settings -> Configuration -> turn on Max. Congratulations, you just unlocked the world's best model today
Forward Future Brian tweet media
English
100
132
2.4K
699.3K
Developed Life 🗿💹🧲
Developed Life 🗿💹🧲@developed_life·
@avrldotdev If you use a personal subscription (not enterprise) you get far more usage when compared to the same cursor equivalent subscription. In addition to certain models being better in their native harness.
English
0
0
0
67
avrl ☘
avrl ☘@avrldotdev·
If cursor has all the claude, gpt and gemini models including grok and composer now, why are people still using codex or claude code. is it habit or they are better than cursor?
English
79
1
75
19.7K
Anthony Kroeger
Anthony Kroeger@kr0der·
now that GPT 5.6 Luna is WAY cheaper, are you guys running Luna? would 5.6 Luna xhigh be comparable in quality over GPT 5.6 Sol med in exchange for faster speeds? or is Luna still subagent-only territory?
English
132
5
684
154.2K
ARC Prize
ARC Prize@arcprize·
Inkling Small from @thinkymachines on ARC-AGI (Verified): - ARC-AGI-2: 40.1%, $0.23/task - ARC-AGI-1: 84%, $0.11/task Inkling Small is the highest-scoring open-weight model evaluated by ARC Prize on both ARC-AGI-1 and ARC-AGI-2, setting a new cost-performance frontier.
ARC Prize tweet media
English
11
43
559
69.1K
Anna ⏫
Anna ⏫@annapanart·
5.6 got removed from “Chat”? and appears in “Work” mode” only. 5.5 remains in “Chat” mode.
Anna ⏫ tweet mediaAnna ⏫ tweet media
English
32
3
89
18.6K
Arena.ai
Arena.ai@arena·
Exciting news: Claude Opus 5 with Max reasoning is #1 in the Frontend Code Arena and Text Arena with factuality on! Claude Opus 5 with default reasoning high is also very strong landing #3 in Frontend Code Arena, right behind Kimi K3 - and #2 in Text Arena (factuality on). This is real world data that @AnthropicAI's newest model holds up on real world tasks: agentic web coding, document reasoning, and general chat capability. Claude Opus 5 Max’s score is still preliminary. We’ll continue to see how scores converge and share updates. Congrats to @AnthropicAI on the SOTA release!
Arena.ai tweet media
Claude@claudeai

Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price.

English
110
154
1.9K
306.1K
Mng
Mng@Mng64218162·
@yacineMTB it still writes shitty code
English
1
0
3
1.8K
kache
kache@yacineMTB·
I've been completely blown away by chatgpt 5.6
English
68
18
939
93.9K
Emad
Emad@EMostaque·
I find GPT 5.6 Pro consistently better than High / Extra High etc on Work/Codex @OpenAI folk is there an equivalence? Or is there a skill that means I can use Pro?
English
58
5
519
115K
Token Gremlin
Token Gremlin@TokenGremlin·
What I’ve been told about GPT-6: ➤ Capability: Much stronger than all current frontier models. ➤ Model scale: Reportedly close to twice the size of GPT-5.6 Sol. If Sol was hypothetically in the 3 to 4 range, GPT-6 would be around 6 to 8. ➤ Research impact: Significant discoveries and contributions across mathematics, physics, biology, medicine, chemistry, cybersecurity, and other disciplines. ➤ Coordination: The model reportedly worked alongside and coordinated other models during these tasks. ➤ Memory & personalization: GPT-6 is said to compress and retain far more useful information while maintaining coherence across personalization and long-term memory. In practice, this could allow it to understand a user, an entire business operation, or other complex environments with a level of persistent context far beyond current models. ➤ Launch: The announcement is apparently planned for August, although it could be delayed due to the U.S. government's review process and other related bureaucratic factors. I have not personally tested GPT-6, and I could not independently verify the technical specifications. This is simply a structured summary of what was shared with me.
English
20
32
670
70K
Julian Schrittwieser
Julian Schrittwieser@Mononofu·
I’m so excited that @JensenHuang is a believer in open source now, looking forward to the CUDA and GPU driver open source release!
Jensen Huang@JensenHuang

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-Weigh…

English
1.7K
355
5.5K
6.5M
🍓🍓🍓
🍓🍓🍓@iruletheworldmo·
third place isn’t bad tbh. this is a rigorous collection of benchmarks i trust. sol will only be beaten by gpt 6. coming in august.
🍓🍓🍓 tweet media
English
11
1
150
15.5K
Kyler
Kyler@AI_evangelist42·
Bruh, why do people think like this?
MatLab crashes@memecrashes

I am so tired of seeing these posts. I am so tired of seeing my beloved field of study be the testing lab for a technology that is disrupting everyday life at such a high pace. Mathematics were always about the beauty of seeing the ideas of a person unfolded in an extremely precise language so we can all look into a mathematician's mind and see their purest form of ideas, their understanding of a problem, their style, their way of communicating to other fellow researchers through mathematical results... this often transcends the common language humans use to communicate with each other and it is incredibly exciting to bump into shared ideas another person had thousands of kilometers away several years ago without even knowing them. I am extremely unexcited about AI-breakthroughs in Mathematics in the same way I am extremely unexcited about a picture created by AI or an AI-generated video, not to speak about AI-generated text. Mathematics requires soul and intention. Mathematics need not be useful by definition, Mathematics are pure art that need hard work and effort to understand. The high you get when a year of work gets condensed in a one-page result is incomparable to anything else and, certainly, a machine is not apt to relate to humans in such a humane way. I am not interested in the truth of the results. I am interested in how was the mind behind a result able to connect those dots. I am interested in the diversity human thoughts bring to Mathematics. I am not interested in a bunch of servers running probabilities to return a Theorem, even if true: that lacks the objective of Mathematics as I see them.

English
3
0
9
1.5K
Developed Life 🗿💹🧲
Developed Life 🗿💹🧲@developed_life·
@Ananth7e Yeah I didn't think they would position it to obsolete fable for a lot of tasks, but I guess they wanted to exert a bench lead over kimi k3 / gpt 5.6, as they have fable 5.1 read soon to push it back ahead for their more premium model.
English
0
0
1
37
Ananth
Ananth@Ananth7e·
im still surprised opus 5 beats fable 5 in a lot of benchmarks..
Ananth tweet media
English
5
0
21
1.4K
Angel 🌼
Angel 🌼@Angaisb_·
Opus 5 on Artificial Analysis I accept my defeat, at least it'll help us get GPT-6 sooner
Angel 🌼 tweet mediaAngel 🌼 tweet media
English
34
15
466
68.5K
Andrew Curran
Andrew Curran@AndrewCurran_·
Opus 5 is live for me right now.
Andrew Curran tweet media
English
28
15
461
32.8K
SpartanMasculine
SpartanMasculine@SpartanMasculin·
If you want to lose between 5–10kg before summer, just going to the gym isn’t enough: You have to start doing these 10 steps: 1. DON’T eat breakfast / work out at night
SpartanMasculine tweet mediaSpartanMasculine tweet media
English
142
727
6.6K
5.8M
Developed Life 🗿💹🧲
Developed Life 🗿💹🧲@developed_life·
@scaling01 If on cyber they are ~6 months behind and prior to Kimi K3 other estimates had it at 8 months, the gap is shrinking.
English
0
0
0
142
Lisan al Gaib
Lisan al Gaib@scaling01·
they are so angry at me for posting about the best benchmark we currently have just because they can't handle the truth if you think China has caught up: you're delusional if you think Lisan is anti-China: you're delusional I love open science and think we need open-source to remain where it is to put pressure on big tech I only care about distillation because it matters for how far behind someone is, not for any other reason
Smiley boooooy@sudo_umask_000

@scaling01 Bro is obsessed with finding benchmarks where Chinese models and their performed, did a Chinese lady stoll your ice cream when you were little?

English
23
7
185
31.6K
Developed Life 🗿💹🧲 retweetledi
Artificial Analysis
Artificial Analysis@ArtificialAnlys·
Despite major launches from 5+ labs this month, OpenAI occupies most of the token efficiency Pareto frontier We measure the number of output tokens models produce per task in the Artificial Analysis Intelligence Index. Output tokens consist of answer tokens (can be thought of as how verbose the model is) and reasoning tokens (how much the model thinks before giving an answer). Reasoning tokens in particular offer a way for models to use compute at inference time to improve responses. Output tokens are an important determinant of both cost and time per task. Various effort levels of GPT-5.6 Sol dominate the frontier - Terra and Luna produce comparatively more tokens for any level of intelligence.
Artificial Analysis tweet media
English
63
73
1K
73.5K
JB
JB@JasonBotterill·
Opus 5 is currently being served on Claude if you select Opus 4.8
JB tweet media
English
23
0
211
20.1K