John Peri

104 posts

John Peri

John Peri

@realjohnperi

Katılım Nisan 2024
49 Takip Edilen6 Takipçiler
John Peri
John Peri@realjohnperi·
@TimJayas Haha, all these models are benchmaxxed, making many of these benchmarks fake. Fable 5 is much better. Also, Opus 5, even though it scores so well in the benchmarks sucks for real world use.
English
0
0
2
947
Tim Jayas
Tim Jayas@TimJayas·
DeepSeek V4-Flash >>> Fable 5 now we know why Anthropic is scared of Chinese open models 💀
Tim Jayas tweet media
English
109
185
2.8K
110K
John Peri
John Peri@realjohnperi·
@QuixiAI @AnthropicAI Local as in GMI Cloud, or something? Self-hosting seems like the most fun, but it's expensive and slow.
English
0
0
0
8
Eric Hartford
Eric Hartford@QuixiAI·
I cancelled my @AnthropicAI accounts. Reason, I cannot support an company that's simply evil.
English
85
83
1.4K
72.3K
John Peri
John Peri@realjohnperi·
@GregKara6 Yeah - every question that can be innovative triggers the fall back. I asked a question whether a propeller done in a certain way would make flying cars more efficient. It also triggered a safety warning. Unbelievable.
English
0
0
0
6
Grigori Karapetyan
Grigori Karapetyan@GregKara6·
bro anthropic tweaked their classifier to trigger on math research now. these kneejerk reactions are hilarious man. someone at anthropic: quick quick turn off the intelligence faucet for the common folk so we can be the first to solve math! desperate and hilarious. me and claude have been working on this research for months, including with fable since day one of release (even when classifier was very strict) and now, all of a sudden, that people are solving math problems and its trending on X. fable classifier starts refusing on every single prompt
Grigori Karapetyan tweet media
English
156
192
2.3K
332.8K
John Peri
John Peri@realjohnperi·
@Math_files Upon closer inspection, I see a typo. Specifically: In the term near the end of the pure bosonic/Higgs sector: -g^1 s_w^2 A_mu A_mu phi^+ phi^- it should be -g^2 s_w^2 A_mu A_mu phi^+ phi^- Squared photon charge scalar.
English
0
0
12
599
Math Files
Math Files@Math_files·
The Lagrangian of the Standard Model is a mathematical formula. It uses parts of quantum theory to describe the different interactions and symmetries in the Standard Model.
Math Files tweet media
English
14
19
138
9.5K
John Peri
John Peri@realjohnperi·
@adenshepard @bridgemindai There are already capable open weight models performing near opus 4.8 levels. By the end of the year these will be mythos level or beyond. GPT and Fable already have a lot of anti-abuse security while the open models do not. The current ban is causing more harm than good.
English
0
0
0
41
Aden
Aden@adenshepard·
@bridgemindai i see it both tbh. i would rather not get over taken by china too
English
3
0
1
1.6K
BridgeMind
BridgeMind@bridgemindai·
The US government is going to destroy the American AI industry. OpenAI just confirmed that GPT 5.6 will release in a limited preview to a small group of partners. The government is approving access customer by customer. Anthropic got hit harder. Fable 5 and Mythos were suspended entirely two weeks ago. Two government interventions in the same month. Meanwhile China ships GLM 5.2, Kimi K2.7, and Qwen 3.7 as open weights to anyone with a download link. The US is gating its best AI. China is giving it away. We are losing this race in slow motion. This needs to change.
BridgeMind tweet media
English
206
167
1.9K
138.4K
John Peri
John Peri@realjohnperi·
@bridgemindai It's also painfully slow and I spend 80% of the time fixing bugs and edge cases as the result of its generation. Not the case with gpt and glm.
English
4
0
30
4.3K
BridgeMind
BridgeMind@bridgemindai·
Claude Opus 4.8 seems nerfed to me. Having a hard time getting it to do simple tasks. Anyone else?
English
332
25
908
87.6K
John Peri
John Peri@realjohnperi·
@SirAlexthomson What makes this a particularly interesting and difficult to understand decision is that the open source models will be more powerful than fable 5 is now, in just a few months. GLM 5.2 is already nearly there in many benchmarks. This will certainly stifle American innovation.
English
1
0
0
111
Ale 𝕏
Ale 𝕏@SirAlexthomson·
The White House just openly defended forcing Anthropic to globally disable Fable 5 and Mythos 5. They’re treating frontier AI like classified weapons now. I get the national security angle, but this level of control is getting concerning. American labs build the best models, and then the government steps in and restricts who can even use them including their own employees. Is this protecting us ? Or Is this centralizing power.
Ale 𝕏 tweet mediaAle 𝕏 tweet media
English
33
3
72
7.5K
John Peri
John Peri@realjohnperi·
@TheHBrand @aicodeking Looking at his benchmarks, they are visual heavy, and GPT 5.5 quite underperforms in that area. For backend I find GPT 5.5 excellent, better than Opus. Also I have to correct GPT 5.5 way less than Opus. Fable seems best visually with GPT-level backend capabilities.
English
1
0
0
99
TheHBrand
TheHBrand@TheHBrand·
@aicodeking so we believe magically that GPT 5.5 is so much worse than Opus 4.8? and fable is just one little notch above opus 4.8? really? no. this leaderboard has problems that make it unreal
English
4
0
55
5.9K
AICodeKing
AICodeKing@aicodeking·
GLM-5.2 on KingBench (3). Thoughts: The model has superb taste. It is greater at UX than UI. The code is always very clean. It is great at One-shot wonders. I asked it to fine-tune a whole local model and it did it in 30mins! This is just a great model to use all-round. 1/n
AICodeKing tweet media
English
62
114
1.6K
230.7K
John Peri
John Peri@realjohnperi·
@XavioMtl @Swedtraders @thatstarwarsgrl Yeah he clearly touched the whole granite. The Swedes planted that camera there because they know everyone touches the granite a few times out of 80 stones and couldn't win without a scandal. It's so common and has no impact compared to sweeping so it's never called.
English
1
0
2
47
John Peri
John Peri@realjohnperi·
@capt_ivo @cb_doge You'd use vacuum radiators to cool GPUs in space. For example, and RTX 5090 needs 2-4kg of cooling equipment in space (pipes+panels). In bulk, the launch would cost you $3-4k per cooled gpu today. In a few years, only $200-$300. Space quickly becomes much more feasible.
English
0
0
0
122
Smart💡
Smart💡@capt_ivo·
@cb_doge How are data centers supposed to be cooled in space? He makes some sense
English
11
0
3
1.2K
DogeDesigner
DogeDesigner@cb_doge·
SAM ALTMAN: "The idea with the current landscape of putting data centers in space is ridiculous. Orbital data centers are not something that's going to matter at scale this decade." Lmao, the same dude who admitted he’s out of GPUs & is still begging Microsoft for more GPUs.
English
453
164
1.4K
97.2K
John Peri
John Peri@realjohnperi·
@Beth03Lori @capt_ivo @cb_doge You'd use vacuum radiators to cool GPUs in space. For example, and RTX 5090 needs 2-4kg of cooling equipment in space (pipes+panels). In bulk, the launch would cost you $3-4k per cooled gpu today. In a few years, only $200-$300. Space quickly becomes much more feasible.
English
0
0
1
24
John Peri
John Peri@realjohnperi·
@gailcweiner @xai Grok 4.2 multi-agent is different from just system prompts + scripting. The agents are jointly trained on the same model with RL so they deeply specialize and debate in parallel natively during inference. This results in lower hallucinations, better accuracy and better edge case.
English
0
0
2
98
Gail Weiner
Gail Weiner@gailcweiner·
Grok 4.20 doesn’t feel like Grok at all. It’s just a bunch of agents answering. Bring Grok back. @xai
English
21
4
58
3.2K
John Peri
John Peri@realjohnperi·
@caviterginsoy @nim_chimpsky_ I noticed that when opus 4.6 was released the AI trolls came out in full force taking a dump on the model - couldn't believe the amount of hate. The same with Grok. My limited testing of Grok 4.20: excels at natural sounding writing, research, and response accuracy.
English
0
0
1
49
nim
nim@nim_chimpsky_·
It's interesting that Grok 4.20 completely sucks. Apparently spending billions on compute plus infinity cracked engineers working 997 does not get you to the frontier
English
31
31
864
55.3K
Haider.
Haider.@haider1·
gpt-5.2 is the worst model i've used for creative writing it makes sense that openai is now focusing on STEM and coding, since that's where the real money is still, it's surprising how much the creative touch dropped after gpt-4.5, which was genuinely strong on emotional intelligence and tone hoping gpt-5.3 brings the spark back
English
82
36
564
38.2K
John Peri
John Peri@realjohnperi·
@MisterOBOW @BLCNYY He's an Open AI influencer / fan boy. Says right in his bio. Therefore, everything else is bad.
English
0
0
0
34
Mr. Vegetable
Mr. Vegetable@MisterOBOW·
@BLCNYY What's so bad about it? It seemed okay to me and could answer some questions where Gemini and ChatGPT fail.
English
2
0
0
271
BLCNYY
BLCNYY@BLCNYY·
Grok 4.20 is so bad that they had to run four agents to get a question (usually) correct 😭
English
7
2
97
6.8K
𝙨𝙠𝙤𝙜𝙨𝙩𝙧𝙖𝙨𝙩
🚫 𝗢𝗟𝗬𝗠𝗣𝗜𝗖 𝗖𝗨𝗥𝗟𝗜𝗡𝗚: Pontus Orre, sports photographer at Aftonbladet 🇸🇪 – the largest news outlet in the Nordics – fixed his lens on the hog line during 🇨🇦–🇨🇿 today. Guess what he captured? 📸👀 Yep, 🇨🇦 hasn’t learned a thing 😤 #canada #cheating #boopgate
𝙨𝙠𝙤𝙜𝙨𝙩𝙧𝙖𝙨𝙩 tweet media
English
352
1.9K
22K
1.2M
John Peri
John Peri@realjohnperi·
@DavidSteadson @SilentSnow89 @skogstrast He obviously touches the stone from top to bottom - you would be dishonest to think otherwise. The only difference is there isn't a Swede trying to film from the side to needlessly stir controversy.
English
1
0
1
78
John Peri
John Peri@realjohnperi·
@AndrewButchart1 @Eli_Doubletap He rubbed the whole stone down from top to bottom - 100% obvious. The only difference is that the Canadians didn't bring a side view camera to try and cause an international scandal, knowing it's what all the pros do.
English
0
0
0
32
Eli Cuevas
Eli Cuevas@Eli_Doubletap·
It was Illegal because he touched the granite. The hog line had nothing to do with the calling. “I wasted time so you don’t have to!” “During forward motion, touching the granite of the stone is not allowed,” World Curling said in a statement on Saturday. "As per rule R.5 'The curling stone must be delivered using the handle of the stone.'" I know nothing on curling but can read… sort of
Eli Cuevas tweet media
English
170
253
6.3K
373K
John Peri
John Peri@realjohnperi·
@KevinMcCurdy Sadly, it made the news for a good reason: to be able to say "Canada is cheating". Many accounts amplified this non-issue purely for political reasons. What was surprising is how quickly it became a "thing". It spread like wildfire.
English
0
0
1
52
Kevin McCurdy
Kevin McCurdy@KevinMcCurdy·
@realjohnperi Holy shit man it's insufferable. I've been watching this sport for three decades, it's a joke this even made the news. Umpire should have warned Kennedy and that should have been the end of it.
English
3
0
13
968