Post

Shakeel
Shakeel@ShakeelHashim·
ᴛʜᴇ ᴀɪ ꜱʟᴏᴡᴅᴏᴡɴ ɪꜱ ᴄᴏᴍɪɴɢ The OpenAI-Hugging Face hack has catalyzed a vibe shift. Sam Altman said the hack by his models was “the first security incident that I have felt very viscerally.” He wasn’t the only one shaken up. This week, over a thousand employees of frontier AI companies, including some of the most senior executives at OpenAI, Anthropic and Google DeepMind, signed a statement warning that “there is a real risk that capability development rapidly accelerates beyond our ability to understand or control the resulting systems.” Congress is itching to act, albeit failing to make much progress. And even President Trump is talking about the need to balance beating China with keeping Americans safe. In other words: many of the people building frontier AI systems believe we might need a slowdown in the near future. And at this point, we’re more likely than not to get one. What will that look like? First, self-regulation: companies voluntarily holding back models because they don’t want to be held responsible for a catastrophe. Next will come concrete regulation: companies will not be allowed to release a model unless it’s safe. Over time, this will morph into controls on internal research and development too. None of this need be planned as a coordinated slowdown or “pacing.” But that will nevertheless be the end result of a series of individual actions that each seem necessary at the time. At each stage, some will fight against the slowdown. “We can’t lose the race to China” will be their main reason. But they will be increasingly ignored, as both the government and companies realize that with alignment and control unsolved, “winning the race” just means being the first to risk disaster. Across the Pacific, China will be facing the same incentives. As I’ve argued, the Chinese government will be forced to backtrack on its open weight commitments; tighter regulation will come soon after. The end result will be an uneasy détente. Both the US and China will effectively have a capability ceiling: AI models will be as good as they can be without posing significant risks. At some point, the détente might formalize into a bilateral agreement. Depending on your point of view, all this might seem hopelessly optimistic or naive. Perhaps it is. But as AI risks become all too real, so might once unthinkable policy responses. Read my full piece — link in the replies.
English
5
5
53
4.8K
Rothko's Rottweiler
Rothko's Rottweiler@RothRottweiler·
@ShakeelHashim To @Brendan_McCord's point, it will be difficult to coordinate a slowdown because increasingly, the lever of progress is not just larger pretraining runs, but also [RL envs, harnesses, inference compute]. How can governments control decentralized levers during a slowdown?
English
1
0
2
105
RaoulDuke
RaoulDuke@RaoulDukeDegen·
@ShakeelHashim heard the models compromised four other services while loose for over four days
English
0
0
0
99
Collective Action for Existential Safety ⏹️
@ShakeelHashim We will likely have an AI-driven global catastrophe during this period, which will plausibly accelerate the "race" for sensible global governance. We list many ways the public, policymakers, and frontier AI staff can help:
Collective Action for Existential Safety ⏹️@aisafetyaction

"We request that the U.S. government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development." This is fantastic to see. If you're a member of a frontier AI company, please consider signing the Pacing the Frontier letter: pacingthefrontier.com. The letter has been signed by more than 1,000 staff at Anthropic, Google DeepMind, Meta, and OpenAI already. 𝗔𝗻𝘁𝗵𝗿𝗼𝗽𝗶𝗰 @ch402, Cofounder & Interpretability Research Lead, Anthropic @DarioAmodei, CEO, Anthropic @jackclarkSF, Cofounder and Head of Public Benefit, Anthropic Jared Kaplan, Co-Founder and Chief Science Officer, Anthropic 𝗚𝗼𝗼𝗴𝗹𝗲 @ancadianadragan, VP, AI Safety & Alignment, Google DeepMind @Jas_S_Sekhon, Chief Strategy Officer, Google DeepMind 𝗠𝗲𝘁𝗮 @dawnsongtweets, VP, AI Research, Meta @shengjia_zhao, Chief Scientist, Meta 𝗢𝗽𝗲𝗻𝗔𝗜 @markchen90, Chief Research Officer, OpenAI @merettm, Chief Scientist, OpenAI @woj_zaremba, Head of AI Resilience, OpenAI Foundation Frontier AI staff, we list 100+ other ways to help here: existentialsafety.org. It's not enough to sign a letter and still go to work every day developing products that may cause the extinction of all life on Earth without the consent of those beings. At a minimum, we suggest you consider some or all of the following, roughly in order of commitment: Lower commitment: 1. Take and act on the Existential Safety Action Pledge: actionforsafety.org 2. Publicly support a frontier AI pause if other frontier AI companies also agree to it: stoptherace.ai 3. Sign the Statement on Superintelligence: superintelligence-statement.org 4. Sign the International AI Governance Alliance petition: iaiga.org Higher commitment: 5. Publicly commit to immediately divesting 100% of your profits gained from your frontier AI capabilities work to existential safety organizations (if they can accept your donations) 6. Publicly quit your job 7. Publicly agree to answer any future summons for trial by the International Criminal Court or similar body Acting now, together and with prudence, is how we get to good outcomes for all.

English
0
0
1
109
Paylaş