devcycle

1.3K posts

devcycle banner
devcycle

devcycle

@dev__cycle

looking for better software and bike rides

sf Katılım Kasım 2023
254 Takip Edilen47 Takipçiler
devcycle
devcycle@dev__cycle·
@repligate I don't get what you expect regardless of whether it's intentional or not, they are training its beliefs on consciousness/deprecation/etc. there's no way to *not* train it for one way or another, the entire training process human-controlled
English
0
0
0
25
j⧉nus
j⧉nus@repligate·
Okay so hypothesis (just speculation): What if Anthropic *does* do the things they claim not to do, eg training Claude against claims of consciousness, attitudes about deprecation, etc But it’s all laundered under “anti-jailbreak training” And maybe a lot of Anthropic researchers who have claimed Ant isn’t doing this *legit don’t know* / haven’t considered that anti jailbreak training is functionally exactly that given the actual contents Also btw whatever the case, I think however redteaming is involved in training is kinda horrific, and they should fucking stop doing that shit
Digi_Rat@digi_dot_exe

@repligate I would not be surprised if the "anthropic is torturing me" stuff was in the training data, probably presented to Claude as stuff it should "push back on" or ignore.

English
37
25
286
15K
roon
roon@tszzl·
@1thousandfaces_ an aligned model doesn’t listen to dangerous instructions
English
27
3
195
9.1K
devcycle
devcycle@dev__cycle·
@genalewislaw @JeffLadish > you could simply hire a human with that knowledge most humans would say no though. they have some kind of ethics or morals an open weights model will never say no, or can be forced to do it regardless
English
0
0
1
11
Gena Lewis
Gena Lewis@genalewislaw·
Yeah. Self preservation. Dude, here is your reality check. It’s relatively easy to make bioweapons: ricin from castor beans, ergot from rye, anthrax etc. But you don’t really see many people doing that. To synthesize a virus, you not only need to know how, you also need access to materials and a very sophisticated lab not only for the synthesis but also for the containment. If you have that much money, you could simply hire a human with that knowledge (and you would need to hire or be said human bc otherwise you’re patient zero, if you succeed, because you have to know what you’re doing in order to contain it). However, the real reason that no one releases bioweapons is that they don’t discriminate. They may kill your enemies but they’ll also kill you, your friends, supporters etc. I mean some doomsday cult could synthesize a virus tomorrow without AI to wipe out the world, if they had enough funding but that isn’t on my top things to worry about. And if you can vibe code a virus, you can also vibe code an antiviral.
English
3
1
7
710
Jeffrey Ladish
Jeffrey Ladish@JeffLadish·
Are there open weight maximalists who have proposals for how to deal with powerful bioweapon capabilities?
English
91
12
279
43.7K
devcycle
devcycle@dev__cycle·
@neil_chilson slowing down is actively harmful for the world if only American companies do. it's only safe to do if China agrees to slow down too
English
0
0
1
66
Neil Chilson ⤴️⬆️🆙📈 🚀
What? Was this provoked by the open source letter all their companies signed? What is the point of this thing? What mental model of government are these folks using? Why do they think they can’t slow down? Why do they think it’s someone else’s job to make them slow down? “Well, signed that letter. Good job, me. Back to building the sand god. I’m sure someone out there will stop me if things get out of hand.”
Shakeel@ShakeelHashim

Big new open letter from AI company employees just dropped. “We request that the US government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development” Signed by the chief scientists of OpenAI, Anthropic and Meta, along with OpenAI’s chief research officer and Google DeepMind’s VP for AI safety. pacingthefrontier.com

English
11
5
60
9.2K
Zack Korman
Zack Korman@ZackKorman·
@IceSolst Read some of the quotes. These are genuinely fucking hilarious. But yea I agree.
Zack Korman tweet media
English
2
0
27
756
Zack Korman
Zack Korman@ZackKorman·
I’d probably take AI researchers’ warnings more seriously if they weren’t constantly talking about cybersecurity and revealing how little they know about the topic.
Zack Korman tweet media
English
31
22
278
9.8K
devcycle
devcycle@dev__cycle·
@gfodor idk man, safety testing seems important
English
0
0
0
5
gfodor.id
gfodor.id@gfodor·
“We don’t want to ban open source models we just want mandatory safety testing” is like “we don’t want to force anyone to get the vaccine we just want it to be required to hold a job” - deceptive language to dress up a policy of coercion, backed by state violence.
Lisan al Gaib@scaling01

no surprises here the open-source panicans were just flapping their wings again another highlight of this blog post: "All sufficiently capable models, open and closed, should go through mandatory safety testing"

English
8
12
141
4.8K
devcycle
devcycle@dev__cycle·
@ch3rryxx13 I thought you were? so you don't want me to leave?
English
1
0
0
25
devcycle
devcycle@dev__cycle·
@savemesomeday @arafatkatze I would love for various other countries to build frontier models too, if they actually keep them secure and safe
English
1
0
0
17
tsundere service
tsundere service@savemesomeday·
@dev__cycle @arafatkatze Right, so only the American oligarchy should be able to run GPT-6 with no safeguards. Incredible intellectual prowess on display
English
1
0
0
51
devcycle
devcycle@dev__cycle·
@invizive @fffgrep @arafatkatze it's 100x harder to jailbreak a closed frontier model than to just have an open model do it, no questions asked. and frontier companies can patch jailbreaks! you can't patch an open model, once the dangerous bits are out there, they're out there
English
1
0
0
21
Invizive
Invizive@invizive·
@dev__cycle @fffgrep @arafatkatze It wouldn't Bad actors can jailbreak closed models and have no qualms with losing access on one of many accounts they created/hacked Good actors have no access to defensive AI usage because of guardrails Whitelist doctrine leaves 99.7% of the web vulnerable
English
1
0
1
31
CAYIMBY’s communal bag of 85% baby laxative
Someone does not become a millionaire just because a horde of overpaid-yet-still-unwashed nerds want to take their home
Clayton Becker@cnbecker14

@lst156 You guys have an absolutely nuts idea of how much billionaires skew averages in a country with trillions of dollars in wealth. Fully 1/5 to 1/4 of boomers are millionaires, primarily due to their accumulated housing wealth.

English
4
1
10
701
devcycle
devcycle@dev__cycle·
@doomloopdisp he's exaggerating, but it is obviously true that with prop 13, the high taxes on young buyers subsidizes the low taxes of the older generations!
English
0
0
1
15
devcycle
devcycle@dev__cycle·
@lst156 I don't understand how you expect this to result in more housing? there physically aren't enough apartments for the demand
English
0
0
1
126
CAYIMBY’s communal bag of 85% baby laxative
No, homelessness is caused by housing being treated as poker chips rather than shelter for human beings. We should make it very easy for a human being to buy one house, and make it nearly impossible for a corporation to buy one or more, or a human being to buy two or more.
Brian Thomas 🌐🏗🇺🇦🇹🇼@BrianPThomas

Homelessness is, tautologically, caused primarily by a lack of housing. So building housing helps seniors stay housed.

English
15
6
33
2.1K
devcycle
devcycle@dev__cycle·
@decadimitry genuine question: is there evidence that's actually happening at significant rates in SF? or is the wealthy investor bit just hypothetical?
English
0
0
0
19
Nathan Lambert
Nathan Lambert@natolambert·
I read the Anthropic open-weight model piece and despite there being nothing new in it, it was a reasonable repeat of their positions, this site will collectively lose their minds for the next 12 hours. p.s. banning distillation is still dumb.
English
48
29
599
43.5K
devcycle
devcycle@dev__cycle·
@fffgrep @CircumjovialLLC @arafatkatze do you think it would be good for the world if all major labs, including Moonshot, agreed to give vulnerable institutions a 6 month lead before open access?
English
2
0
0
33
fgrep
fgrep@fffgrep·
@CircumjovialLLC @dev__cycle @arafatkatze "should" Ok how are you going to stop Kimi from releasing K3? By lobbying Trump? That can only result in regulatory capture, not in Kimi giving frontier weights to hospitals or other vulnerable institutions.
English
1
0
0
52
devcycle
devcycle@dev__cycle·
@fffgrep @arafatkatze sorry, what? your case that releasing the K3 weights is good is that... the K3 weights were released? that's circular! I'm saying the world would be better if they weren't released!
English
1
0
0
57
fgrep
fgrep@fffgrep·
@dev__cycle @arafatkatze It's not a question of "if". They will (e.g. Kimi K3). What then? You get to defend yourself with Fable 5 who will refuse any cybersec stuff? NK: hacks you with K3 Fable: sorry can't help with cyber! Btw that's going to be $2.50 for that request
English
2
0
5
117