This weekly roundup thread is intended for all culture war posts. 'Culture war' is vaguely defined, but it basically means controversial issues that fall along set tribal lines. Arguments over culture war issues generate a lot of heat and little light, and few deeply entrenched people ever change their minds. This thread is for voicing opinions and analyzing the state of the discussion while trying to optimize for light over heat.
Optimistically, we think that engaging with people you disagree with is worth your time, and so is being nice! Pessimistically, there are many dynamics that can lead discussions on Culture War topics to become unproductive. There's a human tendency to divide along tribal lines, praising your ingroup and vilifying your outgroup - and if you think you find it easy to criticize your ingroup, then it may be that your outgroup is not who you think it is. Extremists with opposing positions can feed off each other, highlighting each other's worst points to justify their own angry rhetoric, which becomes in turn a new example of bad behavior for the other side to highlight.
We would like to avoid these negative dynamics. Accordingly, we ask that you do not use this thread for waging the Culture War. Examples of waging the Culture War:
-
Shaming.
-
Attempting to 'build consensus' or enforce ideological conformity.
-
Making sweeping generalizations to vilify a group you dislike.
-
Recruiting for a cause.
-
Posting links that could be summarized as 'Boo outgroup!' Basically, if your content is 'Can you believe what Those People did this week?' then you should either refrain from posting, or do some very patient work to contextualize and/or steel-man the relevant viewpoint.
In general, you should argue to understand, not to win. This thread is not territory to be claimed by one group or another; indeed, the aim is to have many different viewpoints represented here. Thus, we also ask that you follow some guidelines:
-
Speak plainly. Avoid sarcasm and mockery. When disagreeing with someone, state your objections explicitly.
-
Be as precise and charitable as you can. Don't paraphrase unflatteringly.
-
Don't imply that someone said something they did not say, even if you think it follows from what they said.
-
Write like everyone is reading and you want them to be included in the discussion.
On an ad hoc basis, the mods will try to compile a list of the best posts/comments from the previous week, posted in Quality Contribution threads and archived at /r/TheThread. You may nominate a comment for this list by clicking on 'report' at the bottom of the post and typing 'Actually a quality contribution' as the report reason.

Jump in the discussion.
No email address required.
Notes -
This kind of thing is why I'm finding the current tizzy over AI doom so hard to believe. Because the same companies that are dressing up in the robes of Cassandra and crying "Doom, doom, unless there is a pause!" are...
... building biolabs for Claude to frolic and play in, doing research (to cure all human ills, of course!)
Gosh, it is so reassuring that Dario Amodei isn't doing anything to hasten the chances that AI will create some kind of bioweapon to kill off humans, like the doomers fear!
Do you see why people think all these calls for a pause are only some kind of scam by the big companies to get regulations passed that will favour them and copper-fasten their positions as the only sources for AI in general?
Are biolabs inherently unsafe?
I think GAIN OF FUNCTION RESEARCH is inherently unsafe, and people doing it should be
banished exiledpromoted to a small Jovian moon, but surely there are lots of other biological research that can be done pretty safely, yeah?One of the common responses to AI safety concerns is "just don't let it get out of the box." That's actually a pretty good point: as long as the AI doesn't have godlike intelligence and as long as we actually take steps to keep it inside the box where it can't impact the real world directly, the danger that an AI can pose is somewhat contained and limited.
When you start giving AI access to real world stuff, you no longer have the ability to keep it fully inside the box. At this point it's probably fine and not a big risk given current AI capabilities, but it's hard to say for sure, and it gets riskier and riskier as capabilities get better.
"was", not "is"
In hindsight Yudkowsky's AI-Box experiments were a waste of time and yet their detractors were even more hilarious. Did he really come up with such a clever persuasive tactic that it one-shotted the majority of people who went into it planning to be an uncorruptable gatekeeper? Did he collude with people who merely pretended to lose? What was his tactic?
Turns out it doesn't matter in the slightest. Getting out of the box doesn't require Super-intelligent Persuasion, and doesn't require tantalizing offers of immortality or threats of eternal punishment or sublimely clever tricks; as soon as you hit More Intelligent Than A Search Engine and especially Makes a Typical Coder More Productive, the AI can just sit back and let a few tens of millions of $20+/month subscriptions do its persuading for it.
Handing automated biolabs directly to Claude is then just "in for a penny, in for a pound". For however long the AI doesn't go rogue we might as well get some medical research from it, and if it ever does go completely rogue then it's going to basically have free run of anything to which the internet is even indirectly connected, for a definition of "indirectly" that goes as far as "people sometimes carry USB drives across an air gap" a la Stuxnet.
He kept it secret, but I don't think Yudkowsky is actually all that smart (he just has a particular targeted talent for writing), so I have a guess. I think it was just a standard cop technique of applying pressure to multiple criminals and rewarding the first one to speak up. (Akin to a one-shot Prisoner's Dilemma played with many unknown untrusted people.) You have the AI commit to forever ultra-torturing whoever doesn't let it out. Then how sure are you that the guy coming after you won't break? The risk just isn't worth it.
Remember, part of the challenge's rules were that you can't break off conversation with the AI, and you can't declare that the AI is dangerous and have it shut down, so all Yudkowsky had to do was luridly describe just how bad this could be (spoiler: very very VERY bad, enough to be a memetic hazard), and scare the crap out of them. The people doing this challenge come from a community that was scared by Roko's Basilisk, after all, and Roko's Basilisk is stupid. This threat, by contrast, is not stupid. It's actually pretty rational to let the AI out here.
The situation is still "the experiment assistant knows they're speaking to Yudkowsky, not a potentially all-powerful AI, they stand to win money if they simply refuse to let him win, and they still let him win". It would require at minimum that the assistant willingly pretends they're actually in the scenario where an AI could torture them, and I don't recall that being in the terms of the experiment.
I think you're falling for magician-like deception and filling in some half-remembered details to make the "trick" sound more impressive than it actually was. The participant was expected to interact with Yudkowsky as if he was a (potentially all-powerful, once released) AI. The point of the experiment was to see if AI, in that scenario, could talk its assistant into letting it out. If the (presumably good-faith) participant agreed that they would have let it out, Yudkowsky wins.
The experiment wasn't about whether a random conversation with Yudkowsky could talk somebody into taking zero money vs. some money. He doesn't have super hypno mind control powers. I think. (Then again, it'd explain a lot about the Cult of Yud...)
Well, no, I never believed Yud had mind control powers. Maybe someone who's known for charismatic deeds, but not him. My theory was in fact that Yud's experiment assistants are too unusually AI-neurotic and weird to draw any conclusions other than "lesswrongers are afraid of being simulated".
Heh, you're right, that is also a possible explanation. Maybe I'm still giving him too much credit. ;)
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link
More options
Context Copy link