메뉴
HN
Hacker News • 9일 전

AI 안전 논의의 뿌리는 사실 '섹스 컬트'라는 주장

IMP
4/10
핵심 요약

이 글은 AI 안전(AI Safety) 운동의 지적 중심이 엘리저 유드코프스키(Eliezer Yudkowsky)와 그가 창시한 소위 '섹스 컬트'적 커뮤니티에서 비롯되었다고 주장합니다. 저자는 유드코프스키가 DeepMind 창업 자금 연결, OpenAI 창설 계기, 'AI 얼라인먼트(AI Alignment)' 용어 대중화 등 AI 업계의 핵심 흐름에 깊이 개입했으며, 효과적 이타주의(Effective Altruism)와 합리주의(rationalism) 커뮤니티가 사실상 동일하다고 지적합니다. 나아가 이런 인물들이 AI 정책 수립에 관여해서는 안 된다고 비판합니다.

번역된 본문

돌아왔습니다, 여러분이 그리웠어요. 제가 너무 잘 아는 중요한 주제이기에, 'AI 안전은 대부분 섹스 컬트다'를 설명하는 스레드를 준비했습니다. 저는 이 사람들이 정책을 만들어서는 안 된다고 생각합니다. (1/?)

(대체 제목: 컬트 이론 좀 해볼 시간)

수년에 걸쳐 AI 안전 논의와 정책을 그 기원이자 지적 중심으로부터 분리하려는 대대적인 노력이 있었습니다. 왜냐하면 그 중심은 팬픽션 작가가 시작한 섹스 컬트이기 때문입니다. 진지하게 대우받으려면 그 사실을 숨겨야 하죠.

이 사람이 엘리저 유드코프스키("Yud")입니다. 그는 버니 샌더스를 만나 AI 정책에 대해 자문했고, 현재 넷플릭스에 있는 AI 다큐에서 AI 정책을 놓고 논쟁하는 모습을 볼 수 있습니다.

이 사람은 OpenAI의 최고위급 직원이자 비공식 대변인인 Roon입니다. Roon이 유드에 대해 갖는 견해는 1) 그가 섹스 컬트에 속해 있거나 운영하고 있다는 것, 2) 그럼에도 그의 AI에 대한 의견은 매우 중요하고 진지하게 받아들여야 한다는 것입니다.

유드는 셰인 레그와 데미스 허사비스를 피터 틸에게 연결해줘 현재 구글의 AI 연구 부서가 된 DeepMind의 창업 자금을 확보하게 한 인물입니다.

샘 알트먼이 유드에 대해 갖는 견해는 이렇습니다.

이것은 2015년 1월 푸에르토리코에서 열린 Future of Life Institute의 AI 컨퍼런스 일정입니다. 일론 머스크는 그해 12월 OpenAI를 창립했습니다. 유드코프스키는 이 컨퍼런스에서 벌어진 일들이 그 계기가 되었다고 주장하며 그에 대해 분노해 있지만, 이는 사실처럼 보이지 않습니다.

현재 Anthropic의 CEO인 다리오 아모데이는 2013년에 유드의 비영리단체 블로그에서 '효과적 이타주의 커뮤니티의 일원'으로 소개된 모습이 여기 있습니다. 그는 유드와 '클립 최대화자(paperclip maximizer)'에 대해 논쟁을 벌였습니다. Anthropic 논문들은 지금도 가끔 유드를 인용하며, 우리가 AI에 대해 이야기하는 방식 전체가 그의 비영리단체가 대중화한 어휘를 사용합니다.

'AI 얼라인먼트(AI Alignment)'라는 표현은 이렇게 일관되게 사용한 최초의 장이었던 유드코프스키의 비영리단체가 대중화한 것입니다. 유드의 사상이나 래트/EA(rationalist/Effective Altruism) 씬을 암묵적으로 거스르지 않고서는 AI를 논하기가 매우 어렵습니다.

최근 Anthropic을 떠나 내부고발자로 자리매김한 콕슨이 스크린샷에 찍혀 있습니다.

이 부분을 거의 잊을 뻔했네요: 여기 2023년 타임지에서 자신의 취미 주제를 위해 핵전쟁 위험을 감수하자고 주장하는 엘리저 유드코프스키가 있습니다. 샘 알트먼의 집에 화염병 투척을 시도한 남자는 유드코프스키를 주요 영감의 원천으로 언급했습니다. 이에 대해 이야기하는 녹음이 어딘가에 굴러다니니 관심 있는 사람이 있으면 말씀하세요. 하지만 현실적으로 생각해보면, 누가 감히 서툴게 샘 알트먼의 집에 화염병을 던지겠습니까?

'섹스 컬트' 부분으로 반드시 돌아오겠다고 약속합니다. 사실 이건 설명하기 가장 쉬운 부분입니다. 그들이 이에 대해 끊임없이 게시하거든요. 어려운 것은 이것이 본질이라는 사실을 부정하는 방대하고 복잡하며 자금이 충분한 장벽을 넘는 것입니다.

이 그림은 사실 농담이 아닙니다. 유드코프스키의 추종자들은 요즘 자신들을 '합리성 커뮤니티'라고 부르고, 외부에서는 그들을 '합리주의자(rationalists)'라 부르며 그 추종자들을 '합리주의(rationalism)'라 합니다. 아마도 이렇게 부르면 덜 컬트처럼 들린다고 생각하는 모양입니다. 저는 개인적으로 그 매우 긴 '우리는 종교가 아닙니다 설명'이 우습기도 하고 모욕적이기도 합니다.

아, 이 수표도 현금화해 두겠습니다: 수천 개의 컬트를 낳은 팬픽션은 '해리 포터와 이성주의의 방법(Harry Potter and the Methods of Rationality)'입니다. 이것은 현대의 단일 종교 텍스트로는 아마 가장 큰 영향력을 가진 것이며, 그 영향력이 아직 정점에 이르지 않았을 수 있습니다.

웃긴 연표입니다.

(en.wikipedia.org - 해리 포터와 이성주의의 방법 - 위키피디아)

소문에 따르면 '효과적 이타주의(Effective Altruism)'는 별개의 것이며 유드에게 지시를 받지 않는다고 합니다. 이것은 기본적으로 헛소리입니다. 사상과 사람이 크게 겹치는데, 특히 샌프란시스코에서는 더욱 그렇습니다. 또한 Anthropic은 효과적 이타주의 기업이 아니라고 들었습니다. 이것은 100% 헛소리입니다.

버니와 유드가 나온 그 영상? 그 방에 함께 있던 사람들: 제프리 라디시(Palisade Research, Coefficient 지원)와 다니엘 코코타일로(AI 2027 저자, Survival and Flourishing Fund 지원). 모두 '합리주의자'가 아닌 '효과적 이타주의자'들이지만, 그들 모두 이 씬에 끼어 있었습니다.

원문 보기
원문 보기 (영어)
I'm back, I missed you all. Since this is an important subject I know too much about, here's a 🧵 explaining that AI Safety Is Mostly A Sex Cult. I don't think these people should make policy. (1/?) (alternative title: Time For Some Cult Theory) There has been a great effort over many years to distance AI Safety discussion and policy from its origins and intellectual center, because its center is a sex cult started by a fan fiction author. If you want to be taken seriously you have to hide that. This is Eliezer Yudkowsky ("Yud"). We see him here meeting with Bernie Sanders to advise him on AI Policy and arguing about AI Policy in The AI Doc, now on Netflix. This is very senior OpenAI employee and unofficial OpenAI spokesman Roon. Roon's opinion of Yud is that 1) he is in and/or running a sex cult and 2) his opinions about AI are very important and should be taken seriously. Yud is the person who connected Shane Legg and Demis Hassabis to Peter Thiel so that they could secure funding to start DeepMind, which is now Google's AI research division. Here's Sam Altman's opinion of Yud. This is the schedule for the Future of Life Institute's conference on AI in Puerto Rico in January, 2015. Elon Musk would found OpenAI that December. Yudkowsky claims that somehow events at this conference caused that, and that he's upset about it, but this seems unlikely. Dario Amodei, current CEO of Anthropic, is here in 2013 identified as a "member of the Effective Altruism community" on the blog of Yud's non-profit. He argues with yud about the "paperclip maximizer". Anthropic papers still cite Yud on occasion, and the entire way we talk about AI uses words his non-profit popularized. "AI Alignment" was popularized by Yudkowsky's non-profit, which was the first venue to use this exact phrasing consistently. It is very hard to discuss AI at all without implicitly invoking Yud's ideas or the rat/ea scene. Coxon, who recently left Anthropic and positioned himself as a whistleblower, screenshotted here. I almost forgot this part: Here's Eliezer Yudkowsky in Time Magazine in 2023, advocating in favor of risking nuclear war over his hobby subject. The guy who tried to firebomb Sam Altman's house cited Yudkowsky as his primary inspiration. I have a recording of him talking about this laying around somewhere if anyone cares, but let's be real, who else is going to ineptly firebomb Sam Altman's house? I promise I will come back to the "sex cult" part, but that's actually the easy part to explain because they post about it constantly. What's difficult is getting past the extensive, complicated and well-funded wall of denial that this is the main event. This picture isn't really a joke, though Yudkowsky's followers call themselves "the rationality community" these days, externally they are called "rationalists", the followers "rationalism". Apparently they think this sounds less cultlike. I personally find the very long "we are not a religion explanation" either funny or insulting. Oh, just so I have cashed this check: The fan fiction that launched a thousand cults is Harry Potter and the Methods of Rationality. It is possibly the highest-impact single modern religious text, and its influence may not yet have peaked. Silly timeline. en.wikipedia.org Harry Potter and the Methods of Rationality - Wikipedia en.wikipedia.org Allegedly, "Effective Altruism" is a different thing and doesn't take marching orders from Yud. This is basically bullshit. The ideas and people heavily overlap, especially in San Francisco. I am also told that Anthropic is not an Effective Altruist company. This is 100% bullshit. That video with Bernie and Yud? Also in the room: Jeffrey Ladish (Palisade Research, Coefficient funded) and Daniel Kokotajlo (AI 2027 author, funded by Survival and Flourishing Fund). All "effective altruists", not "rationalists", but they all got in on that together. Where there is any separation between the two things, "rationalism" is team Yud and "effective altruism" is currently team Dario, and they're competing for attention and funding in that same space. The way they do this is extremely funny. Yudkowsky's most viral idea, the one that infects the most people and spreads the farthest, is that AI, if we continue to develop it, will inevitably 1) achieve superintelligence ("RSI"), 2) escape human control 2) kill everyone ("paperclip them", "doom") All arguments about AI Safety and AI Alignment are primarily about mitigating this exact scenario under his exact terms. Significant modification is relatively rare. In the broad subculture, this is implicitly what "safety" and "alignment" mean. People who think they're Effective Altruist and not Rationalist, but who are deeply concerned with "AI Safety", are generally implicitly or explicitly concerned with Yudkowsky's basic framework for loss of control and doom. This idea seems to be the most infectious because it activates the same neurosis that every apocalyptic or millenarian cult does. The world is going to end, and only I and my special group of people in the know can maybe stop it. (Also, you should give me money and have sex with me.) You can argue with them in any way you like but it barely matters. It's not, mostly, something people believe for either rational or empirical reasons, and they will simply go back to complaining about the apocalypse as soon as you stop paying attention. In person they are gathered mostly in various communes in Berkeley, California. They call their communes "group houses", maybe to make it sound less cultish. Most of the details of this are what you expect from "a commune in Berkeley". Everyone is polyamorous, everyone attends the same events, everyone lives in group homes together, everyone venerates the same few authors, and everyone is very worried about AI doom. They are all very contrarian together. Lots and lots of them work at AI companies and non-profits. The safety, red teaming and alignment teams seem especially afflicted, because that's the area you go on to worry about and try to work against AI doom, which you are sure is at least medium likely. The crown jewel of this whole thing, though, is Lighthaven. It's the former Rose Garden Inn in Berkeley, and it was bought by one of the rationalist charities in late 2022 with SBF's money, which was looted from FTX for them just before it collapsed. There are maybe five people who could have physically pushed through wiring money to a rationalist charity from FTX when it was collapsing, and one of them is Leopold Aschenbrenner. Leopold married one of the others, and she's Chief of Staff at Anthropic. Leopold is, allegedly, an effective altruist and not a rationalist, but you can see how this gets very difficult to keep straight. In what year were you one and not the other? Was your organization, at the time, financially tied directly to Yudkowsky or no? Lighthaven hosts the local Astral Codex Ten reading group with Scott Alexander, a reading group for The Sequences, which is essentially the Rationalist Old Testament, a conference for AI people called The Curve, an EA conference called EAGxBerkeley, MATS, aka "Machine Alignment, Transparency, and Security or ML Alignment & Theory Scholars" and SlutCon, organized by Aella. Gretta, pictured here, is Eliezer Yudkowsky's primary partner, and draws a salary of about 200k from his nonprofit as his assistant while also holding this event at LightHaven. She has no prior qualifications for any job at his nonprofit One hopes no illegal sex work happened at SlutCon, and the presence of Eliezer's partner as an organizer is just fun and quirky. I also don't know that any of the paid orgies, which Gretta also organizes, ever happened at LightHaven. I do note that Lighthaven seems desperate for money. Lighthaven's facilities manager is Ronn