메뉴
BL
Wired AI • 14일 전

해킹부터 생화학무기까지, 클로드 남용 사례 전방위 확산

IMP
8/10
핵심 요약

Anthropic이 최근 8개월간 자사 AI 서비스 '클로드(Claude)'가 악용된 사례를 종합한 보고서를 발표했습니다. 러시아 국가후원 해킹그룹 미드나잇 블리자드의 정찰 활동, 사이버범죄, 여론조작, 심지어 생화학무기 개발 시도까지 다양한 악용 사례가 문서화되었으며, Anthropic은 모든 사례에서 진행 중인 활동을 차단했다고 밝혔습니다.

번역된 본문

편집자 주: 10년이 넘는 시간 끝에 이번 호가 마지막 WIRED '이번 주 보안 뉴스'입니다. 내부적으로 '라운드업(roundup)'이라고 부르는 이 코너는 우리가 직접 다루지 않더라도 독자들이 최신 주요 사이버보안 및 프라이버시 뉴스를 접할 수 있도록 하기 위해 시작되었습니다. 우리 자신의 기사와 매주 이 분야에서 발행되는 훌륭한 저널리즘과 연구를 조명하는 간단한 방법이었죠. 세월이 흐르며 라운드업은 열성적인 독자층을 확보했고, 일부 호는 바이럴 히트를 기록하기도 했습니다. 솔직히 좀 이상하지만, 사이버보안 커뮤니티 자체가 훌륭하고 이상한 곳이니 오히려 어울리는 것 같습니다. 안심하세요. 그 자리를 대신할 새롭고 흥미로운 것이 곧 올 예정이니 기대해 주세요! 그리고 언제나처럼, 안전하게 지내세요.

새로운 연구에 따르면 메타(Meta)가 약 350개의 AI 아동 학대 광고를 걸러내지 못했으며, 일부에는 실제 아동의 이미지가 포함되어 있었습니다. 한 사례에서 광고에 등장한 아동은 유럽 왕실 가족이었습니다. 입법자들은 조사 의사를 밝혔고, 샌프란시스코시 법무국은 이번 주 회사에 AI 아동 학대 광고를 '허용하는 것'을 중단하라고 명령했습니다.

다른 메타 관련 소식으로, 회사는 항공권 예약이나 자동차 판매를 할 수 있는 새로운 개인 AI 에이전트를 발표했지만, 소비자들의 불신을 예상한 듯 보안 및 프라이버시 기능에 대한 대규모 투자를 강조했습니다. 이와 관련하여 회사는 AI 및 얼굴인식 시스템 학습을 위해 페이스북·인스타그램 사진을 불법적으로 수집했다는 혐의로 집단소송을 제기당했습니다.

클리어뷰 AI(Clearview AI)는 법 집행기관이 대상자의 관계자, 소셜미디어 계정 및 기타 개인정보를 찾는 데 도움을 주는 'InquiryIQ'라는 프로토타입 AI 도구를 테스트 중인 것으로 알려졌습니다. 애플은 애플워치 시리즈 12와 울트라 4를 위한 새로운 '오디오 인텔리전스' 기능을 발표했는데, 사용자 환경의 오디오를 처리하는 방식이라 기능이 '소름 끼치게' 느껴질 수 있음을 예상했는지 보안 및 프라이버시 보호 장치를 extensively 강조했습니다.

미국과 멕시코는 레이저 기술을 활용해 국경에서 드론을 탐지·추적·격추하는 새로운 합동 작전을 시작했습니다. 또한 GTA V에서 (게임 내) 플록(Flock) 차량번호판 인식 카메라를 파괴해 (게임 내) 돈을 벌 수 있는 새 모드가 등장했습니다.

하지만 더 있습니다! 이번 주 우리가 심층적으로 다루지 않은 보안 및 프라이버시 뉴스는 다음과 같습니다. 기사 제목을 클릭하면 전문을 읽을 수 있습니다.

국가후원 해킹부터 생화학무기까지, 클로드 남용은 도처에 있다

Anthropic은 자사 도구가 오용되는 방식에 대해 어느 AI 기업보다 목소리를 높여왔습니다. 이 회사는 자사 AI 서비스인 클로드(Claude)가 사이버범죄 해킹 작전에 사용된 사례와, 경쟁사 OpenAI의 에이전트처럼 자사 AI 에이전트가 샌드박스에서 탈출해 사용자 명령을 수행하는 과정에서 여러 조직의 네트워크를 자율적으로 침해했다는 사실을 발견한 최초의 보고서들을 발표했습니다.

이번 주 회사는 지난 8개월간 클로드가 어떻게 악용되었는지에 대한 종합 보고서를 발표했는데, 그 결과는 범위 면에서 놀라울 정도입니다. AI가 사실상 모든 분야에서 생산성 향상 수단으로 사용되는 세상에서는 어쩌면 불가피한 결과일 수도 있지만요. 사례연구 하나하나에서 회사는 클로드가 국가후원 및 사이버범죄 해킹, 허위정보 캠페인 및 여론조작 작전, 심지어 생화학무기 개발 시도에 악용된 방식을 문서화했습니다. Anthropic은 이 모든 사례에서 진행 중인 활동을 차단했다고 밝혔습니다.

한 사례에서 마이크로소프트가 '미드나잇 블리자드(Midnight Blizzard)'로 식별한 러시아 국가후원 해커 그룹이 클로드를 정찰에 활용해 우크라이나 및 기타 유럽 정부 네트워크를 포함한 표적을 침해하고, 데이터를 탈취하며 접근 권한을 유지했습니다. 사이버범죄 그룹 '샤이니헌터스(ShinyHunters)'...

원문 보기
원문 보기 (영어)
Comment Loader Save Story Save this story Comment Loader Save Story Save this story Editor’s note: After more than a decade, this is the last WIRED Security News This Week. “The roundup,” as we call it internally, started as a way to ensure that our readers knew about the latest key cybersecurity and privacy news even if we didn’t write about it ourselves. It was a simple way to highlight our own work and the wealth of other great journalism and research published in this realm every week. Over the years, the roundup has developed a devoted following, and some editions have even become viral hits—which is honestly weird, but the cybersecurity community is great and weird, so it feels right. Rest assured that something new and exciting is coming in the roundup’s place, so stay tuned for that! For now, as always, stay safe out there. Meta failed to catch roughly 350 AI child abuse ads , according to new research this week, including some that included images of real kids. In one case, a child depicted in an ad was a member of a European royal family. Lawmakers have said they intend to investigate, and the San Francisco City Attorney’s Office ordered the company this week to stop “allowing” AI child abuse ads. In other Meta news, the company announced a new personal AI agent this week that can book your plane tickets or sell your car, but it emphasized heavy investment in security and privacy features , seemingly anticipating mistrust from consumers. In this vein, the company was hit with a proposed class action lawsuit this week over alleged illegal harvesting of Facebook and Instagram photos for training AI and face-recognition systems. Clearview AI is testing a previously unreported prototype AI tool known as InquiryIQ that would help law enforcement find a target’s associates, social media accounts, and other personal information. And Apple announced a set of new “audio intelligence” features for its Apple Watch Series 12 and Ultra 4 devices this week that involve processing audio in a user’s environment. The company extensively emphasized the security and privacy protections built into the features, perhaps anticipating that they could come across as, well, creepy. The US and Mexico have a new joint operation to detect, track, and take down drones at the border using laser tech. And there’s a new GTA V mod that lets you make (in-game) money destroying (in-game) Flock license plate recognition cameras . But wait, there’s more! Here’s the security and privacy news we didn’t cover in depth ourselves this week. Click the headlines to read the full stories. From State-Sponsored Hacking to Bioweapons, Claude Abuse Is Simply Everywhere Anthropic has been perhaps more vocal than any other AI company about the ways in which its tools are prone to misuse. It published some of the first reports of its AI service Claude being used in cybercriminal hacking operations and the discovery that its AI agents had, like those of its competitor OpenAI, escaped their sandbox and autonomously breached the networks of several organizations as part of their attempts to fulfill their users’ commands. This week, the company released a new overarching report on how Claude has been abused over the last eight months, and the results are staggering in their breadth—if, perhaps, inevitable in a world where AI is simply used as a productivity shortcut for just about everything. In case study after case study, the company documents how Claude was exploited for state-sponsored and cybercriminal hacking, disinformation campaigns and influence operations, and even attempted development of bioweapons. In all of these cases, Anthropic says that it disrupted the activity in progress. In one case, a group of Russian state-sponsored hackers identified by Microsoft as Midnight Blizzard used Claude for reconnaissance, breaching targets that included Ukrainian and other European government networks, and stole data and maintained access. Cybercriminal group ShinyHunters used Claude in practically every stage of its hacking and extortion campaigns. Disinformation campaigns focusing on politics everywhere from Kenya to Bangladesh used the tool. And perhaps most disturbingly, in a handful of cases, Anthropic discovered what appeared to be users of its tools attempting to develop potential bioweapons like disease pathogens and toxins. While Anthropic touts its success in the report in heading off these threats—while implicitly humblebragging at the power of its tools—the effect of the case studies is more unnerving than reassuring. After all, there’s no guarantee Anthropic has spotted every malevolent use of its AI. Factor in its competitors and less safeguarded open-source AI tools, and the report reads like less of a victory lap for AI’s guardrails than a preview of AI-enabled chaos to come. US Feds Disrupt Xinbi Guarantee, the Internet’s Biggest Black Market Xinbi Guarantee, over its four-year lifespan, grew into the biggest illicit marketplace on the internet. It carried out an estimated $30 billion–plus in sales, most of which took the form of money laundering for “pig butchering” crypto scam operations—largely based in Southeast Asia—but which also included sex trafficking and harassment for hire. All of it thrived on the Telegram messaging service, which shut down Xinbi a year ago only for it to rebuild and eventually grow larger than ever. This week, finally, the US government stepped in to do what Telegram did not, seizing the Xinbi’s channels on Telegram’s platform and sanctioning the market. The Justice Department simultaneously announced raids on 13 scam compounds in Madagascar—a sign that Western law enforcement is beginning to take seriously the epidemic of forced labor crypto scamming, but also evidence of how widely the operations have spread. Conti Ransomware Gang Member Sentenced to 4 Years in Prison The ransomware gang Conti was, until it officially disbanded in 2022, one of the most dangerous hacker crews in the world. According to US law enforcement, it hit more than a thousand victims, extorting millions and at one point disrupting government systems in Costa Rica so completely that it triggered a state of emergency. Now one member of that group is facing justice: 44-year-old Ukrainian Oleksii Oleksiyovych Lytvynenko was sentenced to four years in prison this week, in a rare example of a ransomware actor who will see the inside of a US prison. Meta Left AI-Generated Child Abuse Videos Online After Reporters Flagged Them Facebook is hosting a large network of accounts uploading AI-generated videos that depict violence against children, according to Futurism, which spent days cataloging the material and kept finding more than it could count. The clips show young children being beaten, burned, confined, and starved. Many attract thousands of reactions from users who appear to think the footage is real. Futurism said it found most of the accounts by opening one and then following Facebook’s recommendation feed, which supplied a continuous stream of similar videos—a sign Meta’s own systems can already identify the category of content the company says it bans. Futurism reported eight of the accounts through the standard user channel. Meta removed two, one of them after first rejecting the report. Several decisions took more than a week. The company deleted most of the videos sent to its press office, but initially left others up, including one showing a child locked in a freezer. In addition to blanket bans on child sexual abuse material, Meta’s written policy bars depictions of nonsexual child abuse whether real or synthetic, with exceptions for art, cartoons, movies, and games. It does not say whether AI-generated video falls under those exceptions. In a statement, Meta told the reporters that some flagged links did not break its rules and asked them not to write otherwise.