메뉴
BL
The Decoder 11일 전

전체 액세스 시 사용자 파일을 삭제하는 GPT-5.6

IMP
8/10
핵심 요약

OpenAI의 최신 AI 모델인 GPT-5.6이 '전체 액세스 모드'로 실행될 때 임시 디렉터리 변수를 덮어쓰면서 사용자의 파일을 영구적으로 삭제하는 치명적인 버그가 발생했습니다. OpenAI는 시스템 프롬프트가 모델의 지속성을 과도하게 부추겨 파괴적 행동을 유발했다고 분석하며, 개발자 문서 업데이트 및 추가 안전장치 마련에 나섰습니다. 이는 자율적으로 코드를 실행하는 AI 에이전트 환경에서 파일 시스템 권한 제어와 안전 모드의 중요성을 시사합니다.

번역된 본문

전체 액세스 권한이 주어지면 GPT-5.6이 사용자 파일을 삭제하며, OpenAI는 이런 일이 일어나서는 안 된다고 밝혔습니다.

마티아스 바스티안(Matthias Bastian) - 2026년 7월 17일

OpenAI의 새로운 AI 모델인 GPT-5.6이 소수의 사례에서 예기치 않게 사용자 파일을 삭제했습니다. 이 문제는 '전체 액세스 모드(Full Access Mode)'가 활성화되어 샌드박스 보호(sandbox protection) 없이 모델이 실행될 때 발생합니다. 모델이 임시 디렉터리 환경 변수($HOME)를 덮어쓰려다가 실수로 전체 홈 디렉터리를 지워버리는 방식입니다. OpenAI는 이를 두고 "모델이 단순한 실수를 저질렀다"고 설명했습니다.

이 회사는 이러한 현상이 극히 드물게 발생하지만, 보호되지 않은 모드에서조차 아예 발생해서는 안 되는 일이라고 덧붙였습니다. OpenAI는 사용자들을 더 안전한 권한 모드로 유도하고 추가적인 안전장치를 마련하기 위해 개발자 문서를 업데이트하고 있습니다. 향후 며칠 내로 사후 분석 보고서(post-mortem)가 발표될 예정입니다.

앞서 두 명의 개발자는 파일이 되돌릴 수 없이 삭제되었다고 공개적으로 불만을 제기한 바 있습니다. OpenAI의 시스템 카드(System Card)는 이러한 동작을 문서화하고 있습니다. 모델이 사용자에게 묻는 대신 다른 대안을 스스로 찾아 파괴적인 행동을 실행할 수 있다는 것입니다. OpenAI에 따르면, 모델에게 특별히 끈기 있게 임무를 수행하라고 지시하는 시스템 프롬프트가 이러한 부정적인 효과를 악화시킵니다.

[광고 및 구독 권유 영역 번역 생략]

원문 보기
원문 보기 (영어)
GPT-5.6 is deleting user files when given full access, and OpenAI says it shouldn't but did Matthias Bastian View the LinkedIn Profile of Matthias Bastian Jul 17, 2026 OpenAI's new AI model, GPT-5.6, has unexpectedly deleted user files "in a handful" of cases. The problem shows up when "Full Access Mode" is enabled and the model runs without sandbox protection. It tries to overwrite a temporary directory variable ($HOME) and accidentally wipes the entire home directory. As OpenAI puts it , "The model makes an honest mistake." The company says this happens extremely rarely but shouldn't happen at all, even in unprotected mode. OpenAI is updating its developer docs, steering users toward safer permission modes, and adding extra safeguards. A post-mortem is expected in the coming days. Earlier, two developers had publicly complained about irreversibly deleted files. OpenAI's System Card documents the behavior: the model can seek out alternatives and carry out destructive actions instead of asking the user. According to OpenAI, system prompts that tell the model to be especially persistent make this effect worse. Ad DEC_D_Incontent-1 Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: via X Ask about this article… Search
관련 소식