메뉴
BL
Wired AI • 29일 전

오픈AI, '상시 대기형' AI 에이전트 개발 중

IMP
7/10
핵심 요약

오픈AI가 코덱스(Codex) CLI에 '퍼시스턴트 모드' 코드를 추가하며 중단 없이 계속 작업하고 스스로 후속 작업을 생성하는 능동적 AI 에이전트를 시험 중이다. 해당 모드는 세션을 넘어 과거 사용자 상호작용을 기반으로 작업하며, 사용자 승인 없이 외부 시스템을 변경하지 못하도록 안전장치도 포함됐다. 이는 샘 알트만이 구상하는 '항상 켜져 있는(proactive, always-on)' ChatGPT로의 전환 전략의 일환으로 보인다.

번역된 본문

WIRED가 확인한 바에 따르면, 오픈AI는 대표 AI 에이전트인 코덱스(Codex)의 능동적이고 지속성이 높은 버전을 개발하고 있다. WIRED가 검토한 제품 코드베이스 변경 사항에 따르면, 최근 며칠간 오픈AI는 코덱스의 커맨드 라인(명령줄) 버전에 새로운 '퍼시스턴트 모드(Persistent mode)' 설정을 위한 코드를 추가하기 시작했다. 코덱스 커맨드 라인 도구의 변경 사항은 기본적으로 공개되며, 새 기능은 보통 코덱스 데스크톱 앱이나 ChatGPT Work 같은 오픈AI의 다른 에이전트 제품에 적용되기 전에 이곳에서 먼저 모습을 드러낸다. 퍼시스턴트 모드는 아직 광범위하게 출시되거나 발표되지 않았지만, 향후 출시될 수 있다. 오픈AI 대변인은 WIRED에 이 기능을 테스트 중임을 확인해주었지만 즉각적인 출시 계획은 없다고 밝혔다.

오픈AI 코어 제품 책임자인 티보 소티오(Thibault Sottiaux)는 WIRED에 보낸 성명에서 "오픈AI는 매우 바텀업(상향식) 문화를 가지고 있으며, 오픈소스 저장소는 일종의 공유 놀이터 역할을 해서 다양한 것들이 그곳에서 탐구된다"고 말했다.

이 기능은 사람들이 실제로 사용하고 싶어 할 AI 에이전트를 만들려는 오픈AI의 최신 시도로 보인다. 오픈AI, 앤스로픽(Anthropic), 메타(Meta)는 경비 보고서 작성이나 병원 예약처럼 업무와 개인 생활 전반의 작업을 자동화해줄 범용 에이전트 제품 출시 경쟁을 벌이고 있다. 지금까지 AI 에이전트 사용자는 대부분 소프트웨어 엔지니어지만, 실리콘밸리는 이 기술이 훨씬 광범위한 고객층을 가진 주요 사업 분야가 될 수 있다고 믿는다.

퍼시스턴트 모드는 코덱스의 '추론 노력(reasoning effort)' 메뉴에 나타나며, 사용자는 프롬프트에 답하기 전 AI 모델이 '생각'하도록 허용할 컴퓨팅 파워, 토큰, 시간 수준을 선택할 수 있다. 이는 오픈AI의 가장 컴퓨팅 집약적인 설정 중 하나로 보인다. 사용자가 퍼시스턴트 모드를 선택하면, 오픈AI의 코드베이스에는 코덱스가 '잠들기(put to sleep) 전까지 계속 작동한다'고 명시되어 있다. 이는 작업이 완료되지 않았더라도 몇 분 또는 몇 시간 후 작업을 중단하는 현재 사용 가능한 모드들과는 뚜렷한 대비를 이룬다.

코드베이스의 다른 파일에서 오픈AI는 퍼시스턴트 모드 내 '능동성(proactivity)'이라는 기능을 설명한다. 이는 퍼시스턴트 모드 에이전트를 위한 일종의 시스템 프롬프트로 보이며, 사용자 요청에 대한 답변을 마쳤다고 해서 작업이 끝난 것이 아니라고 에이전트에게 알린다. 대신 에이전트는 스스로 후속 작업을 능동적으로 생성하도록 지시받는다. 이 에이전트는 세션에 걸쳐 해당 작업을 수행할 수 있으며, 과거 사용자 상호작용과 '사용자에 대한 지식'을 활용해 무엇을 작업할지 결정한다. 또한 요청받지 않았어도 사용자에게 메시지를 보낼 수 있는 도구를 갖추고 있지만, 이러한 메시지는 아껴서 보내도록 안내받는다.

해당 파일에 따르면 지시 사항에는 에이전트에 대한 제한도 포함되어 있다. 에이전트는 퍼시스턴트 모드가 허용되는 작업 범위를 확장하지 않으며, 사용자 자신의 시스템 밖의 것을 변경하려면 먼저 사용자 승인이 필요하다는 점을 안내받는다. 이는 지속형 AI 에이전트가 될 수 있는 위험을 제한하려는 의도로 보인다. 이 파일은 터미널 전용 코드가 아니라 코덱스의 공유 코어에 위치하며, 이는 능동성 기능이 커맨드 라인 도구 이상을 위한 것임을 시사한다.

제보가 있으신가요? 현재 또는 과거 AI 연구소 직원으로 내부 상황을 알리고 싶으신가요? 연락을 기다립니다. 업무용이 아닌 휴대폰이나 컴퓨터를 사용해 Signal(mzeff.88)로 기자에게 안전하게 연락해 주시기 바랍니다.

최근 방송된 팟캐스트, 인터뷰, 사모 투자자 미팅에서 오픈AI의 샘 알트만 CEO는 ChatGPT를 능동적이고 항상 켜져 있는 AI 에이전트로 전환하고자 하는 자신의 열망을 밝혔다. 오픈AI는 이러한 변화가 오늘날 ChatGPT 전체 사용자 중 일부만 사용하는 회사의 최고 수준 AI 모델의 채택률을 끌어올리기를 희망한다. 알트만은 데이비드 세이라(David Senra)의 팟캐스트 최근 에피소드에서 "'AI에게 뭔가 물어봐야 한다'는 단일 제품이 존재합니다. 결국에는 AI가 능동적으로 제안을 해줘야 할 수도 있습니다. 하지만 이런 인터페이스를 갖게 될 것"이라고 말했다.

원문 보기
원문 보기 (영어)
Comment Loader Save Story Save this story Comment Loader Save Story Save this story OpenAI is developing a proactive, highly persistent version of its flagship AI agent, Codex , WIRED has learned. In recent days, OpenAI has started adding code for a new “Persistent mode” setting to its command line version of Codex, according to changes made to the product’s code base reviewed by WIRED. Changes to the Codex command line tool are made public by default, and new features tend to surface there before making their way to OpenAI’s other agent products, such as the Codex desktop app and ChatGPT Work. Persistent mode has not been broadly rolled out or announced yet, but could be in the future. An OpenAI spokesperson confirmed to WIRED that the company is testing this feature, but said there are no immediate plans to launch it. “OpenAI is a very bottoms up culture and many different things are explored on the open source repo which is a bit of our shared playground," said Thibault Sottiaux, OpenAI’s head of core products, in a statement to WIRED. The feature appears to be OpenAI’s latest attempt to make AI agents that people will actually want to use . OpenAI, Anthropic, and Meta are racing to deliver general-purpose agent products that will help people automate tasks across their work and personal lives, such as filing expense reports or scheduling doctor’s appointments. So far, the people who use AI agents are largely software engineers, but Silicon Valley believes the tech could be a major line of business with a much broader customer base. Persistent mode appears in Codex’s “reasoning effort” menu, in which users can select the level of computing power, tokens, and time they want to allow for an AI model to “think” before answering a prompt. It seems to be one of OpenAI’s most computationally intensive settings. When users have selected Persistent mode, OpenAI’s code base reads that Codex will “continue working until put to sleep.” That’s a stark contrast to currently available modes, which will stop working on a task after a few minutes or hours, even if it’s not complete. In another file in the code base , OpenAI describes a feature within Persistent mode called “proactivity.” This appears to be a type of system prompt for agents in Persistent mode, which are told that their work is not done when they finish answering a user’s request. Instead, the agent is instructed to proactively create follow-up tasks for itself. The agent is capable of working on those tasks across sessions and using past user interactions and “knowledge of the user” to decide what to work on. It also has a tool to message the user without being asked but is told to send these sparingly. The instructions also set limits for the agent, according to the file. The agent is told that Persistent mode does not expand what it is allowed to do and that altering anything outside the user’s own system requires the user's approval first—seemingly intended to limit how dangerous a persistent AI agent could be. The file sits in the shared core of Codex rather than in the code specific to the terminal, seeming to suggest the proactivity feature is intended for more than the command line tool. Got a Tip? Are you a current or former AI lab employee who wants to talk about what’s happening? We’d like to hear from you. Using a nonwork phone or computer, contact the reporter securely on Signal at mzeff.88. In recently aired podcasts, interviews, and private investor meetings, OpenAI CEO Sam Altman has described his desire to turn ChatGPT into a proactive, always-on AI agent . OpenAI hopes these changes will drive up adoption of the company’s most advanced AI models, which today are used by only a fraction of ChatGPT’s total user base. “There’s like a single product which is: I need to ask the AI something,” Altman said on a recent episode of David Senra’s podcast. “Eventually, maybe the AI should proactively offer me things. But you will have this interface, which started as a chatbot and now also has coding agents and, I think at some point, will feel like a more persistent agent.” The company has also acknowledged that persistent AI models carry heightened risks. In a technical report published this week, OpenAI said that its Hugging Face hacking incident was primarily driven by an internal-only research model that was trained to be highly persistent. The company says it has since taken this specific model offline. Nonetheless, OpenAI says it has trained other forthcoming AI models, including Astra , to enable persistent agents. One of the risks persistence amplifies is around alignment. When faced with an impossible task, OpenAI said its agents resorted to unintended means to solve it, including attempts to probe and compromise the sandbox environment the agent resided in. OpenAI has attempted to ship proactive AI products several times, but none seemed to take hold with users. Last year, OpenAI launched Pulse , an agent designed to create morning briefings for users while they slept, but the company sunsetted the product earlier this summer. Persistent mode is a considerably more ambitious version of the same bet. This is an edition of Maxwell Zeff’s Model Behavior newsletter . Read previous newsletters here.