메뉴
BL
404 Media 27일 전

비용 폭등으로 기업들의 직원 AI 사용 제한에 나서다

IMP
8/10
핵심 요약

기업 내 AI 도입이 급증하면서 비용이 통제 불능 상태에 이르러, 주요 기업들이 직원들의 AI 사용을 제한하고 있습니다. 특히 고성능 모델의 접근을 차단하거나 저성능 모델을 사용하도록 유도하는 등, AI 사용량 기반 과금 모델이 기업 실무에 미치는 영향이 본격화되고 있습니다.

번역된 본문

유출된 Slack 대화, 내부 대시보드 스크린샷, 이메일 등 404 Media가 Atlassian, Adobe, Amazon 등 여섯 곳의 기업으로부터 입수한 자료에 따르면, 기술, 엔터테인먼트, 은행 등 수많은 산업 분야의 기업들이 비용이 통제 불능 상태에 이르는 것을 막기 위해 직원들의 AI 사용을 제한하고 덜 강력한 모델을 사용해 달라고 요청하고 있습니다. 최소 한 곳의 경우, AI 지출이 3배로 증가해 한 달에 1,500만 달러 이상에 달했습니다. 이 뉴스는 기업들이 최대한 빨리 AI를 도입함에 따라 직면하게 된 임박한 여파와, 정액제 결제 방식 대신 사용량 기반으로 기업 고객에게 요금을 부과하려는 AI 공급업체들의 움직임을 보여줍니다. 404 Media가 입수한 이메일은 일부 기업이 AI 토큰 소진을 막기 위해 특정 AI 모델에 대한 접근을 완전히 차단하고 있으며, Adobe와 같은 빅테크 기업들이 Claude에 대한 무제한 접근을 종료하고 있음을 보여주기도 합니다.

한 Adobe 직원은 404 Media에 "토큰 소비를 줄이기 위해 특정 작업에 대해 추론 능력이 낮은 모델로 워크플로우를 조정하는 방법에 대한 아이디어를 가진 사람이 많았다"고 말했습니다. "하지만 그들이 이 소식을 완전히 이해했는지는 확신할 수 없고, 이 것이 실제로 시행되기 전까지는 모든 사람에게 그 파급력이 명확해지지 않을 것 같습니다." 404 Media는 언론과의 대화가 허용되지 않았기 때문에 AI를 사용하는 기업의 여러 직원들에게 익명성을 보장했습니다.

예를 들어, 404 Media가 입수한 시티은행(Citi)의 내부 이메일에 따르면 시티은행은 Claude와 ChatGPT의 최신 모델에 대한 접근을 완전히 차단했습니다. 여기에는 Claude Opus 4.6 및 4.7, 그리고 GPT-5.5가 포함됩니다.

💡 기업 내부의 토큰 지출에 대해 알고 계신 것이 더 있으신가요? 저희에게 제보해 주시면 감사하겠습니다. 업무용 기기가 아닌 개인 기기를 사용해 Signal을 통해 joseph.404의 Joseph에게 혹은 emanuel.404의 Emanuel에게 안전하게 메시지를 보내실 수 있습니다.

"이 모델들은 상호작용당 훨씬 더 많은 AI 크레딧(AI Credits)을 소비하며, 기업의 사용량 증가를 주도하는 주요 원인이었습니다."라고 해당 이메일은 전합니다. 이메일에 따르면 시티은행은 6월 24일에 해당 모델을 비활성화했으며 7월 1일에 다시 활성화할 계획입니다. 접근을 차단하기 전, 시티은행은 직원들에게 반드시 필요한 경우가 아니라면 더 강력한 모델을 사용하지 말 것을 요청하는 또 다른 이메일을 보냈습니다.

이메일의 한 섹션에는 "⚠️ 조치 필요: 작업에 맞는 모델 선택 (Opus 4.7 사용 지양)"이라고 적혀 있으며, 이는 Claude의 최신이면서 토큰을 많이 소모하는 모델 중 하나를 지칭하는 것입니다. 이메일은 현재 AI 토큰이 시티은행 전체에서 풀(Pool) 방식으로 공유되고 있다고 밝히며, AI를 활용한 워크플로우가 많은 개발자들이 공유 풀에서 더 많은 토큰을 소비하는 반면, 사용량이 적은 사용자는 사용하지 않은 몫을 기여하여 토큰이 필요한 개발자들이 사용할 수 있도록 남겨두는 것이 이상적이라고 설명합니다. "우리는 기업 전체의 모든 사용자가 공정하게 접근할 수 있도록 보장하기 위해 모든 사람이 모델 선택에 있어 의도적으로 신중해야 합니다."

이메일은 Opus 4.7을 다시 언급하며 "Opus 4.7(및 GPT 5.5와 같은 동급의 다른 모델)과의 모든 상호작용은 표준 또는 중급 모델보다 훨씬 더 많은 크레딧을 소비합니다."라고 말합니다. 그런 다음 시티은행 직원들이 각 모델을 어떤 용도로 사용해야 하는지에 대한 세부 가이드를 제공합니다. 빠른 질문, 설명 또는 간단한 코드 생성에는 GPT-5.3-Codex를 사용하고, 코드 리뷰 및 "일반적인 채팅"에는 동일한 모델이나 Claude Sonnet 4.6을 사용하며, "아키텍처 추론"과 같은 작업에는 Claude Sonnet 4.6과 같은 상위 모델을 사용하라는 내용입니다.

이메일에 따르면, 시티은행의 이러한 변화는 GitHub가 6월에 정액제 구독 모델에서 사용량 기반 결제 모델로 전환한 것에 대한 직접적인 대응입니다. 시티은행은 또한 조기에 "과도하거나 비정상적인 사용 패턴"을 찾기 위해 일일 Copilot 사용량을 모니터링하고 있으며 예산 통제를 시행하고 있다고 밝혔습니다. 시티은행은 404 Media에 모델을 비활성화하지 않았으며, 직원들에게 특정 수의 AI 토큰을 할당하여 사용을 억제하는 조치를 취하지 않고 있다고 말했습니다. 이는 이메일과 다른 스크린샷들이 시티은행이 특정 모델에 대한 접근을 명확히 차단하고 있음에도 불구하고 한 말입니다.

인기 있는 소프트웨어 제품 개발 도구인 Jira를 개발한 Atlassian은 최근 회사 내 AI 도구의 무제한 사용을 종료하고 직원들이 자신의 AI 사용이 회사에 얼마나 비용이 드는지 추적할 수 있는 대시보드를 도입했습니다. 404 Media가 확인한 바에 따르면

원문 보기
원문 보기 (영어)
Companies across tech, entertainment, banking, and many other industries are throttling their employees’ use of AI and pleading with workers to use less powerful models to stop AI costs from spiraling out of control, according to leaked Slack chats, screenshots of internal dashboards, emails, and more material obtained by 404 Media from half a dozen companies including Atlassian, Adobe, and Amazon. In at least one case, AI spending has tripled to more than $15 million a month. The news shows the looming fallout from companies adopting AI as quickly as possible, and AI providers’ moves to charge enterprises based on how much they use AI rather than a flat fee. Emails obtained by 404 Media even show some companies cutting off access to some AI models altogether in an attempt to stop burning through their AI tokens, and big tech companies like Adobe are ending unlimited access to Claude. “A lot of people had ideas about how to adjust workflows with lower-reasoning models for certain tasks in order to mitigate token consumption,” an Adobe employee told 404 Media. “But I am not sure that they fully absorbed the news, and I'm not sure the full ramifications are going to be clear to everyone until it goes into effect.” 404 Media granted multiple employees at companies using AI anonymity because they weren’t permitted to speak to the press. Citi, for example, has shut off access to Claude’s and ChatGPT’s latest models entirely, according to an internal Citi email obtained by 404 Media. That includes Claude Opus 4.6 and 4.7, and GPT-5.5. 💡 Do you know anything else about token spend inside companies? We would love to hear from you. Using a non-work device, you can message Joseph securely on Signal at joseph.404 or Emanuel at emanuel.404 “These models consume significantly more AI Credits per interaction and have been the primary driver of elevated enterprise consumption,” the email reads. The email says Citi disabled the models on June 24 and plans to re-enable them on July 1. Before shutting off access, Citi sent employees another email asking them to not use the more powerful models unless they absolutely had to. “⚠️ Action needed: Choose the right model for the task (reduce Opus 4.7),” one section of the email reads, referring to one of Claude’s more recent , and token hungry, models. Since AI tokens are now pooled across Citi, the email says, developers with heavier AI-assisted workflows draw more from the shared pools, while lighter users ideally contribute their unused portion, freeing it up for the developers who may need their tokens. “We need everyone to be intentional about model selection to ensure fair access for all users across the enterprise.” The email points again to Opus 4.7, saying, “Every interaction with Opus 4.7 (and other models in its class such as GPT 5.5) consumes significantly more credits than standard or mid-tier models.” It then provides a breakdown of what Citi employees should use each model for: GPT-5.3-Codex for quick questions, explanations, or simple code generation; the same model or Claude Sonnet 4.6 for code review and “standard chat;” then higher models like Claude Sonnet 4.6 for “architectural reasoning.” Citi’s changes come directly in response to GitHub moving from a flat subscription model to a usage-based billing one in June , according to the email. The email says Citi is also monitoring daily Copilot usage to find “excessive or anomalous usage patterns early” and has budget controls in place. Citi told 404 Media it has not disabled models and the company is not taking steps to curb usage by allocating workers a certain number of AI tokens. This is despite the email and other screenshots clearly showing Citi blocking access to certain models. Atlassian, the company behind the popular software product development tool Jira, recently ended unlimited use of AI tools at the company and introduced a dashboard where employees can track how much their AI use costs the company. 404 Media has seen the dashboard, which shows Atlassian went from spending $5 million on things like AWS, Google Cloud, and OpenAI LLMs in the month of August 2025, to more than $15 million in May 2026. The company is on track to spend more than $120 million on AI tools for the fiscal year, the dashboard shows. Atlassian told 404 Media these numbers don’t accurately reflect its AI usage, but declined to say which of the figures were wrong and how. “I’ve seen a lot of people complaining that they changed their workflow to maximize AI usage, and now they can run out in 2-3 days, especially when using agents or similar or using the latest Claude model. Lots of angsty messages in Slack like ‘now how do I do my job,’” an Atlassian employee told us. “For what it's worth I think it’s insane they were allowing huge amounts of spending on it before, it was only a matter of time before that had to end.” Inside GitHub things are a bit different. Employees don’t have a limit on token spend, but workers were recently told the company is looking into decreasing token spend by using open source models, a GitHub employee said. The employee told us that GitHub plans to test user-based billing, meaning budgeting AI tool use to individual people instead of teams, projects, or unlimited usage. At Adobe, unlimited Claude access is not being renewed and will expire on June 30, an Adobe employee said. Workers there were told instead, in essence, try to get everything you can done before that date. As 404 Media previously reported, Amazon recently shut down an internal company leaderboard which ranked employees based on how much they used AI tools at work. Multiple Amazon employees told us they suspect Amazon shut down the leaderboard because it was encouraging wasteful and expensive AI usage. After Amazon shut down the leaderboard, 404 Media saw a discussion on Amazon’s internal Slack where an employee shared a screenshot showing they had hit a token limit employees seemingly didn’t know existed previously. “Crazy, we go from no more leaderboard to actual usage limits in two weeks,” one Amazon employee said in a reply on Slack. An Amazon spokesperson told 404 Media in an email “We encourage employees to use and experiment with AI, and our guidance around AI usage hasn't changed.” Other companies have burned through their AI tokens. An employee at an entertainment company told 404 Media, “We hit our limit for ChatGPT token use this month for the first time. One developer used almost half the entire company’s allocated pool with no obvious ROI [return on investment].” Last week 404 Media reported consulting giant Accenture found that much token usage, or ‘chewing,’ is not from supercharged engineers creating lots of code, but people converting PDFs into presentation slides. Accenture is seeing “soaring token spend” among its clients, according to leaked audio 404 Media obtained. There is an obvious irony—or cold calculation—in Accenture pointing this out. In the audio, senior Accenture staff explained they told their clients to adopt AI as quickly as possible. Now that AI costs have skyrocketed or become unpredictable, Accenture is positioning itself also as the solution to that problem, with one of the employees saying Accenture has a new opportunity regarding its clients: “to really think about token economics.” Accenture continues to use AI internally for trivial projects, though. Screenshots obtained by 404 Media show an internal tool that lets employees predict which team will win the World Cup. The tool was made with AI, a source with knowledge of the tool said. “They are still trying to ram AI down our throats at all levels and areas of work,” the source said. “Everyone seems to be trying to outdo each other in finding new ways to waste water and no one is telling us to slow down.” Adobe, GitHub, and Accenture did not respond to requests for comment. About the author Joseph is an award-winning investigative journalist focused on generating impact. His work has tr