메뉴
HN
Hacker News 41일 전

미 국방부, 의회 보고서 작성에 AI 활용 자랑

IMP
8/10
핵심 요약

미 국방부가 매년 의회에 제출해야 하는 방대한 양의 의무 보고서를 생성형 AI를 활용해 작성하고 있다고 밝혔습니다. 실무자들은 작업 시간을 200시간에서 5시간으로 단축할 수 있다며 AI 도입의 효율성을 강조했지만, 인간의 검토가 없을 경우 발생할 수 있는 오류와 환각(Hallucination)으로 인한 리스크도 함께 지적되고 있습니다.

번역된 본문

미 국방부는 매년 다양한 국가 안보 주제에 대해 수백 건의 보고서를 의회에 제출해야 하는 의무 과제를 안고 있습니다. 하지만 펜타곤 관계자들은 의회 보고서를 작성하기 위해 생성형 AI 도구를 사용하는 새로운 지름길을 자랑스럽게 소개하고 있습니다.

미 국방부 최고기술책임자(CTO)인 에밀 마이클(Emil Michael)은 6월 12일 워싱턴 DC의 허드슨 연구소 주최 행사에서, 트럼프 행정부 시절 '전쟁부(Department of War)'로 불렸던 국방부가 생성형 AI를 어떻게 도입했는지 보여주는 핵심 사례로 AI가 작성한 의회 보고서를 언급했습니다. 국방부는 2025년 12월부터 구글 클라우드의 '제미니 포 버번먼트(Gemini for Government, Gemini for Government)'를 시작으로 독자적인 GenAI.mil 플랫폼을 통해 6개 군 병력 모두에게 AI 도구를 폭넓게 제공하고 있습니다.

마이클은 "매년 의회에 이 사안에 대해 보고해야 한다"며, "관련 서류들을 모두 업로드해 의회 보고서 초안을 작성하게 했는데, 원래라면 200시간의 인력이 필요했을 작업을 5시간 만에 해냈다"고 말했습니다.

이러한 AI 활용에 대한 추가적인 증거는 지난 4월 23일 워싱턴 DC에서 열린 박스 연방 정상 회닩(Box Federal Summit)에서 제이콥 글래스만(Jacob Glassman) 미 국방부 과학기술기반 차관보의 이전 발언에서도 나타났습니다. 디펜스스쿱(DefenseScoop)의 보도에 따르면, 글래스만은 의무 보고서를 제출해야 했으나 인력이 부족했던 팀에게 "GenAI.mil을 사용해서 최선을 다해 보라"고 지시했다고 밝혔습니다. 일주일 후 팀은 돌아와 해당 AI 생성 보고서가 "지난 5년간 우리가 쓴 보고서 중 최고"라고 주장했다고 합니다. 단, 디펜스스쿹은 글래스만이 구체적으로 어떤 보고서인지 밝히지 않았다고 전했습니다.

미 국방부는 의회에 이러한 보고서를 효율적이고 적시에 전달하는 데 오랫동안 어려움을 겪어 왔습니다. 특히 의회가 통과시키는 새로운 국방 수권 법안마다 의무 보고서의 수가 증가하는 추세입니다. 미국 정부책임국(GAO)에 따르면 보고서 수는 2000년 500여 건에서 2020년에는 1,400건 이상으로 급증했습니다.

전 정부책임국 최고 임원이었던 엘리자베스 필드(Elizabeth Field)는 2023년 연방뉴스네트워크와의 인터뷰에서, 입법담당차관보실 관계자들이 최신 보고 요건을 찾기 위해 국방수권법을 "거의 줄 단위로" 검토해야 한다고 말했습니다. 그녀가 작성한 GAO 보고서에 따르면, 펜타곤이 보고 요건을 식별하고 적절한 팀에 보고서를 할당하는 까다로운 과정에는 3~6개월이 걸릴 수 있으며, 일부 의회 의무 보고서는 1년 이내에 제출해야 합니다.

AI 도입을 서두르는 위험성 이처럼 지루한 과정을 고려할 때, 현재 펜타곤의 지도부가 AI 생성 보고서를 매력적인 지름길로 여기는 것은 놀라운 일이 아닙니다. 하지만 로펌이나 주요 컨설팅 펌과 같은 다른 조직들 역시 적절한 인간의 검토와 감독 없이 오류투성이인 AI 작성물에 의존할 때 발생하는 수많은 함정을 이미 경험한 바 있습니다.

최근의 경고 사례 중 하나는 다국적 컨설팅 거대 기업 KPMG가 기업의 AI 활용에 대한 보고서를 발행했다가, 수많은 AI 생성 오류와 허위 주장이 포함된 사례 연구가 포함되어 있었던 사건입니다. 이는 연구 그룹 GPTZero에 의해 밝혀졌고 파이낸셜 타임스가 보도했습니다. 이 사건으로 인해 KPMG는 '에이전틱 AI 시대의 탁월함 재정의(Redefining excellence in the age of agentic AI)'라는 제목의 보고서를 철회하게 되었습니다.

현재 펜타곤이 의회에 제출하는 AI 생성 보고서의 정확성을 검토하기 위해 어떤 절차를 마련해 두고 있는지는 불분명합니다. 하지만 이러한 보고서는 미군이 납세자의 돈을 어떻게 사용하는지 책임을 묵기 위한 의회의 감시를 위한 핵심 요소입니다. 따라서 AI로 인한 오류나 잘못된 서술은 감시 체계를 훼손할 수 있습니다.

원문 보기
원문 보기 (영어)
Text settings Story text Size Small Standard Large Width * Standard Wide Links Standard Orange * Subscribers only Learn more Minimize to nav The US Department of Defense has a lot of congressionally mandated homework to do every year involving hundreds of required reports on various national security topics. But Pentagon officials have been proudly describing a new shortcut—using generative AI tools to write such reports for Congress. Pentagon Chief Technology Officer Emil Michael highlighted AI-generated reports to Congress as a key example of how the Department of Defense—stylized as the Department of War under the Trump administration—has adopted generative AI during an event hosted by the Hudson Institute think tank in Washington, DC, on June 12. The Pentagon has made AI tools, starting with Google Cloud’s Gemini for Government, widely available to members of all six military branches through the department’s bespoke GenAI.mil platform since December 2025. “I have to report to Congress every year on this thing,” Michael said. “Let me load all the papers onto it and have it draft me a congressional report that would otherwise take 200 hours of staffing time and do it in five hours.” More evidence of such AI usage came from previous comments by Jacob Glassman , deputy assistant secretary of defense for science and technology foundations at the US Department of Defense, during the Box Federal Summit held in Washington, DC, on April 23. According to DefenseScoop coverage, Glassman described how he told a short-staffed team responsible for delivering a congressionally mandated report to “use GenAI.mil, do the best you can.” The team supposedly came back to Glassman a week later, claiming that the AI-generated report was “the best report we’ve written in the past five years.” As DefenseScoop notes, Glassman did not identify the report in question. The Department of Defense has long struggled to deliver such reports to Congress efficiently and in a timely manner, especially as the number of mandated reports generally rises with every new defense appropriations bill passed by Congress. The number of reports had soared from just over 500 reports in 2000 to more than 1,400 reports by 2020, according to the US Government Accountability Office . Officials at the Office of the Assistant Secretary of Defense for Legislative Affairs typically have to go through defense authorization statutes “almost line by line” to find the latest reporting requirements, said Elizabeth Field , former senior executive director at the Government Accountability Office, in a Federal News Network interview in 2023. Her GAO report showed how the Pentagon’s painstaking process of identifying the reporting requirements and assigning reports to the appropriate team could take between three and six months—and some of the Congressionally mandated reports are due within a year. The perils of pushing AI adoption Given that tedious process, it’s not surprising that the Pentagon’s current leadership may find AI-generated reports to be a tempting shortcut. But other organizations, such as law firms and major consulting firms , have already discovered the many pitfalls of relying on error-ridden AI-generated writing without adequate human vetting and oversight. One of the latest cautionary tales involved the multinational consulting giant KPMG publishing a report about AI use in businesses that featured case studies with numerous AI-generated errors and false claims , as revealed by the research group GPTZero and reported by the Financial Times. The revelations led KPMG to pull the report titled “Redefining excellence in the age of agentic AI.” It’s unclear what processes the Pentagon has in place to review the accuracy of its AI-generated reports to Congress. But such reports are a crucial element of congressional oversight intended to hold the US military accountable for how it uses taxpayer dollars—and so any AI-induced errors or mischaracterizations could undermine the accountability mechanism of such reports. This also comes at a time when the Pentagon has requested an unprecedented $1.5 trillion budget for the 2027 fiscal year. Members of the US military have also been using generative AI tools to write personnel evaluation reports for non-commissioned officers and commissioned officers, generate commendation medal citations, and create counseling statements, according to a Small Wars Journal article. The number of Department of Defense personnel using commercial AI tools such as Gemini through GenAI.mil has significantly increased from just 80,000 in December 2025 to 1.5 million in June 2026, the Pentagon CTO claimed during his remarks at the Hudson Institute. The Department of Defense has an overall workforce of approximately 3.5 million. Google is among multiple US tech companies that signed agreements in 2025 with the US General Services Administration to make their AI tools available across federal government agencies for deeply discounted prices. On May 1, the Department of Defense announced new agreements with “eight of the world’s leading frontier artificial intelligence companies” to deploy more AI tools on classified networks for “lawful operational use.” Those companies include SpaceX , OpenAI , Google , Nvidia , Reflection AI, Microsoft , Amazon Web Services , and Oracle . The US government has not divulged how much it is paying the companies under the new contracts. But the list notably excludes Anthropic, which was blacklisted by the Trump administration after the tech company supposedly refused to allow its Claude AI models to be used in an unrestricted manner for autonomous warfare and mass surveillance. Jeremy Hsu Tech Reporter Jeremy Hsu Tech Reporter Jeremy Hsu is a reporter exploring a wide range of topics across deep tech and AI. He has previously written for New Scientist, Scientific American, IEEE Spectrum, Wired, Undark Magazine and MIT Tech Review, among many other publications, about topics such as deepfakes, data centers, drones, battery tech, robotics, and GPS jamming. He also has a Master of Arts in Journalism from NYU, and a bachelor's degree from University of Pennsylvania in History and Sociology of Science, with a minor in English. 96 Comments