마이크로소프트(Microsoft)는 다중 에이전트 보안 플랫폼인 MDASH 내에 통합된 새로운 보안 특화 AI 모델, MAI-Cyber-1-Flash를 발표했습니다. 이 모델은 기존 최고 수준의 모델들보다 뛰어난 취약점 발견 및 수정 성능을 보여주면서도 비용은 절반으로 줄여주며, 이를 통해 실시간 위협 모니터링과 패치가 가능한 자율형 보안 시스템인 Project Perception의 역량을 강화합니다.
번역된 본문
오늘 우리는 다중 에이전트 취약점 식별 및 수정 하네스(harness)인 MDASH 내부에 통합된 MAI-Cyber-1-Flash를 발표합니다. 이 결합은 최고 수준의 성능을 주도 모델들 대비 절반의 비용으로 제공합니다.
AI의 발전은 놀라웠지만, 그것이 촉발하는 새로운 세대의 사이버 위협 역시 마찬가지입니다. 공격자들은 이제 점점 더 강력한 능력을 행사하며, 자신들이 침투할 수 있는 단 하나의 약점을 찾기 위해 끊임없이 쌓여가는 방대한 코드를 탐색하고 있습니다. 취약점을 발견하는 비용이 급격히 낮아짐에 따라, 가끔씩 스캔하고 결국에는 패치하는 이전의 보안 방식은 이제 더 이상 쓸모가 없습니다. 우리가 AI의 진정한 이점을 활용하려면, 먼저 우리 모두가 세상이 돌아가는 소프트웨어를 강화하는 데 도움을 줄 수 있는 훌륭한 사이버 보안 모델을 구축해야 합니다.
이것이 복잡한 코드베이스에서 까다로운 취약점을 찾기 위해 구축된 MAI-Cyber-1-Flash의 개발 동기입니다. 이 모델은 MDASH에 깊이 통합되었으며, 업계 최고의 사이버 보안 전문가들에 의해 다듬어졌고 지구상에서 가장 큰 보안 환경 전체에 걸쳐 강화되었습니다. 이 결합된 전문성은 시스템이 대규모 코드베이스를 추론해 코드 내의 실제 취약점을 찾는 방식을 평가하는 황금 표준 벤치마크인 CyberGym에서 Mythos, Gemini, GPT를 능가하는 탁월한 보안 보호 기능을 제공합니다.
작업에 맞는 모델 선택
보안은 항상 켜져 있는 작업이며, 들어오는 공격의 막대한 양을 고려할 때 토큰 비용은 이제 방어자들에게 실질적인 제약이 되었습니다. MAI-Cyber-1-Flash는 모든 작업의 최대 90%를 효율적으로 처리하도록 설계되어, MDASH가 실제로 그것이 필요한 10%의 예외적으로 어려운 작업에 대해서만 당사의 보유 모델 중 더 크고 비싼 모델(이 경우 GPT-5.4)을 사용할 수 있게 합니다.
그 결과, MDASH와 MAI-Cyber-1-Flash의 통합 시스템은 CyberGym에서 96%의 달성률(Mythos 대비 +12pt)을 기록했습니다. 이러한 조합은 오늘날 MDASH에서 제공하는 최상의 조합(GPT 5.4 + 5.4 mini + 5.3 codex)과 비교할 때 50%의 비용 절감 효과를 제공합니다. 이것이 바로 독특하고 풍부한 과거 학습 데이터에 접근할 수 있는 잘 조정된 다중 모델 시스템의 힘입니다. 이를 통해 항상 모든 작업에 가장 좋은 가격으로 최고의 모델을 사용할 수 있습니다.
이 새로운 환경에서 새로운 취약점을 식별하는 것에서 실시간으로 이를 해결하는 것으로 넘어갈 수 있는 것이 중요합니다. 소프트웨어 취약점에 대한 AI 수정 작업이 핵심 보안 워크플로우가 된 반면, 보안 실무자들이 직접 수행해야 하는 많은 작업들이 남아 있습니다. 그래서 오늘 우리는 MDASH 내에서 다양한 보안 워크플로우를 위한 에이전트 팀을 제공하여 새로운 위협 벡터를 지속적으로 모니터링, 패치 및 차단하는 자율형 보안 시스템인 Perception을 출시합니다. 또한 Perception은 소프트웨어 취약점 작업을 넘어 곧 더 많은 보안 워크플로우에 MAI-Cyber-1-Flash를 사용하게 될 것입니다.
오늘날 중요한 세 가지는 다음과 같습니다: 모델(Model), 데이터(Data), 하네스(Harness)입니다. 우리는 고객이 독보적으로 강력한 보안 기능을 누릴 수 있도록 세계 최고 수준의 모델, 비할 데 없는 과거 데이터, 그리고 전문가가 조정한 하네스를 공동으로 최적화했습니다.
모델(Model): MAI-Cyber-1-Flash는 고품질 데이터를 기반으로 사내에서 처음부터 구축된 MAI-Thinking-1 계열에서 파생된 코드 중심의 간결한 보안 모델입니다. 자세한 내용은 기술 보고서를 참조하십시오.
데이터(Data): 우리의 가장 큰 장점입니다. 세계 최고 수준의 보안 시스템을 수십 년간 구축해 온 덕분에 이제 우리는 ID, 엔드포인트, 클라우드 및 네트워크에 걸친 매일 수조 개의 신호와 실제 해킹 및 수정에 대한 비할 데 없는 기록을 보유하고 있습니다. 이러한 역사는 그 누구도 인위적으로 만들어 낼 수 없습니다.
하네스(Harness): 당사의 다중 에이전트 취약점 식별 및 수정 하네스인 MDASH는 업계 최고의 보안 전문가들에 의해 조정되었습니다. 이들은 취약점을 찾고, 검증하고, 수정하기 위해 여러 주요 모델을 사용하는 100개 이상의 에이전트를 만들었습니다. 에이전트 기반 코드 스캐닝은 보안 운영 센터(SOC)의 핵심 기능이며, 우리의 새로운 자율형 보안 시스템인 Project Perception에 데이터를 제공합니다.
Copilot --> Models Introducing MAI-Cyber-1-Flash inside MDASH World-class security at half the cost Mustafa Suleyman & Hayete Gallot July 27, 2026 Models Mustafa Suleyman & Hayete Gallot LI X FB Today we're announcing MAI-Cyber-1-Flash inside of MDASH, our multi-agent vulnerability identification and remediation harness. Together they deliver world-class performance at 50% of the cost of leading models. Progress in AI has been startling and so has the new generation of cyber threats it's unleashing. Attackers now wield increasingly powerful capabilities, probing an ever-growing mountain of code for just a single weakness that lets them in. As the cost of finding a flaw collapses, the old model of security, where you scan occasionally and patch eventually, is now obsolete. If we're to unlock the true benefits of AI, we must first build outstanding cyber models that help all of us harden the software the world runs on. That's the motivation behind MAI-Cyber-1-Flash, which has been built to find challenging vulnerabilities in complex codebases. It's been deeply integrated into MDASH, honed by the best cybersecurity experts in the industry and hardened across the largest security estate on the planet. This combined expertise delivers exceptional security protection, beating Mythos, Gemini and GPT on CyberGym, the gold standard benchmark for evaluating how systems reason over large codebases to find real vulnerabilities in the code. Picking the right model for the task Security is an always-on mission, and given the enormous volume of inbound attacks, token cost is now the real constraint for defenders. MAI-Cyber-1-Flash was designed to efficiently handle up to 90% of all tasks, enabling MDASH to use the larger and most costly models in our fleet (in this case GPT-5.4) for the 10% of exceptionally hard tasks that truly need them. The result is that the unified system of MDASH with MAI-Cyber-1-Flash delivers 96% on CyberGym (+12 pt above Mythos) . This combination delivers a 50% cost saving when compared against our best offering in MDASH today (GPT 5.4 + 5.4 mini + 5.3 codex). That's the power of a well-tuned, multi-model system with access to uniquely rich historical training data. It ensures you always have the best model at the best price for every task. In this new environment, being able to go from identifying a new vulnerability to addressing it in real-time is critical. And while AI remediation of software vulnerabilities is now a key security workflow, there are many jobs to be done by Security practitioners themselves. That’s why today we’re also launching Perception , our agentic security systems, that provides teams of agents for a variety of security workflows in MDASH, to continuously monitor, patch, and close new threat vectors. Perception will also soon use MAI-Cyber-1-Flash for many more security workflows, beyond the software vulnerability work. Three things matter today: Model. Data. Harness. We have jointly optimized our world-class models, our unmatched historic data, and our expert-tuned harness to ensure that our customers have a uniquely powerful security offering. Model. MAI-Cyber-1-Flash is a compact, code-heavy security model derived from the MAI-Thinking-1 lineage, which was built from scratch, in-house, on the highest quality data. Details in our technical report . Data. Our deepest advantage. Decades of building world-class security systems now give us trillions of daily signals across identity, endpoint, cloud, and network, and an unmatched record of real exploits and remediations. No one can manufacture this history. Harness. MDASH , our multi-agent vulnerability identification and remediation harness, is tuned by the best security experts in the industry, who have created 100+ agents using multiple leading models to find, validate, and remediate vulnerabilities. Agentic code scanning is a critical function in the Security Operating Center and feeds Project Perception, our new agentic security system. Built with safety first Because MAI-Cyber-1-Flash is Microsoft's first cyber model, we built trust into every layer of the system, from model training to customer deployment. The model was developed with a security-first calibration, rigorously evaluated by Microsoft's AI Red Team, tested through automated and expert-led adversarial exercises, and independently assessed by a third party. Trust extends beyond the model itself. Through MDASH, customers get enterprise-grade controls including Role-Based Controls, tenant isolation, encryption, auditability, and sandboxed execution environments with no internet access. The result is a cyber model that delivers powerful capabilities to defenders while maintaining the governance, security, and control enterprises expect from Microsoft. Our hill-climbing machine Cybersecurity is not just a data-rich domain; it is a live reinforcement learning loop. Every day, defenders investigate threats, triage alerts, hunt adversaries, remediate vulnerabilities, deploy protections, and learn from the outcome. Microsoft sees that loop end to end: vulnerabilities through Microsoft Security Response Center ; attacks and defenses across identity, endpoint, cloud, data, browser, and applications; more than 100 trillion security signals every day ; and operational insight from 1.6 million customers. Because we can connect actions to outcomes; what was exploitable, what was contained, what was blocked, and what actually worked; we have more than data. Our MAI reinforcement learning loop gives us the foundation to build cyber models that improve continuously and become expert cyber defenders. That'll remain our commitment to our customers for years to come. Build the Future With Us We’re a lean, talent-dense team of explorers, researchers, and full-stack engineers. We move fast, sweat the details, and operate at frontier scale with a roadmap to build the world’s most powerful AI models. Most importantly, we’re united by the belief that doing this right is the only way to do it at all. If our mission resonates with you, we’d love to talk. Explore all jobs Related Stories Hill-climbing MAI models for GitHub Copilot and Excel models Introducing MAI-Image-2.5-Pro and MAI-Voice-2-Flash models Bringing Ode Poetry to life with MAI’s audio models partnerships