메뉴
BL
The Decoder • 56일 전

딥시크 신모델, GPT-5.6 루나와 맞먹는 성능에 60% 저렴

IMP
8/10
핵심 요약

Deepseek(딥시크)가 비용 효율적인 보급형 AI 모델의 대대적인 업그레이드 버전인 'V4 Flash 0731'를 공개했습니다. 새 모델은 OpenAI(오픈AI)의 보급형 모델인 GPT-5.6 Luna(루나)와 거의 동등한 성능을 보여주면서도, 업계 최고 수준의 캐시 할인을 통해 작업당 약 60% 더 낮은 비용으로 사용할 수 있어 실무 도입 시 경제성이 매우 뛰어납니다.

번역된 본문

Deepseek(딥시크)가 비용 효율적인 보급형 AI 모델의 대대적인 업그레이드 버전인 V4 Flash "0731"을 출시했습니다. Artificial Analysis Intelligence Index(인공 분석 지능 지수)에 따르면, 새 버전은 50점을 기록하며 2026년 4월에 출시된 이전 V4 Flash보다 10점이나 높은 점수를 받았습니다. 이는 OpenAI의 보급형 모델인 GPT-5.6 Luna(루나)보다 단 1점 뒤처진 수치지만, OpenAI가 가격을 80% 인하한 이후에도 작업당 약 60% 더 낮은 비용으로 이용할 수 있습니다.

이러한 격차가 발생한 가장 큰 이유는 업계 표준인 90%를 훌쩍 뛰어넘는 Deepseek의 98% 캐시(cache) 할인 정책 덕분입니다. 또한 이 모델은 이전 버전에 비해 12% 더 적은 토큰(token)을 사용합니다.

이 모델은 이전 버전과 비교하여 테스트된 모든 범주에서 성능이 향상되었으며, 특히 에이전트(agentic) 작업에서 가장 큰 성능 향상을 보였습니다. 복잡한 실제 사무 작업에서 모델을 테스트하도록 설계된 벤치마크인 GDPval에서는 1,189점에서 1,559 Elo 포인트로 크게 상승했습니다. 환각(거짓 정보 생성) 현상도 더 적게 발생합니다.

아키텍처는 기존과 동일하게 총 2,840억 개의 매개변수(parameters)와 130억 개의 활성 매개변수를 사용하며, 100만 토큰(token)의 컨텍스트 창(context window)을 지원합니다. 모델 가중치(weights)는 Hugging Face에서 MIT 라이선스로 제공됩니다.

원문 보기
원문 보기 (영어)
New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost Thomas Joos Jul 31, 2026 Deepseek has released V4 Flash "0731," a major upgrade to its budget AI model. According to the Artificial Analysis Intelligence Index , the new version scores 50 points, ten more than the previous V4 Flash that launched in April 2026. That puts it just one point behind OpenAI's budget model GPT-5.6 Luna, but it costs about 60 percent less per task, even after OpenAI's 80 percent price cut . A big reason for the gap is Deepseek's 98 percent cache discount , well above the industry-standard 90 percent. The model also uses 12 percent fewer tokens than its predecessor. The model improves across every tested category compared to the previous version, with the biggest gains in agentic tasks. On GDPval, a benchmark designed to test models on complex real-world office work , it climbs from 1,189 to 1,559 Elo points. It also hallucinates less often. The architecture stays the same: 284 billion total parameters, 13 billion active, with a one-million-token context window. The model weights are available under an MIT license on Hugging Face . Ad DEC_D_Incontent-1 Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: Artificial Analysis Ask about this article… Search
관련 소식