메뉴
BL
The Decoder • 43일 전

딥시크, V4-Pro 업그레이드 및 에이전트 오픈소스화

IMP
8/10
핵심 요약

딥시크(Deepseek)가 성능을 개선한 V4-Pro 모델을 정식 출시하고, 자체 에이전트 소프트웨어인 'Deepseek Harness'를 오픈소스로 공개했습니다. 이와 동시에 중국 영업시간(피크 타임)을 기준으로 한 시간제 차등 API 요금제를 도입하여 전반적인 사용료를 인상했습니다.

번역된 본문

딥시크(Deepseek)가 메인 제품을 테스트 단계에서 벗어나 정식 출시했으며, 자체 에이전트 소프트웨어를 오픈소스로 공개하는 동시에 API 가격 인상을 발표했습니다.

deepseek-v4-pro 엔드포인트는 이제 빌드 버전 V4-Pro-0813을 제공합니다. 모델 이름, 매개변수(parameter) 수, 100만 토큰의 컨텍스트 윈도우는 변경되지 않았으며, 기존에 연동된 환경은 수정 없이 계속 작동한다고 딥시크는 밝혔습니다. 앱과 웹 환경에서는 '전문가 모드(Expert Mode)'를 통해 해당 모델을 사용할 수 있습니다. 새롭게 추가된 기능으로는 OpenAI Responses API와 Codex 통합에 대한 네이티브 지원이 있습니다. 추론 노력(reasoning effort)은 '낮음(low)', '높음(high)', '최대(max)'의 3단계로 설정할 수 있으며, 딥시크는 일상적인 에이전트 사용 시에는 중간 단계를 권장하고 있습니다.

딥시크의 자체 비교 테이블에 따르면, Terminal Bench 2.1 점수는 72.1에서 87.9로, DeepSWE 점수는 12.8에서 62.7로 크게 상승했습니다. 여러 에이전트 벤치마크에서 이 모델은 Claude Opus 4.8을 뛰어넘었습니다. 평가 기관인 Artificial Analysis 역시 이러한 성능 향상을 인정하면서도 전반적인 순위를 분석했습니다. V4-Pro는 지능 지수(Intelligence Index)에서 45점에서 53점으로 상승하며 GLM-5.2와 동점을 기록했습니다. 하지만 이는 여전히 57점인 Muse Spark, 58점인 Qwen 3.8 Max, 60점인 Kimi K3에 뒤처진 수치입니다. 현재 Claude Opus 5가 63점으로 1위를 차지하고 있습니다. 딥시크는 아직 새로운 빌드의 가중치(weights)를 공개하지 않았으며, Hugging Face에는 여전히 4월 프리뷰 버전이 올라와 있는 상태입니다.

이번 업데이트는 더 작은 크기의 V4 Flash 모델이 플래그십 모델의 성능에 근접하고 있는 상황에 대응하기 위한 것이기도 합니다. 7월 말, 딥시크가 V4 Flash를 위한 업데이트 0731를 출시하면서 저렴한 비용으로도 Artificial Analysis 지능 지수에서 Pro 프리뷰 버전과 거의 동일한 성능을 보여주었기 때문입니다.

MIT 라이선스로 공개되는 에이전트 하네스(Harness) 모델 업데이트와 함께 'Deepseek Harness v0.1'이 개발자 프리뷰 버전으로 MIT 라이선스 하에 공개되었습니다. 이 오픈소스 에이전트 소프트웨어는 OpenAI의 Codex와 Claude의 대안으로 제시됩니다. 이 소프트웨어는 새롭게 출시된 Cordis 플러그인 시스템을 기반으로 구축되었으며, 도구, 샌드박스, 세션, 사용자 인터페이스(UI)에 이르기까지 모든 기능을 교체 가능한 플러그인 형태로 제공합니다. 지속적인 세션 로그는 모든 프롬프트, 도구 호출 및 결과를 추적합니다. 작업을 일시 중지했다가 다시 시작하거나, 분기(branch)를 나누거나, 과거의 실행을 다시 재생하는 것도 가능합니다. 최소 모드(Minimal mode)는 셸(shell)과 파일 에디터만 남기고 모두 축소하며, 딥시크는 자체적인 벤치마크 테스트를 진행할 때 이 설정을 사용합니다. 로컬 웹 인터페이스를 통해 npx로 실행되지만, 호환성 문제가 있을 수 있다고 딥시크는 경고했습니다. 이 프로젝트는 2026년 3월 퀀트 트레이딩 기업 제인 스트리트(Jane Street)에서 합류한 쿠이 톈이(Cui Tianyi)가 이끌고 있습니다. 팀이 8월 초에 베타 테스터를 모집한 지 불과 3일 만에 712개의 프로젝트가 지원했습니다.

API 가격 인상, 특히 캐시 적중 시 요금 상승 주요 새로운 요금제는 8월 16일 UTC 기준 오후 4시부터 적용됩니다. 딥시크는 6월 말에 피크 및 오프피크 시간제 요금제로 전환한다고 발표했지만, 당시에는 구체적인 금액이나 날짜를 공유하지 않았습니다. 시간제 요금제는 2025년 2월부터 시행되어 왔으며, 당시 회사는 야간 시간대에 V3 및 R1 모델에 대한 할인을 제공했었습니다.

오프피크(비성수기) 시간대 사용 요금은 절반으로 책정됩니다. 피크 타임은 UTC 기준 새벽 1시부터 4시, 오전 6시부터 10시까지로, 이는 중국의 근무 시간대와 일치합니다. 유럽 사용자의 경우 오후 시간대의 대부분이 할인이 적용되는 오프피크 요금 구간에 해당합니다. 오프피크 시간 동안 V4-Pro의 입력(input) 요금은 백만 토큰당 0.435달러에서 0.66달러로 인상되었으며, 출력(output) 요금은 0.87달러에서 1.98달러로 크게 뛰었습니다. 반면, 피크 타임의 요금은 두 배로 증가합니다.

원문 보기
원문 보기 (영어)
Deepseek ships improved V4 Pro, open-sources its agent software, and raises API prices Jonathan Kemper View the LinkedIn Profile of Jonathan Kemper Aug 13, 2026 Nano Banana Pro prompted by THE DECODER Key Points Deepseek has released an updated version of its V4-Pro model. It scores higher on agent benchmarks but still trails top models like Claude Opus 5 in overall rankings. The company is also releasing "Deepseek Harness," open-source software that turns language models into autonomous agents through a modular plugin system. API prices are going up, with new time-based rates that make usage outside Chinese business hours cheaper. Repeated data retrievals are getting pricier. Ask about this article… Search Deepseek has moved its flagship product out of the testing phase, released its proprietary agent software as open source, and announced higher API prices at the same time. The deepseek-v4-pro endpoint now delivers build V4-Pro-0813 . The model name, parameter count, and one-million-token context window remain unchanged, and Deepseek says existing integrations will keep running without any tweaks. In the app and on the web, the model is available under "Expert Mode." A new addition is native support for the OpenAI Responses API with Codex integration. Reasoning effort can be set to three levels: "low," "high," and "max," with Deepseek recommending the middle setting for everyday agent use. According to Deepseek's own comparison table, Terminal Bench 2.1 scores jumped from 72.1 to 87.9, and DeepSWE scores went from 12.8 to 62.7. On several agent benchmarks, the model beat Claude Opus 4.8. Ad Artificial Analysis backs up the improvement but also puts it in context. V4-Pro climbs from 45 to 53 on the Intelligence Index, tying GLM-5.2. That's still behind Muse Spark at 57, Qwen 3.8 Max at 58, and Kimi K3 at 60. Claude Opus 5 sits at the top with 63 points. Deepseek hasn't published the weights for the new build yet, and the April preview version is still up on Hugging Face . Ad The update was also a response to the smaller V4 Flash model closing in on the flagship. At the end of July, Deepseek shipped update 0731 for V4 Flash , which practically matched the Pro Preview on the Artificial Analysis Intelligence Index while costing a fraction of the price. Agent harness ships under the MIT license Alongside the model update, Deepseek Harness v0.1 is shipping as a Developer Preview under the MIT license. The open-source agent software is pitched as an alternative to OpenAI's Codex and Claude. It's built on the newly released Cordis plugin system , where all features are swappable plugins, from tools and sandboxes to sessions and the UI. A continuous session log tracks every prompt, every tool call, and every result. Runs can be resumed, branched, and replayed. Ad Minimal mode strips things down to the shell and file editor, and Deepseek uses this setup for its own benchmark runs. The software launches via npx through a local web interface, though Deepseek warns of compatibility issues. The project is led by Cui Tianyi, who joined Deepseek from quantitative trading firm Jane Street in March 2026. When the team put out a call for beta testers in early August, 712 projects signed up within three days . API prices are going up, especially for cache hits The new rates kick in on August 16 at 4:00 p.m. UTC. Deepseek announced the switch to peak and off-peak pricing at the end of June but didn't share specific figures or a date at the time. Time-based rates have been around since February 2025 , when the company offered a discount on V3 and R1 during nighttime hours. Ad Off-peak usage costs half as much. Peak hours run from 1 a.m. to 4 a.m. and 6 a.m. to 10 a.m. UTC, lining up with the Chinese workday. For users in Europe, nearly the entire afternoon falls under the lower rate. Ad During off-peak hours, V4-Pro input goes from $0.435 to $0.66 per million tokens, and output jumps from $0.87 to $1.98. During peak hours, those rates double to $1.32 and $3.96. Cache hits are seeing the steepest increase, going from $0.003625 to $0.022 off-peak and $0.044 at peak. That shrinks the cache discount from about one-hundred-twentieth to one-thirtieth of the regular input price. For agents that repeatedly read the same files, this is the most expensive part of the change. The new pricing partially undoes the price cut Deepseek rolled out in May , and cache hits will actually cost more than they did before that reduction. The price hike comes as the company is raising new capital and preparing for an initial public offering. AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: X/Deepseek | Harness