메뉴
BL
The Decoder 13일 전

구글, 'Gemma 4' 슬그머니 업데이트... 도구 호출 버그 및 응답 누락 수정

IMP
7/10
핵심 요약

구글이 오픈소스 AI 모델인 Gemma 4에 성능 개선 및 버그 수정 업데이트를 조용히 적용했습니다. 이번 업데이트로 엔비디아 GPU 환경에서의 처리 속도가 크게 향상되었으며, 외부 도구 호출 오류와 답변이 중간에 끊기는 문제가 해결되었습니다. 개발자 커뮤니티에서는 모델 버전을 새로 명명하지 않고 기존 'Gemma 4' 이름 그대로 업데이트를 배포한 구글의 방식에 대해 아쉬움을 나타내고 있습니다.

번역된 본문

구글(Google)이 오픈소스 AI 모델인 'Gemma 4'에 엔비디아 호퍼(Nvidia Hopper) GPU에서의 성능을 가속화하고, 도구 호출(Tool calling) 버그를 수정하며, 응답이 중간에 끊기는 문제를 해결하는 업데이트를 조용히 출시했다.

구글에 따르면, 플래시 어텐션 4(Flash Attention 4)를 활성화하면 모델이 프롬프트를 처리하는 속도가 25~70% 향상되며, 첫 번째 토큰이 생성되기까지 걸리는 시간(Time to first token)은 최대 31% 단축된다. 또한 구글은 모델이 스스로 외부 도구를 실행할 수 있게 해주는 기능인 도구 호출의 버그를 수정했다. 모델이 답변을 너무 짧게 자르거나 불완전한 응답을 반환하는 사례도 줄였다고 덧붙였다.

이미지 처리의 경우, 사용자가 'max_soft_tokens' 파라미터를 280에서 1,120으로 수동으로 늘려 더 선명한 OCR(광학 문자 인식) 결과를 얻고 최대 2.51메가픽셀 해상도를 지원받을 수 있다. 이를 위해 구글은 허깅페이스(Hugging Face)에 인터랙티브 구성기를 게시해 두었다.

공개된 벤치마크는 31B 및 E4B 변형 모델만 이전 버전과 비교했지만, 허깅페이스 저장소에 따르면 이 모델 세대의 모든 파라미터 크기가 최신 12B 버전을 포함하여 모두 업데이트된 것으로 나타났다.

그러나 개발자 커뮤니티는 구글이 이번 업데이트를 'Gemma 4.1'과 같은 별도의 버전으로 명명하지 않고 동일한 'Gemma 4'라는 이름으로 묵묵히 배포한 것에 대해 비판적인 반응을 보이고 있다.

원문 보기
원문 보기 (영어)
Gemma 4 gets a stealth update that fixes tool calling bugs and truncated responses under the same name Jonathan Kemper View the LinkedIn Profile of Jonathan Kemper Jul 16, 2026 Google shipped an update to its open AI model Gemma 4 that speeds up performance on Nvidia Hopper GPUs, fixes tool calling bugs, and addresses problems with truncated responses. Turning on Flash Attention 4 boosts the speed at which the model processes incoming prompts by 25 to 70 percent, according to Google. Time to first token drops by up to 31 percent. Google also fixed bugs in tool calling, the feature that lets the model trigger external tools on its own. Google says it also cut down on cases where the model would cut answers short or return incomplete responses. For image processing, users can manually raise the "max_soft_tokens" parameter from 280 to 1,120 to get sharper OCR results and support resolutions up to 2.51 megapixels. Google put up an interactive configurator on Hugging Face for that. The published benchmarks only compare the 31B and E4B variants against their predecessors, but the Hugging Face repository shows that all parameter sizes in this model generation got updated, including the newest 12B release . The community has pushed back on Google shipping the update under the same "Gemma 4" name instead of tagging it as a separate version like "Gemma 4.1." Ad DEC_D_Incontent-1 Ad AI News Without the Hype – Curated by Humans Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section. Subscribe now Source: X/Google Ask about this article… Search