메뉴
BL
MarkTechPost • 25일 전

그래디움 AI, 신규 기본 TTS 모델 공개

IMP
5/10
핵심 요약

Gradium AI가 새 기본 TTS(텍스트 음성 변환) 모델을 발표했다. 5개 언어의 어려운 문장 500개에 대해 사람 평가 기준 81.0% 통과율을 기록했으며, Coval 환경에서 첫 오디오까지 216ms(P50)라는 빠른 응답 속도를 보였다. 평가 데이터셋은 Hugging Face에서 CC BY 4.0 라이선스로 공개되어 있다.

번역된 본문

텍스트 음성 변환(TTS) 분야에서는 속도와 정확도가 보통 상충합니다. Gradium AI의 새 기본 모델은 두 가지를 모두 달성했다고 보고합니다. 5개 언어의 어려운 문장 500개에 대해 사람 평가 기준 81.0%의 통과율을 기록했고, Coval에서 첫 오디오까지의 시간(Time-to-First-Audio)은 P50 기준 216ms입니다. 평가 세트는 Hugging Face에서 CC BY 4.0 라이선스로 공개되어 있습니다.

「Gradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio」라는 제목의 이 글은 MarkTechPost에 처음 게재되었습니다.

원문 보기
원문 보기 (영어)
Speed and accuracy usually pull against each other in text-to-speech. Gradium AI's new default model reports both: an 81.0% human-rated pass rate on 500 hard sentences across five languages, at 216 ms P50 time-to-first-audio on Coval. The evaluation set is open on Hugging Face under CC BY 4.0. The post Gradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio appeared first on MarkTechPost.