Hume AIがTADAをオープンソース化、競合より5倍高速で幻覚ゼロの音声モデル
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
The Decoder
Hume AIはMITライセンスでTADAを公開した。この高速音声生成モデルはテキストと音声を同期処理し、テストで幻覚を一切発生させなかった。
Source Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。

Hume AIは、MITライセンスの下で高速音声生成モデル「TADA」をオープンソース化しました。このモデルはテキストと音声を同期処理し、テストにおいて幻覚を一切発生させませんでした。
この記事「Hume AI open-sources TADA, a speech model five times faster than rivals with zero hallucinated words」は、The Decoderで最初に公開されました。
原文を表示
Hume AI has open-sourced TADA, an AI system for speech generation that processes text and audio in sync. Unlike previous systems that generate significantly more audio frames per text token, TADA maps exactly one audio signal to each text token. The result, according to Hume AI: TADA is over five times faster than comparable systems and produced zero transcription hallucinations—no made-up or skipped words compared to the source text—across tests with more than 1,000 samples. In human evaluations, the system scored 3.78 out of 5 for naturalness.
Hume AI says TADA is compact enough to run on smartphones, though longer texts can cause the voice to occasionally drift. The system comes in two sizes—1B and 3B parameters—both based on Llama. The smaller model supports English, while the 3B version covers seven additional languages. All code and models are available on GitHub and Hugging Face under the MIT license, and the full technical details can be found in the paper.
AI News Without the Hype – Curated by Humans
Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive "AI Radar" frontier report six times a year, full archive access, and access to our comment section.
Subscribe now
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み