音響的近傍埋め込みの理論的枠組み
本文の状態
日本語全文を表示中
詳細モードで約1分の本文を読めます。
同じ出来事の情報源
この情報源を基点に整理
Apple Machine Learning
研究者らが、可変幅の音声やテキストの音韻内容を固定次元の埋め込み空間で表現する「音響的近傍埋め込み」の理論的枠組みを提案した。単語間の音韻的類似性の定量的定義に基づき、埋め込み間の距離の確率的解釈を提供し、原理に基づいた理解と応用を可能にする。
Continue in AI NEW LAB
このニュースを、実務の判断につなげる
AI NEW LABで、試したことや先に確認したい条件を共有できます。まずはログインなしで読めます。
AI NEW LABで論点を見るSource Article
元記事を日本語で読む
本文に関係しない購読案内、埋め込み通知、サイト内プロモーションは除いています。
本論文は、可変幅の音声またはテキストの音響的内容を固定次元の埋め込み空間で表現する「音響近傍埋め込み(acoustic neighbor embeddings)」を解釈するための理論的枠組みを提供する。単語間の音響的類似性に関する一般的な定量的定義に基づき、埋め込み間の距離の確率的解釈を提案する。これにより、埋め込みを理解し適用するための原理的な枠組みが得られる。一様なクラスごとの等方性(isotropy)の近似を支持する理論的および実証的な証拠を示す。これにより、…
原文を表示
This paper provides a theoretical framework for interpreting acoustic neighbor embeddings, which are representations of the phonetic content of variable-width audio or text in a fixed-dimensional embedding space. A probabilistic interpretation of the distances between embeddings is proposed, based on a general quantitative definition of phonetic similarity between words. This provides us a framework for understanding and applying the embeddings in a principled manner. Theoretical and empirical evidence to support an approximation of uniform cluster-wise isotropy are shown, which allows us to…
今日のまとめ
AIデイリーブリーフで今日の重要ニュースをまとめ読み