Who Spoke When?: Audio-based Speaker Location Estimation for Diarization - Maral Dadvar - 書籍 - LAP LAMBERT Academic Publishing - 9783844386288 - 2011年7月1日
カバー画像とタイトルが一致しない場合、正しいのはタイトルです

Who Spoke When?: Audio-based Speaker Location Estimation for Diarization

価格
¥ 7.541
税抜

遠隔倉庫からの取り寄せ

発送予定日 年8月12日 - 年8月24日
Maral Dadvar の新しいリリースのお知らせを受け取る
iMusicのウィッシュリストに追加

まだ評価がありません

Speaker diarization is the process which detects active speakers and groups those speech signals which has been uttered by the same speaker. Generally we can find two main applications for speaker diarization. Automatic Speech Recognition systems make use of the speaker homogeneous clusters to adapt the acoustic models to be speaker dependent and therefore increase recognition performance. Speaker indexing and rich transcription systems also use the speaker diarization output as one of information extracted from a recording, which allow its automatic indexation and other further processing. In this study a speaker diarization application is developed ? using multiparty binaural speech recordings ? to track speaker activity based on interaural time difference (ITD) cues. These cues, for a given speech signal frame, are computed using gammatone filtering and cross-correlation technique. Their values are used to determine which speaker in the recording produce the considered speech fragment. This study has been supervised by Dr. Jon Barker, and defended to fulfill the requirements for the degree of Master in Advanced Computer Science, University of Sheffield, United Kingdom, 2007.

メディア 書籍     Paperback Book   (ソフトカバーで背表紙を接着した本)
リリース済み 2011年7月1日
ISBN13 9783844386288
出版社 LAP LAMBERT Academic Publishing
ページ数 68
寸法 150 × 4 × 226 mm   ·   119 g
言語 ドイツ語  

Maral Dadvarの他の作品を見る

すべて表示