Unsupervised Speaker Diarization that is Agnostic to Language, Overlap-Aware, and Tuning Free

Conference of the International Speech Communication Association (INTERSPEECH)(2022)

引用 0|浏览35
暂无评分
摘要
Podcasts are conversational in nature and speaker changes are frequent -- requiring speaker diarization for content understanding. We propose an unsupervised technique for speaker diarization without relying on language-specific components. The algorithm is overlap-aware and does not require information about the number of speakers. Our approach shows 79% improvement on purity scores (34% on F-score) against the Google Cloud Platform solution on podcast data.
更多
查看译文
关键词
overlap-aware
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要