diarization


정의: Speaker diarisation is the process of partitioning an audio stream containing human speech into homogeneous segments according to the identity of each speaker. It can enhance the readability of an automatic speech transcription by structuring the audio stream into speaker turns and, when used together with speaker recognition systems, by providing the speaker’s true identity. It is used to answer the question "who spoke when?" Speaker diarisation is a combination of speaker segmentation and speaker clustering. The first aims at finding speaker change points in an audio stream. The second aims at grouping together speech segments on the basis of speaker characteristics.


📄 키워드 상세정보

핵심 연구 분야Artificial Intelligence
주요 연도2024년
주요 연관 키워드neural
좋아요 수0

키워드별 논문 목록 (1건)