Title
A comparison study on patient-psychologist voice diarization.
Abstract
Conversations between a clinician and a patient, in natural conditions, are valuable sources of information for medical follow-up. The automatic analysis of these dialogues could help extract new language markers and speed up the clinicians’ reports. Yet, it is not clear which model is the most efficient to detect and identify the speaker turns, especially for individuals with speech disorders. Here, we proposed a split of the data that allows conducting a comparative evaluation of different diarization methods. We designed and trained end-to-end neural network architectures to directly tackle this task from the raw signal and evaluate each approach under the same metric. We also studied the effect of fine-tuning models to find the best performance. Experimental results are reported on naturalistic clinical conversations between Psychologists and Interviewees, at different stages of Huntington’s disease, displaying a large panel of speech disorders. We found out that our best end-to-end model achieved 19.5 % IER on the test set, compared to 23.6% achieved by the finetuning of the X-vector architecture. Finally, we observed that we could extract clinical markers directly from the automatic systems, highlighting the clinical relevance of our methods.
Year
DOI
Venue
2022
10.18653/v1/2022.slpat-1.4
Workshop on Speech and Language Processing for Assistive Technologies (SLPAT)
DocType
Volume
Citations 
Conference
Ninth Workshop on Speech and Language Processing for Assistive Technologies (SLPAT-2022)
0
PageRank 
References 
Authors
0.34
0
9
Name
Order
Citations
PageRank
Rachid Riad100.68
Hadrien Titeux201.01
Laurie Lemoine300.34
Justine Montillot400.34
Agnes Sliwinski500.34
Jennifer Bagnou600.34
Xuan Cao701.01
A-C Bachoud-Levi8101.76
Emmanuel Dupoux923837.33