Convolutive ICA-Based Forensic Speaker Identification Using Mel Frequency Cepstral Coefficients and Gaussian Mixture Models
2013; Volume: 8; Issue: 1 Linguagem: Inglês
10.5769/j201301004
ISSN1980-7333
AutoresMatheus Silveira, Cezar Schroeder, João Paulo C. L. da Costa, Celso José Bruno de Oliveira, José A. Apolinário, Antonio Serrano, Paulo Quintiliano, Rafael Sousa Júnior,
Tópico(s)Speech Recognition and Synthesis
ResumoAutomatic speaker identification techniques are widely used nowadays in forensic applications, but its accuracy harshly drops when the voice of the speaker of interest is immersed in a recording containing more than one voice, common situation of investigations where the targets voice are obtained through ambient recordings. In forensic applications where microphones are hidden, such interferent sound sources in recordings are common and they degrade severely the performance of speaker identification techniques. In this paper, we propose a method to mitigate this problem by spatially separating the voice of each speaker using a Blind Source Separation technique called Convolutive Independent Component Analysis, and then applying the separated speech signals to a speaker identification system based on Mel Frequency Cepstral Coefficients and Gaussian Mixture Models. For identifying more than one speaker, the proposed system has a better accuracy than the state-of-the-art solutions.
Referência(s)