Voice privacy using CycleGAN and time-scale modification. (July 2022)

Record Type:: Journal Article
Title:: Voice privacy using CycleGAN and time-scale modification. (July 2022)
Main Title:: Voice privacy using CycleGAN and time-scale modification
Authors:: Prajapati, Gauri P.
Singh, Dipesh K.
Amin, Preet P.
Patil, Hemant A.
Abstract:: Abstract: Extensive use of Intelligent Personal Assistants (IPA) and biometrics in our day-to-day life asks for privacy preservation while dealing with personal data. To that effect, efforts have been made to preserve the personally identifiable characteristics from human voice using different speaker anonymization techniques. In this paper, we propose Cycle Consistent Generative Adversarial Network (CycleGAN) to modify (transform) the speaker's gender as well as the other prosodic aspects using their Mel cepstral coefficients (MCEPs) and fundamental frequency (i.e., F 0 ). For effective anonymization in the context of voice privacy, we propose two-level (i.e., double) anonymization, where first-level anonymization is done using CycleGAN, followed by second-level anonymization using time-scale modification. The speaker anonymization and intelligibility are measured objectively using the automatic speaker verification (ASV) and automatic speech recognition (ASR) experiments, respectively, on development and test sets of Librispeech and VCTK datasets. For CycleGAN-based anonymization, the average % EERs (% WERs) are 40.3% (8.89%) and 40.95% (9.37%) with original enrollments and anonymized trials of the development and test datasets, respectively. The average % EERs (% WERs) for double anonymization are 46.19% (9.95%) and 44.76% (10.34%) with original enrollments and anonymized trials of the development and test datasets, respectively. For the voice privacy evaluation, the … (more)
Is Part Of:: Computer speech & language. Volume 74(2022)
Journal:: Computer speech & language
Issue:: Volume 74(2022)
Issue Display:: Volume 74, Issue 2022 (2022)
Year:: 2022
Volume:: 74
Issue:: 2022
Issue Sort Value:: 2022-0074-2022-0000
Page Start:
Page End:
Publication Date:: 2022-07
Subjects:: Voice privacy -- Double anonymization -- Time-scale modification -- Speech perturbation -- Cycle Consistent Generative Adversarial Network (cycleGAN)
Speech processing systems -- Periodicals
Automatic speech recognition -- Periodicals
Computers -- Periodicals
Linguistics -- Periodicals
Speech-Language Pathology -- Periodicals
Traitement automatique de la parole -- Périodiques
Reconnaissance automatique de la parole -- Périodiques
Automatic speech recognition
Speech processing systems
Electronic journals
Periodicals
006.454
Journal URLs:: http://www.journals.elsevier.com/computer-speech-and-language/ ↗
http://www.elsevier.com/journals ↗
DOI:: 10.1016/j.csl.2022.101353 ↗
Languages:: English
ISSNs:: 0885-2308
Deposit Type:: Legaldeposit
View Content:: Available online (eLD content is only available in our Reading Rooms) ↗
Physical Locations:: British Library DSC - 3394.276600
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store
Ingest File:: 21011.xml