A silent speech system based on permanent magnet articulography and direct synthesis. (September 2016)
- Record Type:
- Journal Article
- Title:
- A silent speech system based on permanent magnet articulography and direct synthesis. (September 2016)
- Main Title:
- A silent speech system based on permanent magnet articulography and direct synthesis
- Authors:
- Gonzalez, Jose A.
Cheah, Lam A.
Gilbert, James M.
Bai, Jie
Ell, Stephen R.
Green, Phil D.
Moore, Roger K. - Abstract:
- Abstract : Highlights: This paper introduces a 'Silent Speech Interface' with the potential to restore the power of speech to people who have completely lost their voices. Small, unobtrusive magnets are attached to the lips and tongues and changes in magnetic field are sensed as the 'speaker' mouths what s/he wants to say. The sensor data is transformed to acoustic data by a speaker-dependent, learned transformation over parallel acoustic and sensor data. The machine learning technique used here is Mixture of Factor Analysis. Results are presented for 3 speakers which demonstrate that the SSI is capable of producing 'speech' which is both intelligible and natural. Abstract: In this paper we present a silent speech interface (SSI) system aimed at restoring speech communication for individuals who have lost their voice due to laryngectomy or diseases affecting the vocal folds. In the proposed system, articulatory data captured from the lips and tongue using permanent magnet articulography (PMA) are converted into audible speech using a speaker-dependent transformation learned from simultaneous recordings of PMA and audio signals acquired before laryngectomy. The transformation is represented using a mixture of factor analysers, which is a generative model that allows us to efficiently model non-linear behaviour and perform dimensionality reduction at the same time. The learned transformation is then deployed during normal usage of the SSI to restore the acoustic speech signalAbstract : Highlights: This paper introduces a 'Silent Speech Interface' with the potential to restore the power of speech to people who have completely lost their voices. Small, unobtrusive magnets are attached to the lips and tongues and changes in magnetic field are sensed as the 'speaker' mouths what s/he wants to say. The sensor data is transformed to acoustic data by a speaker-dependent, learned transformation over parallel acoustic and sensor data. The machine learning technique used here is Mixture of Factor Analysis. Results are presented for 3 speakers which demonstrate that the SSI is capable of producing 'speech' which is both intelligible and natural. Abstract: In this paper we present a silent speech interface (SSI) system aimed at restoring speech communication for individuals who have lost their voice due to laryngectomy or diseases affecting the vocal folds. In the proposed system, articulatory data captured from the lips and tongue using permanent magnet articulography (PMA) are converted into audible speech using a speaker-dependent transformation learned from simultaneous recordings of PMA and audio signals acquired before laryngectomy. The transformation is represented using a mixture of factor analysers, which is a generative model that allows us to efficiently model non-linear behaviour and perform dimensionality reduction at the same time. The learned transformation is then deployed during normal usage of the SSI to restore the acoustic speech signal associated with the captured PMA data. The proposed system is evaluated using objective quality measures and listening tests on two databases containing PMA and audio recordings for normal speakers. Results show that it is possible to reconstruct speech from articulator movements captured by an unobtrusive technique without an intermediate recognition step. The SSI is capable of producing speech of sufficient intelligibility and naturalness that the speaker is clearly identifiable, but problems remain in scaling up the process to function consistently for phonetically rich vocabularies. … (more)
- Is Part Of:
- Computer speech & language. Volume 39(2016)
- Journal:
- Computer speech & language
- Issue:
- Volume 39(2016)
- Issue Display:
- Volume 39, Issue 2016 (2016)
- Year:
- 2016
- Volume:
- 39
- Issue:
- 2016
- Issue Sort Value:
- 2016-0039-2016-0000
- Page Start:
- 67
- Page End:
- 87
- Publication Date:
- 2016-09
- Subjects:
- Silent speech interfaces -- Speech rehabilitation -- Speech synthesis -- Permanent magnet articulography -- Augmentative and alternative communication
Speech processing systems -- Periodicals
Automatic speech recognition -- Periodicals
Computers -- Periodicals
Linguistics -- Periodicals
Speech-Language Pathology -- Periodicals
Traitement automatique de la parole -- Périodiques
Reconnaissance automatique de la parole -- Périodiques
Automatic speech recognition
Speech processing systems
Electronic journals
Periodicals
006.454 - Journal URLs:
- http://www.journals.elsevier.com/computer-speech-and-language/ ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.csl.2016.02.002 ↗
- Languages:
- English
- ISSNs:
- 0885-2308
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 3394.276600
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 2467.xml