ePrints@IIScePrints@IISc Home | About | Browse | Latest Additions | Advanced Search | Contact | Help

The impact of speaking rate on acoustic-to-articulatory inversion

Illa, Aravind and Ghosh, Prasanta Kumar (2020) The impact of speaking rate on acoustic-to-articulatory inversion. In: COMPUTER SPEECH AND LANGUAGE, 59 . pp. 75-90.

[img] PDF
com_spe_lan_59_75_2020.pdf - Published Version
Restricted to Registered users only

Download (1MB) | Request a copy
Official URL: http://dx.doi.org/10.1016/j.csl.2019.05.004

Abstract

Acoustic characteristics and articulatory movements are known to vary with speaking rates. This study investigates the role of speaking rate on acoustic-to-articulatory inversion (AAI) performance using deep neural networks (DNNs). Since fast speaking rate causes fast articulatory motion as well as changes in spectro-temporal characteristics of the speech signal, the articulatory-acoustic map in a fast speaking rate could be different from that in a slow speaking rate. We examine how these differences alter the accuracy with which different articulatory positions could be recovered from the acoustics. AAI experiments are performed in both matched and mismatched train-test conditions using data of five subjects, in three different rates - normal, fast and slow (fast and slow rates are at least 1.3 times faster and slower than the normal rate). Experiments in matched cases reveal that, the errors in estimating vertical motion of sensors on the tongue articulators from acoustics with fast speaking rate, is significantly higher than those with slow speaking rate. Experiments in mis-matched conditions reveal that there is consistent drop in AAI performance compared to the matched condition. Further experiments performed by training AAI with acoustic-articulatory data pooled from different speaking rates reveal that a single DNN based AAI model is capable of learning multiple rate-specific mapping.

Item Type: Journal Article
Publication: COMPUTER SPEECH AND LANGUAGE
Publisher: ACADEMIC PRESS LTD- ELSEVIER SCIENCE LTD
Additional Information: Copyright of this article belongs to ACADEMIC PRESS LTD- ELSEVIER SCIENCE LTD
Keywords: Acoustic-to-articulatory inversion; Speaking rate; Electromagnetic articulograph
Department/Centre: Division of Electrical Sciences > Electrical Engineering
Date Deposited: 21 Nov 2019 07:36
Last Modified: 21 Nov 2019 07:36
URI: http://eprints.iisc.ac.in/id/eprint/63796

Actions (login required)

View Item View Item