Statistical methods for estimation of direct and differential kinematics of the vocal tract

详细信息查看全文

作者：Adam Lammert^a ; ^{lammert@usc.edu} ; Louis Goldstein^b ; ^c ; Shrikanth Narayanan^a ; ^b ; Khalil Iskarous^b ; ^c
关键词：Speech production ; Direct kinematics ; Differential kinematics ; Task dynamics ; Articulatory synthesis ; Kinematic estimation ; Statistical machine learning ; Locally-weighted regression ; Artificial neural networks
刊名：Speech Communication
出版年：2013
出版时间：January, 2013
年：2013
卷：55
期：1
页码：147-161
全文大小：510 K

文摘

We present and evaluate two statistical methods for estimating kinematic relationships of the speech production system: artificial neural networks and locally-weighted regression. The work is motivated by the need to characterize this motor system, with particular focus on estimating differential aspects of kinematics. Kinematic analysis will facilitate progress in a variety of areas, including the nature of speech production goals, articulatory redundancy and, relatedly, acoustic-to-articulatory inversion. Statistical methods must be used to estimate these relationships from data since they are infeasible to express in closed form. Statistical models are optimized and evaluated - using a heldout data validation procedure - on two sets of synthetic speech data. The theoretical and practical advantages of both methods are also discussed. It is shown that both direct and differential kinematics can be estimated with high accuracy, even for complex, nonlinear relationships. Locally-weighted regression displays the best overall performance, which may be due to practical advantages in its training procedure. Moreover, accurate estimation can be achieved using only a modest amount of training data, as judged by convergence of performance. The algorithms are also applied to real-time MRI data, and the results are generally consistent with those obtained from synthetic data.

地址：北京市海淀区学院路29号邮编：100083

电话：办公室：(+86 10)66554848；文献借阅、咨询服务、科技查新：66554700