Related Experiment Video
Updated: May 10, 2026

Manufacturing Process for Non-Adhesive Super-Soft Vocal Fold Models
Published on: January 5, 2024
Improved vocal tract reconstruction and modeling using an image super-resolution technique
Xinhui Zhou1, Jonghye Woo, Maureen Stone
1Speech Communication Laboratory, Institute of Systems Research and Department of Electrical and Computer Engineering, University of Maryland, College Park, Maryland 20742, USA. zxinhui@umd.edu
Abstract:
Magnetic resonance imaging has been widely used in speech production research. Often only one image stack (sagittal, axial, or coronal) is used for vocal tract modeling. As a result, complementary information from other available stacks is not utilized. To overcome this, a recently developed super-resolution technique was applied to integrate three orthogonal low-resolution stacks into one isotropic volume. The results on vowels show that the super-resolution volume produces better vocal tract visualization than any of the low-resolution stacks. Its derived area functions generally produce formant predictions closer to the ground truth, particularly for those formants sensitive to area perturbations at constrictions.

