Related Experiment Video
Updated: Apr 2, 2026

Asthma Detection Research Based on Voice Signal Processing and Machine Learning
Published on: July 22, 2025
A Multichannel Flexible Interface for Environmental-Robust Laryngeal Signal Decoding
Yusen Guo1, Pengyu Huo1, Sisi Huang1
1School of Advanced Manufacturing and Robotics, Peking University, Beijing 100871, China.
Abstract:
Achieving robust human-machine interaction in noisy, constrained, or speech-impaired environments remains a significant challenge for conventional voice-based systems. Here, we present a wearable, flexible, and multichannel piezoresistive interface capable of decoding laryngeal and submandibular motion during complex speech behaviors. The system integrates a micropyramid polydimethylsiloxane (PDMS) sensing layer coated with conductive polypyrrole (PPy) onto a multichannel electrode array supported by a flexible polyimide (PI) substrate, providing superior skin conformity, high strain sensitivity, and robust long-term stability. We developed a fully integrated hardware platform enabling four-channel synchronous data acquisition, wireless transmission, and real-time on-device processing. A modified Audio Spectrogram Transformer (AST) combined with a multichannel fusion mechanism enables end-to-end semantic recognition. Using a 14-word core English vocabulary, we constructed two structured datasets─Microphone and Vocal─comprising a total of 3,840 samples. The system achieved classification accuracies of 99.6% and 96.4%, respectively, highlighting strong generalizability, semantic clarity, and robustness against signal variability. Real-world evaluations confirm stable performance under motion, facial expressions, and background noise. By unifying soft materials engineering, flexible circuit integration, and multimodal deep learning, this work advances speech recognition in complex environments and offers a scalable solution for assistive communication, wearable AI, and silent interaction under extreme conditions.
Related Concept Videos
Design Example
Classification of Signals
A continuous-time signal holds a value at every instant in time, representing information seamlessly. In contrast, a discrete-time signal holds values only at specific moments, often denoted as x(n), where...
Signal Sequences and Sorting Receptors
Larynx
Anatomy of the Larynx
The larynx consists of various components, including cartilage, muscles, and vocal cords. Its structure includes three large unpaired cartilages—the thyroid, cricoid, and epiglottis—and three smaller paired cartilages—the arytenoids,...
Reconstruction of Signal using Interpolation
Multi-input and Multi-variable systems
In the absence of...

