2023
Synthesizing audio from tongue motion during speech using tagged MRI via transformer
Liu X, Xing F, Prince J, Stone M, Fakhri G, Woo J. Synthesizing audio from tongue motion during speech using tagged MRI via transformer. Proceedings Of SPIE--the International Society For Optical Engineering 2023, 12464: 1246410-1246410-5. PMID: 38009135, PMCID: PMC10669779, DOI: 10.1117/12.2653345.Peer-Reviewed Original ResearchMotion fieldAudio waveformAdversarial training approachImprove synthesis qualityConvolutional decoderAudio dataSynthesis qualityTranslation networkData structureSpeech waveformTemporal modelTagged MRITongue motionTraining approachSpectrogramMuscle deformationSource of informationSpeechIntelligible speechFrameworkDecodingInformationPredictive informationEncodingNetwork
2022
Tagged-MRI Sequence to Audio Synthesis via Self Residual Attention Guided Heterogeneous Translator
Liu X, Xing F, Prince J, Zhuo J, Stone M, El Fakhri G, Woo J. Tagged-MRI Sequence to Audio Synthesis via Self Residual Attention Guided Heterogeneous Translator. Lecture Notes In Computer Science 2022, 13436: 376-386. PMID: 36820764, PMCID: PMC9942274, DOI: 10.1007/978-3-031-16446-0_36.Peer-Reviewed Original ResearchAudio waveformEnd-to-end deep learning frameworkAdversarial training approachDeep learning frameworkEnd-to-endTwo-dimensional spectrogramAdversarial networkIntermediate representationLearning frameworkResidual attentionDisentanglement strategyAudio synthesisDataset sizeImprove realismHeterogeneous representationsHeterogeneous translationAttentional strategiesTraining approachExperimental resultsMuscle deformationIntelligible speechMotor control theoriesTagged-MRIRelated-disordersSpeech acoustics