Sitemap
A list of all the posts and pages found on the site. For you robots out there, there is an XML version available for digesting as well.
Pages
Posts
Future Blog Post
Published:
This post will show up by default. To disable scheduling of future posts, edit config.yml and set future: false.
Blog Post number 4
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 3
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 2
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 1
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
portfolio
Portfolio item number 1
Short description of portfolio item number 1
Portfolio item number 2
Short description of portfolio item number 2 
publications
Re-Sonance: A Dysarthric Asynchronous Real-Time Speech Conversion System Based on a Three-Stage Cascaded ASR-LLM-TTS Architecture
Published in NCMMSC 2025, 2025
Individuals with dysarthria face major difficulties in professional speaking scenarios requiring real-time communication. Existing AAC systems often suffer from high latency and unnatural speech. We introduce Re-Sonance, an LLM-enhanced, speech-driven AAC system integrating Whisper ASR, Qwen LLM, and CosyVoice TTS for real-time use. Evaluations on Mandarin dysarthric speech show improved intelligibility and naturalness while preserving semantics for mild to moderate dysarthria, highlighting the promise of LLM-based AAC systems.
Recommended citation: Wu, Y., Xu, Y., Wang, J., Zhao, X., Jiang, J., & Luo, Z. (2025). Re-Sonance: A Dysarthric Asynchronous Real-Time Speech Conversion System Based on a Three-Stage Cascaded ASR-LLM-TTS Architecture. Proceedings of the 2025 National Conference on Man–Machine Speech Communication (NCMMSC).
Download Paper | Download Slides
PhoenixDSR: Phoneme-Guided and LLM-Enhanced Dysarthric Speech Recognition
Published in ICASSP 2026, 2026
Automatic speech recognition remains weak on dysarthric speech because of limited data and speaker variability. We propose PhoenixDSR, a phoneme-mediated framework that separates acoustic variability from linguistic decoding. A Wav2Vec2-CTC model trained on healthy speech yields stable phonemes, while a weighted confusion matrix captures global and speaker-specific dysarthric patterns. A lightweight LLM decoder performs multi-task phoneme–text repair. PhoenixDSR achieves strong, data-efficient, and robust results on CDSD.
Recommended citation: Wu, Y., Xu, Y., Wang, J., Zhao, X., Jiang, J., & Luo, Z. (2026). PHOENIXDSR: Phoneme-Guided and LLM-Enhanced Dysarthric Speech Recognition. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).
Download Paper
talks
Talk 1 on Relevant Topic in Your Field
Published:
This is a description of your talk, which is a markdown file that can be all markdown-ified like any other post. Yay markdown!
Conference Proceeding talk 3 on Relevant Topic in Your Field
Published:
This is a description of your conference proceedings talk, note the different field in type. You can put anything in this field.
teaching
Teaching experience 1
Undergraduate course, University 1, Department, 2014
This is a description of a teaching experience. You can use markdown like any other post.
Teaching experience 2
Workshop, University 1, Department, 2015
This is a description of a teaching experience. You can use markdown like any other post.
