學系成員

李安德Ryandhimas Edo Zezario 助理教授 Assistant Professor

  • 辦公室位置:70805

  • (03) 4638800 Ext.

  • ryandhimas@saturn.yzu.edu.tw

  • 學生實驗室 :語音與聽覺智慧實驗室(Speech and Auditory Intelligence Lab) Phone: (03)4638800 Ext. 7011 / Ext. 939
  • 學歷

    國立臺灣大學資訊工程學系博士  PhD, CSIE, NTU


    經歷

    Postdoctoral Researcher, Academia Sinica

    Research Assistant, Academia Sinica

    Applied Scientist Intern, Amazon

    研究專長

    語音處理、語音評估、聽覺輔助技術、多模態系統

    Speech Processing, Speech Assessment, Hearing Assistive Technology, Multimodal Systems


  • 期刊論文

    S. Ahmed, R. E. Zezario, H.-G. Yuan, A. Hussain, H.-M. Wang, W.-H. Chung, and Y. Tsao, "NeuroAMP: A Novel End-to-end General Purpose Deep Neural Amplifier for Personalized Hearing Aids," IEEE Transactions on Artificial Intelligence, volume 7, pages 1610-1625, March 2026

    D. A. M. G. Wisnu, S. Rini, R. E. Zezario, H.-M. Wang, and Y. Tsao, "HAAQI-Net: A Non-intrusive Neural Music Audio Quality Assessment Model for Hearing Aids," IEEE Transactions on Audio, Speech and Language Processing, volume 33, pages 1877-1892, February 2025.

    R. E. Zezario, S. -W. Fu, F. Chen, C. -S. Fuh, H. -M. Wang and Y. Tsao, "Deep Learning-Based Non-Intrusive Multi-Objective Speech Assessment Model With Cross-Domain Features," IEEE/ACM Transactions on Audio, Speech, and Language Processing, volume 31, pages 54-70, September 2022.

     C. Yu*, R. E. Zezario*, S.-S. Wang, J. Sherman, Y.-Y. Hsieh, X. Lu, H.-M. Wang, and Y. Tsao, "Speech Enhancement based on Denoising Autoencoder with Multi-branched Encoders," IEEE/ACM Transactions on Audio, Speech, and Language Processing, volume 28, pages 2756-2769, October 2020, (*equal contributions)

  • 會議論文

    R. E. Zezario, D. A. Wisnu, S.-W. Fu, S. M. Siniscalchi, H.-M. Wang, and Y. Tsao, "Few-Shot and Pseudo-Label Guided Speech Quality Evaluation with Large Language Models," IEEE ICASSP 2026, pages 22482-22486, May 2026.

    G. Lin, X. Wang, R. E. Zezario, Y. Tsao, F. Chen, "Enhancing Speech Intelligibility Prediction for Hearing Aids with Complementary Speech Foundation Model Representations," IEEE ICASSP 2026, pages 15012-15016, May 2026.

    D. A. M. G. Wisnu, R. E. Zezario, S. Rini, H.-M. Wang, and Y.Tsao, "Improving Perceptual Audio Aesthetic Assessment via Triplet Loss and Self-Supervised Embeddings,"IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU), December 2025.

    W. Ren, Y.-C. Lin, W.-C. Huang, R. E. Zezario, S.-W. Fu, S.-F. Huang, E. Cooper, H. Wu, H.-Yu Wei, H.-Min Wang, H.-yi Lee, Y. Tsao, "HighRateMOS: Sampling-Rate Aware Modeling for Speech Quality Assessment," IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU), December 2025.

     R. E. Zezario, D. A.M.G. Wisnu, H.-M. Wang, and Y. Tsao, "Speech Intelligibility Assessment with Uncertainty-Aware Whisper Embeddings and sLSTM," 2025 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), October 2025.

    R. E. Zezario, "Non-Intrusive Intelligibility Prediction for Hearing Aids: Recent Advances, Trends, and Challenges," 2025 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), October 2025.

    S. Ahmed, R. E. Zezario, N. Saleem, A. Hussain, H.-M. Wang, and Y. Tsao, "A Study on Speech Assessment with Visual Cues," Interspeech 2025, pages 5418-5422, August 2025.

    R. E. Zezario, S. M. Siniscalchi, F. Chen, H.-M. Wang, and Y. Tsao, "Feature Importance across Domains for Improving Non-Intrusive Speech Intelligibility Prediction in Hearing Aids," Interspeech 2025, pages 5473-5477, August 2025.

    R. E. Zezario, D. A. M. G. Wisnu, H.-M. Wang, and Y. Tsao, "A Study on Zero-Shot Non-Intrusive Speech Intelligibility for Hearing Aids Using Large Language Models," ICCE-TW 2025, July 2025.

    R. E. Zezario, S. M. Siniscalchi, H.-M. Wang, and Y. Tsao, "A Study on Zero-shot Non-intrusive Speech Assessment using Large Language Models," 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), April 2025.

    W.-C. Huang, S.-W. Fu, E. Cooper, R. E. Zezario, T. Toda, H.-M. Wang, J. Yamagishi, and Y. Tsao, "The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction," IEEE Workshop on Spoken Language Technology (SLT), pages 803-810, December 2024.

    R. E. Zezario, F. Chen, C.-S.Fuh, H.-M. Wang, and Y. Tsao, "Non-Intrusive Speech Intelligibility Prediction for Hearing Aids using Whisper and Metadata," Interspeech 2024, pages 3844-3848, September 2024.

    R. E. Zezario, Y.-W. Chen, S.-W. Fu, Y. Tsao, H.-M. Wang, C.-S. Fuh, "A Study on Incorporating Whisper for Robust Speech Assessment," IEEE International Conference on Multimedia and Expo (ICME), July 2024, (Top Performance on the Track 3 - VoiceMOS Challenge 2023)

    R. E. Zezario, Bo-Ren Brian Bai, Chiou-Shann Fuh, Hsin-Min Wang, Yu Tsao, "Multi-Task Pseudo-Label Learning for Non-Intrusive Speech Quality Assessment Model," 2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 831-835, April 2024.

    R. E. Zezario, S.-W. Fu, F. Chen, C.-S. Fuh, H.-M. Wang and Y. Tsao, "MTI-Net: A Multi-Target Speech Intelligibility Prediction Model," Interspeech 2022, pages 5463-5467, September 2022.

    R. E. Zezario, F. Chen, C.-S. Fuh, H.-M. Wang and Y. Tsao, "MBI-Net: A Non-Intrusive Multi-Branched Speech Intelligibility Prediction Model for Hearing Aids," Interspeech 2022, pages 3944-3948, September 2022.

    R. E. Zezario, C. -S. Fuh, H. -M. Wang and Y. Tsao,, "Speech Enhancement with Zero-Shot Model Selection," 2021 29th European Signal Processing Conference (EUSIPCO, pages 491-495, December 2021.

    R. E. Zezario, S. -W. Fu, C. -S. Fuh, Y. Tsao and H. -M. Wang, "STOI-Net: A Deep Learning based Non-Intrusive Speech Intelligibility Assessment Model," 2020 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), number 482-486, December 2020.

    R. E. Zezario, T. Hussain, X. Lu, H. -M. Wang and Y. Tsao, "Self-Supervised Denoising Autoencoder with Linear Regression Decoder for Speech Enhancement," 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 6669-6673, May 2020.

    R. E. Zezario, J. W. C. Sigalingging, T. Hussain, J. -C. Wang and Y. Tsao, "Comparative Study of Masking and Mapping Based on Hierarchical Extreme Learning Machine for Speech Enhancement," 2019 International Symposium on Intelligent Signal Processing and Communication Systems (ISPACS), December 2019.

    R. E. Zezario, S.-W. Fu, X. Lu, H.-M. Wang, and Y. Tsao, "Specialized Speech Enhancement Model Selection Based on Learned Non-Intrusive Quality Assessment Metric," Interspeech 2019, pages 3168- 3172, September 2019.

    R. E. Zezario, J. Huang, X. Lu, Y. Tsao, H. Hwang and H. Wang,, "Deep Denoising Autoencoder Based Post Filtering for Speech Enhancement," 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), pages 373-377, November 2018.

     C. -Y. Hsu, R. E. Zezario, J. -C. Wang, C. -W. Ho, X. Lu and Y. Tsao, "Incorporating local environment information with ensemble neural networks to robust automatic speech recognition," 2016 10th International Symposium on Chinese Spoken Language Processing (ISCSLP), October 2016.

  • 得獎事蹟

    Third Place, Clarity Prediction Challenge — Hearing Industry Research Consortium, 2025

    Best Performing System of High-sampling-frequencies Track, AudioMOS Challenge (IEEE ASRU Grand Challenge), 2025

    Top 25 Most Downloaded Articles, IEEE/ACM TASLP, 2024

    Best Reviewer Award, IEEE ASRU, 2023

    Best Performing System of Noisy-enhanced Track, VoiceMOS Challenge (IEEE ASRU Grand Challenge), 2023

    Gold Prize & 1st Prize Student Award, Clarity Prediction Challenge — Hearing Industry Research Consortium, 2022

^