Publications
(* denotes equal contribution)
-
Learning-free L2-Accented Speech Generation using Phonological Rules
Thanathai Lertpetchpun, Yoonjeong Lee, Jihwan Lee, Tiantian Feng, Dani Byrd, Shrikanth Narayanan
arXiv preprint
-
Speech Generation Speaker Poisoning: Capability Erasure in Zero-Shot Text-to-Speech
Thanathai Lertpetchpun*, Thanapat Trachu*, Sai Praneeth Karimireddy, Shrikanth Narayanan
Workshop on Responsibly Enabling Data for Foundation Models (ReData), COLM 2026
-
Accent Vector: Controllable Accent Manipulation for Multilingual TTS Without Accented Data
Thanathai Lertpetchpun*, Thanapat Trachu*, Jihwan Lee, Tiantian Feng, Dani Byrd, Shrikanth Narayanan
arXiv preprint
-
Trade-offs Between Capacity and Robustness in Neural Audio Codecs for Adversarially Robust Speech Recognition
Jordan Prescott, Thanathai Lertpetchpun, Shrikanth Narayanan
Submitted to SLT 2026
-
Towards Interpretable Framework for Neural Audio Codecs via Sparse Autoencoders: A Case Study on Accent Information
Shih-Heng Wang, Tiantian Feng, Aditya Kommineni, Thanathai Lertpetchpun, Bowen Yi, Xuan Shi, Shrikanth Narayanan
Interspeech, 2026
-
Quantifying Speaker Embedding Phonological Rule Interactions in Accented Speech Synthesis
Thanathai Lertpetchpun*, Yoonjeong Lee*, Thanapat Trachu, Jihwan Lee, Tiantian Feng, Dani Byrd, Shrikanth Narayanan
ICASSP, 2026
-
ARTI-6: Towards Six-dimensional Articulatory Speech Encoding
Jihwan Lee, Sean Foley, Thanathai Lertpetchpun, Kevin Huang, Yoonjeong Lee, Tiantian Feng, Louis Goldstein, Dani Byrd, Shrikanth Narayanan
ICASSP, 2026
-
VoxGuard: Evaluating User and Attribute Privacy in Speech via Membership Inference Attacks
Efthymios Tsaprazlis, Thanathai Lertpetchpun, Tiantian Feng, Sai Praneeth Karimireddy, Shrikanth Narayanan
ICASSP, 2026
-
Voxlect: A Speech Foundation Model Benchmark for Modeling Dialects and Regional Languages Around the Globe
Tiantian Feng, Kevin Huang, Anfeng Xu, Xuan Shi, Thanathai Lertpetchpun, Jihwan Lee, Yoonjeong Lee, Dani Byrd, Shrikanth Narayanan
KDD 2026
-
Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
Tiantian Feng, Jihwan Lee, Anfeng Xu, Yoonjeong Lee, Thanathai Lertpetchpun, Xuan Shi, Helin Wang, Thomas Thebaud, Laureano Moro-Velazquez, Dani Byrd, Najim Dehak, Shrikanth Narayanan
Submitted to DMLR
-
Developing a High-performance Framework for Speech Emotion Recognition in Naturalistic Conditions Challenge for Emotional Attribute Prediction
Thanathai Lertpetchpun*, Tiantian Feng*, Dani Byrd, Shrikanth Narayanan
Interspeech, 2025
-
Developing a Top-tier Framework in Naturalistic Conditions Challenge for Categorized Emotion Prediction: From Speech Foundation Models and Learning Objective to Data Augmentation and Engineering Choices
Tiantian Feng*, Thanathai Lertpetchpun*, Dani Byrd, Shrikanth Narayanan
Interspeech, 2025
-
Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing Developing
Thanapat Trachu*, Thanathai Lertpetchpun*, Ekapol Chuangsuwanich
Interspeech, 2025
-
Instance-based Temporal Normalization for Speaker Verification
Thanathai Lertpetchpun, Ekapol Chuangsuwanich
Interspeech, 2023