Scheduled Maintenance Notice
Please note that Researcher Profiles will be undergoing scheduled maintenance on Wednesday 7th Oct, from 8:00am to 9:00am. During this time, the Researcher Profiles system will be unavailable. We apologise for any inconvenience and appreciate your understanding.
Select Publications
Preprints
, 2025, System X: A Mobile Voice-Based AI System for EMR Generation and Clinical Decision Support in Low-Resource Maternal Healthcare, http://dx.doi.org/10.48550/arxiv.2512.12240
, 2025, Auto-Landmark: Acoustic Landmark Dataset and Open-Source Toolkit for Landmark Extraction, http://dx.doi.org/10.48550/arxiv.2409.07969
, 2025, Mamba in Speech: Towards an Alternative to Self-Attention, http://dx.doi.org/10.48550/arxiv.2405.12609
, 2024, A Joint Spectro-Temporal Relational Thinking Based Acoustic Modeling Framework, http://dx.doi.org/10.48550/arxiv.2409.15357
, 2023, Variational Connectionist Temporal Classification for Order-Preserving Sequence Modeling, http://dx.doi.org/10.48550/arxiv.2309.11983
, 2023, Phonological Level wav2vec2-based Mispronunciation Detection and Diagnosis Method, http://dx.doi.org/10.48550/arxiv.2311.07037
, 2023, Spatial HuBERT: Self-supervised Spatial Speech Representation Learning for a Single Talker from Multi-channel Audio, http://dx.doi.org/10.48550/arxiv.2310.10922
, 2022, Improving Children's Speech Recognition by Fine-tuning Self-supervised Adult Speech Representations, http://dx.doi.org/10.48550/arxiv.2211.07769
, 2022, Speaker- and Age-Invariant Training for Child Acoustic Modeling Using Adversarial Multi-Task Learning, http://dx.doi.org/10.48550/arxiv.2210.10231