ML Research Engineer — Speech-to-Text / Text-to-Speech model training
Performed end-to-end training for Speech-to-Text and Text-to-Speech systems tailored to Malaysian and Arabic dialects, improving robustness in real-world conditions. Worked with transformer architectures, self-supervised learning, and neural vocoders to build and refine ASR/TTS pipelines. Enhanced multilingual performance using data augmentation, noise-robust training, and multilingual transfer learning strategies. • Train and refine STT/TTS deep learning models • Apply transformer-based and self-supervised learning approaches • Use neural vocoders for speech generation • Optimize accuracy and latency for production use