Machine Learning Engineer (Speech/Audio) - Singapore
Core
Building large-scale speech/audio data pipelines and optimizing speech/language models for professional AI tools.
Role type
Senior IC machine learning engineer (speech/audio)
Builds
Data pipelines for model training and fine-tuned speech models for professional productivity tools
Domain
AI / Speech Recognition / Audio Processing
Deliverable
production ML models
Required skills
Speech/audio data pipeline engineering, ASR/SpeechLLM model training and fine-tuning, LLM training and fine-tuning, distributed data processing (Spark, Ray), Python, PyTorch
Preferred skills
Speech SSL concepts, hotword/contextual-biasing optimization, code-switching ASR improvements, publications in top AI venues
Technologies
PyTorch, Spark, Ray, SpeechLLM, StepAudio, Qwen3-Omni
Responsibilities
Own large-scale speech/audio data pipelines for collection, cleaning, filtering, labeling, and augmentation; Support model training and optimization for recognition accuracy in code-switching and domain-specific terms; Perform domain adaptation using scenario-specific data; Build evaluation frameworks and benchmark models against baselines
Seniority
Senior, hands-on IC


