Principal Machine Learning Engineer
Core
Senior technical anchor and solution architect for Grab's core ML and AI infrastructure platform, enabling hundreds of data scientists and ML engineers to build production models.
Role type
Principal Machine Learning Engineer (Platform & Infrastructure)
Builds
Core ML/AI infrastructure including model serving, ML pipelines, data serving, and training systems for Grab's superapp ecosystem.
Domain
Superapp / Ride-hailing / Food delivery / Fintech / AI Infrastructure
Deliverable
production ML models
Required skills
MLOps & ML Platform Engineering, Distributed Systems & Infrastructure, Architecture Design, AI/LLM System Experience, Innovation & AI Fluency, Adaptive Execution & Ownership, Coaching with Care
Preferred skills
None explicitly stated
Technologies
Kubeflow, MLflow, Triton, TorchServe, PyTorch, Ray, Horovod, Kubernetes, vLLM, TensorRT-LLM
Responsibilities
Design end-to-end solutions for AIP users and serve as senior technical escalation point; Drive state of large-scale training (throughput, reliability, cost); Optimize end-to-end model iteration loop from idea to shipped model; Design and drive integrations across AIP surfaces; Translate user pain points into functional requirements for AIP teams; Produce reference architectures and best-practice guidance; Define strategic roadmap and mentor senior engineers.
Seniority
Principal, hands-on IC with strategy & mentorship


