Engineering Manager, Platform
Core
Lead the Platform Engineering and Observability team to build and operate the infrastructure beneath the Firmus AI Cloud, ensuring reliability and developer experience for internal engineering teams.
Role type
Senior Engineering Manager (Platform & Observability)
Builds
Bare-metal GPU compute, high-performance networking, internal platform services, self-service tooling, and observability platforms
Domain
AI Infrastructure / Cloud Computing / Data Center Operations
Deliverable
production ML models
Required skills
Team leadership, org design, end-to-end delivery management, technical direction, incident resolution, cross-functional collaboration
Preferred skills
AI-assisted development advocacy, build/buy/open-source decision making, SLA management
Technologies
GPU compute, high-performance networking, observability platforms
Responsibilities
Hire, coach, and manage engineers; own product roadmap delivery and release gates; set engineering quality standards and technical decisions; ensure production reliability and lead L3 incident resolution
Seniority
Senior, hands-on IC leader