Member of Technical Staff, ML Engineer
Core
Building the inference APIs, batch/compute systems, and services that make the Omnii genome language model fast, reliable, and cheap at genome scale for scientists and partners.
Role type
Senior IC ML Engineer (Inference & Distributed Systems)
Builds
Inference APIs, SDKs, documentation, and compute/orchestration layers for the Omnii model.
Domain
Generative AI / Genomics / Distributed Systems
Deliverable
production ML models
Required skills
Python, typed languages, containers, infrastructure as code, API design, distributed systems, batch job systems, observability, cost/latency optimization
Preferred skills
GPU inference internals (vLLM, TensorRT-LLM, Triton), enterprise VPC/on-prem/air-gapped deployment, genomics/computational biology
Technologies
Python, vLLM, TensorRT-LLM, Triton
Responsibilities
Build real-time and batch inference APIs for genome-scale workloads; design compute orchestration with autoscaling and fault tolerance; develop SDKs and documentation for external teams; optimize inference cost and latency; package platform for on-prem and air-gapped environments; collaborate with scientists to shape API workloads.
Seniority
Senior, hands-on IC