CareerPlanSign in

Staff + Sr. Software Engineer, Scaling

New York City, NY; San Francisco, CA; Seattle, WA💼 Full-time💰 $320,000–$320,000🗓 2026-09-29

Core

Design, build, and maintain distributed systems serving Claude to millions of users, focusing on intelligent request routing, fleet-wide orchestration, and compute efficiency across diverse AI accelerators and cloud platforms.

Role type

Staff/Senior Software Engineer (Distributed Systems & Inference Infrastructure)

Builds

High-performance inference infrastructure, intelligent routing systems, and production-grade deployment pipelines for LLMs.

Domain

Artificial Intelligence / Large Language Model (LLM) Inference / Distributed Systems

Deliverable

production ML models

Required skills

Distributed systems design, load balancing, traffic management, autoscaling, cloud infrastructure, Python, Rust

Preferred skills

LLM inference optimization, Kubernetes, multi-cloud orchestration, high-performance computing

Responsibilities

Design intelligent routing algorithms, manage multi-region deployments, analyze observability data to tune performance, integrate new AI accelerator platforms, build deployment pipelines for model releases.

Seniority

Staff/Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 847,000+ jobs from 20+ sources.