CareerPlanSign in

ML Software Engineer

Seattle, United States of America💼 Full-time🗓 2026-07-30 → 2026-09-28

Core

Engineer the ML-inference stack for Apple Intelligence's Private Cloud Compute, optimizing performance and stability for generative AI on Apple Silicon datacenter hardware.

Role type

Senior IC ML Software Engineer (Inference Stack)

Builds

Private Cloud Compute inference stack for Apple Intelligence

Domain

Cloud infrastructure + Generative AI

Deliverable

production ML models

Required skills

Distributed systems, ML inference optimization, Systems programming (C++/Python/Go/Rust), Hardware/SoC integration, Observability, Concurrency management

Preferred skills

Server-side Swift, RPC/networking stacks (gRPC, Protobuf), Low-level systems programming, GPU acceleration, Production incident management

Technologies

Swift, C++, Python, Go, Rust, gRPC, Protocol Buffers, OpenTelemetry, Splunk, XPC, Instruments

Responsibilities

Engineer continuous improvements in stability and performance for the inference stack, implement new functionality from research, integrate inference code into the full service stack, collaborate with hardware and research teams.

Seniority

Senior, hands-on IC

Sourced via apple · Listed on CareerPlan, which tracks 853,000+ jobs from 20+ sources.