Senior AI Engineer - Services Special Projects
Core
Build, deploy, optimize, and operationalize Small and Large Language Model (LLM)-based applications at Apple scale, focusing on MLOps/LLMOps and scalable production systems.
Role type
Senior IC machine-learning engineer (LLM/MLOps)
Builds
LLM-powered features, CI/CD pipelines, serving infrastructure, observability tools, and governance workflows
Domain
Consumer technology, Large Language Models, Distributed Systems
Deliverable
production ML models
Required skills
LLM fine-tuning, prompt engineering, RAG patterns, model quantization/distillation/compilation, vector search, feature stores, distributed systems, cloud platforms, container orchestration, CI/CD, Python, Java/Scala
Preferred skills
Go, graph databases, SLA definition, open-source contributions, LLM safety guardrails, adversarial evaluation
Technologies
ONNX Runtime, TensorRT, Triton, vLLM, SGLang, TorchServe, Pinecone, Milvus, pgvector, FAISS, Feast, Airflow, MLflow, DVC, Kubernetes, AWS, FastAPI, Flask, Spark, Flink, Kafka, Prometheus, Grafana, Datadog, LangSmith
Responsibilities
Own the full model lifecycle from experimentation to retirement; fine-tune models and optimize for production; design scalable ML infrastructure and experimentation platforms; implement CI/CD methodologies and production serving infrastructure; drive model observability and incident response; implement model governance workflows and safety guardrails; establish robust versioning strategies for datasets and artifacts; mentor engineers and set technical standards.
Seniority
Senior, hands-on IC