CareerPlanSign in

AIML - Machine Learning Research Lead, RL Agents, MLR

Cupertino, United States of America💼 Full-time🗓 2026-09-09 → 2026-09-28

Core

Lead research on reinforcement learning and post-training for agentic AI, managing a small team of senior researchers.

Role type

Senior IC research lead (RL & post-training)

Builds

Interactive agents with tool-calling, coding, and long-horizon capabilities

Domain

Artificial Intelligence / Reinforcement Learning

Deliverable

production ML models

Required skills

Reinforcement learning, post-training, agentic AI, synthetic environment generation, ML infrastructure, team leadership, hands-on coding

Preferred skills

Efficiency-aware modeling, open-endedness, evolutionary computation, multi-agent systems, theoretical RL grounding

Technologies

Large language models, self-supervised learning, generative models

Responsibilities

Lead research on RL and post-training for agentic capabilities; Build and own synthetic data and task-generation pipelines; Drive codebases and infrastructure for the core RL research effort; Manage and mentor a small team of senior researchers; Stay hands-on with experiments and code; Connect post-training research to efficiency and deployment constraints; Collaborate on adjacent directions like world models and self-improvement; Publish in top venues and engage with the research community.

Sourced via apple · Listed on CareerPlan, which tracks 872,000+ jobs from 20+ sources.