Applied Researcher 3 - Multimodal AI
Core
Develop multimodal intelligence and visual recognition models to power search, recommendations, content generation, and video understanding for eBay's global ecommerce platform.
Role type
Senior Applied Researcher (Multimodal AI & Computer Vision)
Builds
Production multimodal models, agentic VLM-powered pipelines, and end-to-end training/inference systems for search and video.
Domain
Ecommerce, Computer Vision, Multimodal AI
Deliverable
production ML models
Required skills
Computer vision, multimodal learning, visual-language modeling, agentic AI workflows, scalable ML infrastructure, distributed training, model serving, Python, PyTorch, TensorFlow, Spark
Preferred skills
Experience with reinforcement learning, model distillation, orchestration of complex inference systems
Responsibilities
Develop and deploy visual recognition models for large-scale applications; Design and optimize end-to-end training and inference pipelines; Advance core capabilities in detection, segmentation, classification, and contrastive learning; Architect agentic VLM-powered solutions; Lead R&D in VLM training and fine-tuning; Partner with product and engineering teams to translate business opportunities into technical roadmaps; Mentor junior and senior team members.
Seniority
Senior, hands-on IC with mentorship responsibilities
