Applied Researcher 1 - Multimodal AI
Core
Develop and deploy computer vision and multimodal machine learning models for production applications in search, recommendations, content generation, and video understanding.
Role type
Applied Researcher 1 (Computer Vision/Multimodal AI)
Builds
Production ML models, agentic/VLM-powered pipelines, and inference/evaluation workflows for large-scale ecommerce applications.
Domain
Ecommerce, Computer Vision, Multimodal AI
Deliverable
production ML models
Required skills
Computer vision techniques, multimodal learning, visual-language modeling, agentic AI workflows, VLM-powered pipelines, experimental methodology, Python, PyTorch, TensorFlow
Preferred skills
Experience in search, recommendations, video or live streaming, classical computer vision, contrastive learning, representation learning
Technologies
PyTorch, TensorFlow, VLMs, Foundation Models
Responsibilities
Develop and deploy CV/multimodal models for production; Build and improve training, inference, and evaluation pipelines; Apply rigorous experimental methodology and metric design; Contribute to core capabilities in detection, segmentation, classification, and visual-language modeling; Build agentic or VLM-powered pipelines; Partner with cross-functional teams to translate business opportunities into ML solutions.
Seniority
Individual Contributor (IC), early career (1+ years experience)
