CareerPlanSign in

LLM大模型评测产品经理-AI创新业务

北京💼 Full-time🗓 2026-09-28

Core

Define product experience and evaluation standards for Large Language Models (LLMs) from a user perspective, building assessment systems for real-world scenarios to drive model iteration.

Role type

Product Manager for LLM Evaluation

Builds

Stable and credible evaluation frameworks and strategies for LLM capabilities

Domain

Artificial Intelligence / Large Language Models / Product Strategy

Deliverable

production ML models

Required skills

User experience definition, evaluation standard formulation, structured analysis of model behavior, data sensitivity, experimental design, cross-functional collaboration

Preferred skills

LLM/Multimodal/Agent product evaluation experience, understanding of AI technology trends, ability to translate research papers into product strategies

Technologies

LLM, Multimodal, Agent

Responsibilities

Define ideal model states and evaluation metrics; Build evaluation systems for real-world application scenarios; Collaborate with R&D and data science to identify model defects and optimization opportunities; Drive cross-team closed loops for model behavior tracking and optimization; Iterate evaluation methodologies based on industry research and business scenarios

Sourced via bytedance · Listed on CareerPlan, which tracks 854,000+ jobs from 20+ sources.