LLM大模型评测产品经理-AI创新业务
Core
Define product experience and evaluation standards for Large Language Models (LLMs) from a user perspective, building assessment systems for real-world scenarios to drive model iteration.
Role type
Product Manager for LLM Evaluation
Builds
Stable and credible evaluation frameworks and strategies for LLM capabilities
Domain
Artificial Intelligence / Large Language Models / Product Strategy
Deliverable
production ML models
Required skills
User experience definition, evaluation standard formulation, structured analysis of model behavior, data sensitivity, experimental design, cross-functional collaboration
Preferred skills
LLM/Multimodal/Agent product evaluation experience, understanding of AI technology trends, ability to translate research papers into product strategies
Technologies
LLM, Multimodal, Agent
Responsibilities
Define ideal model states and evaluation metrics; Build evaluation systems for real-world application scenarios; Collaborate with R&D and data science to identify model defects and optimization opportunities; Drive cross-team closed loops for model behavior tracking and optimization; Iterate evaluation methodologies based on industry research and business scenarios