Agent效果评测专家 - 开发者服务
Core
Define evaluation standards and acceptance criteria for AI Agents in software engineering scenarios, build high-fidelity test sets, and design quantitative metrics to systematically evaluate Agent performance.
Role type
Senior IC AI Agent Evaluation Engineer
Builds
High-fidelity test sets, quantitative evaluation metrics, and automated evaluation pipelines for AI Agents
Domain
Software Engineering + AI Agents
Deliverable
production ML models
Required skills
Go/Python/Java, AI Agent architecture (Multi-Agent, Context Engineering, ReAct), LLM evaluation frameworks, root cause analysis, automated testing
Preferred skills
Agent development experience, AI paper publication, large model training experience
Sourced via bytedance · Listed on CareerPlan, which tracks 860,000+ jobs from 20+ sources.