Agent Evaluation Engineer for an AI Agent Company

Dadaconsultants · Singapore

Sector
AI
Function
Product & Engineering
Level
Mid-Level
Employment type
Full Time
Posted
2026-09-29
Source
mycareersfuture

About our clientOur client is a fast-growing, well-funded AI agent company. The product takes on complex work like deep research, data analysis and software development, and carries it through to a finished result. It is a small team building at the frontier, based in Singapore.About the roleYou will develop evaluations grounded in product needs, user tasks and how AI models and agents actually work. Using metrics, experiments and failure analysis, you'll assess whether capability changes are real, investigate gaps in existing evaluations, and inform system development, post-training and model selection.What you'll doDesign online and offline metrics that translate user tasks, output quality and practical value into measurable, testable evaluation criteriaDesign evaluation tasks and experiments around agent planning, tool use, context and feedback, comparing performance before and after system changesSupport post-training evaluations and comparisons of third-party model quality, defining use cases and the conditions results apply toDevelop new tasks, metrics or experimental methods for issues existing evaluations miss, and test their validity, bias and reproducibilityAnalyze evaluation results and failures, distinguish score changes from real capability changes, and work with product, engineering and model teams to validate improvementsWhat you bringPractical experience evaluating agent products or model post-training, with concrete evidence of metric design and validationDeep understanding of AI model and agent mechanisms — task planning, tool use, context management and feedbackStrong data analysis, engineering and research skills, including experiment design and handling uncertainty in resultsAbility to investigate open-ended problems, develop well-reasoned new evaluation methods, and examine experimental biasIndependent judgement on AI output quality and real user value, with the ability to explain findings and their limitations clearlyNice to haveFamiliarity with user research methods (interviews, observation, usability testing), and the ability to translate findings into evaluation criteriaWhat we offerBuild at the frontier of AI agents with a fast-paced teamUnlimited access to our own AI toolsCompetitive salary, reviewed for the right personIf it sounds like your next move, please don’t hesitate to apply. Kindly note that only shortlisted candidates will be contacted. Appreciate your understanding. Data provided is for recruitment purposes only.About UsDada Consultants was established in 2017, with the commitment of providing the best recruitment services in Singapore. We are comprised of a dynamic head-hunting team dedicated to sourcing for highly competent professionals in IT industry. We provide enterprises with customized talent solutions, and bring talents to career advancement.EA Registration Number: R23118712Business Registration Number: 201735941W. Licence Number: 18S9037www.dadaconsultants.com

Apply on mycareersfuture →
AI Machine Learning AI Agents Data Analysis A/B testing Artificial Intelligence Model Evaluation AI Evaluation