AI Evaluation Engineer
在手机上跟进您的申请 Whileresume 免费应用,支持 iPhone 和 Android。
职位介绍
Mindrift connects specialists with project-based AI opportunities in testing, evaluating, and improving AI systems. The role involves building datasets to evaluate AI coding agents, creating realistic developer environments, and crafting challenging tasks and evaluation criteria. Candidates will review agent solutions, analyse failures, and refine tasks based on feedback, requiring a deep understanding of where models fail. Experience in software development with Python, JavaScript/TypeScript, Docker, and databases is essential. The position requires English proficiency of B2+ and offers flexible hours at competitive rates. This role is not about data labelling, prompt engineering, or writing code from scratch, but rather guiding and evaluating AI-generated code.
查看完整职位
工作职责、任职要求、技能与福利 — 免费创建账号即可查看。
或
已有账户?
登录您可能也感兴趣的职位
暂无高度相似的职位 — 以下是最新职位。