Open role
Research Engineer, Benchmarks
Clera
Singapore, SingaporePosted Oct 3, 2026 · 2h ago$150k – $250k
Full-time$150k – $250kMid LevelOn-siteAI / ML
About this role
Clera is seeking a Research Engineer to join our small, dedicated team focused on building benchmarks for AI agents. In this role, you will design and implement rigorous evaluations for realistic, domain-specific workflows, helping technical teams understand agent performance in real-world scenarios. You'll own the end-to-end process, from translating expert knowledge into tasks to building scalable infrastructure and analyzing results.
What we are looking for
6- Design and implement AI agent benchmarks for domain-specific tasks
- Collaborate with subject-matter experts to define real-world workflows
- Build and operate scalable infrastructure for model and agent evaluation
- Develop metrics and analyses to assess benchmark reliability and performance
- Validate benchmark results against real-world performance and user needs
- 2+ years experience building AI benchmarks or evaluation infrastructure
