We are sharing a specialised full-time consulting opportunity for machine learning engineers and research practitioners with hands-on experience training, evaluating, and experimenting with ML models end to end. This role supports the development of advanced agentic evaluation benchmarks for frontier AI systems. Selected professionals will transform real machine learning research ideas into rigorous multi-step tasks, implement and run experiments, analyse training behaviour, and evaluate where model-generated solutions fall short of technically correct results.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level