This is an early-stage AI company building technology that helps enterprises evaluate, validate, and deploy increasingly capable AI systems with confidence. The company is focused on solving difficult problems around AI evaluation, benchmarking, and the reliability of language models. Its work helps determine how advanced AI systems perform across real-world enterprise use cases and provides the methodologies needed to measure and improve their quality. The engineering and research teams work at the intersection of applied AI research and production software engineering. You will evaluate new models as they are released, build benchmarks from scratch, develop evaluation methodologies, and work closely with engineering teams to turn research ideas into scalable systems. As a Member of Technical Staff — Research Scientist, you will have significant ownership over how AI systems are evaluated. You will work with large datasets, language models, evaluation pipelines, and automated assessment methods while collaborating with AI labs and enterprise customers. This is a highly hands-on applied research role for someone who enjoys building practical evaluation systems, experimenting with modern AI models, and turning research ideas into production-ready methodologies. The role prioritizes applied impact over purely academic research or publishing papers for its own sake.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Entry Level
Education Level
No Education Listed