Innodata is a global data engineering company focused on enabling the responsible advancement of artificial intelligence. We provide data, evaluation frameworks, and human expertise to build trustworthy AI systems. We offer a range of solutions, platforms, and services for Generative AI / AI builders and adopters, leveraging our 36+ year legacy of delivering high-quality data and outstanding outcomes. This role involves evaluating the performance of AI models through voice-based conversations. Specialists will interact with two different AI models using the same assigned scenario, compare their responses, and provide structured evaluations based on defined quality criteria. The primary goal is to ensure a fair and consistent comparison between models and identify which model offers a superior conversational experience.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Part-time
Career Level
Entry Level