Innodata is a global data engineering company focused on enabling the responsible advancement of artificial intelligence. We provide the data, evaluation frameworks, and human expertise needed to build trustworthy AI systems at scale. We offer a range of solutions, platforms, and services for Generative AI/AI builders and adopters, leveraging our 36+ year legacy of delivering high-quality data and outstanding customer outcomes. This role involves evaluating the performance of AI models through real-time, voice-based conversations. Specialists will interact with two different AI models using the same assigned scenario, compare their responses, and provide structured evaluations based on defined quality criteria. The primary goal is to ensure a fair and consistent comparison between models and identify which model delivers a stronger overall conversational experience.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Part-time
Career Level
Entry Level