Innodata is a global data engineering company focused on the intersection of data and Artificial Intelligence (AI). Our mission is to enable the responsible advancement of AI by providing essential data, evaluation frameworks, and human expertise for building trustworthy AI systems at scale. We offer a range of solutions, platforms, and services for Generative AI/AI builders and adopters, leveraging our 36+ year legacy of delivering high-quality data and outstanding customer outcomes. This role focuses on the scientific aspects of speech and audio data for AI models. You will partner directly with customers and frontier labs working on ASR, text-to-speech, speech-to-speech, conversational voice, diarization, and audio-language models. Your work will involve critical judgment on benchmark design, transcription conventions, and the appropriate use of automated versus human evaluation metrics. You will also collaborate closely with our transcription and linguistics lead to ensure model learning is shaped by precise standards.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior