Productive Playhouse is building a talent pool of Indonesian speakers for an upcoming project testing and evaluating leading AI chatbots. The objective is to enhance response quality through direct user interaction with the various AI models. This is an independent contractor engagement (task-based - project-based). The project is iterative, with work arriving in batches, and there may be pauses between them. As an AI Evaluator, you will play a critical role in shaping and improving the next generation of Generative AI. You will participate in structured, hands-on evaluations by interacting directly with various AI models to assess their capabilities, safety, and helpfulness. Your insights and data will directly inform model development and optimization. Exact tasks vary by project and will be spelled out in that project's Statement of Work (SOW) before you start. Depending on the role, your work may include testing and evaluating AI chatbots or language models through structured conversations, using assigned goals and prompts; voice-based interactions with AI models (sessions are recorded, and audio may be shared with the client as part of the evaluation deliverable); and submitting deliverables (write-ups, ratings, screenshots, or recordings) in the format each task specifies.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Career Level
Entry Level
Education Level
No Education Listed