We are seeking highly experienced software engineers (Senior+ level) to evaluate the quality of interactions with modern coding agents such as OpenAI Codex and Claude Code. This is not a traditional engineering role where you will be writing production code. Instead, you will be assessing a more complex aspect: whether the AI model 'thinks' like a great engineer. Your role will involve assessing how AI coding agents behave in real-world scenarios, focusing on the sensibility of their responses, the usefulness of their preambles and reasoning, whether their output reflects strong engineering judgment, and if the interaction feels right to an experienced developer. This role emphasizes engineering 'taste' over mere syntax correctness.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Part-time
Career Level
Senior
Education Level
No Education Listed