We are looking for a Machine Learning Engineer to join our core research and development team focused on recovering accurate 3D human body and hand motion from egocentric (first-person) video. Human demonstration data is the foundation of robot learning, and its quality depends on accurately reconstructing human motion. In this role, you will develop models and production pipelines that transform head-mounted and body-mounted camera streams—including wide-FOV, stereo, motion-blurred, and heavily self-occluded video—into metrically accurate, temporally consistent 3D pose representations for robot policy training and human-to-robot motion retargeting. You will work across the entire perception stack, including camera calibration, data annotation, model training, evaluation, and large-scale deployment. This role is ideal for engineers with strong expertise in both computer vision and deep learning who enjoy solving challenging real-world perception problems.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level