Chakra Labs' mission is to encode human taste into intelligence. We build high-fidelity environments, evals, and datasets for frontier AI research, working with several of the top labs. Our work sits at the frontier of post-training, agent environments, data quality, and research infrastructure. We care about building systems that make models better in ways that are measurable, useful, and hard to fake. The hardest problems at the frontier. A new environment modality, an eval targeting a failure mode nobody's measured, a dataset that doesn't exist yet. You take problems like these from a researcher's hunch to a shipped deliverable, working at the edge of what agents can currently do. Environments, evals, and datasets. One project is a high-fidelity environment, the next is a task distribution with grading logic, the next is a dataset built to a demanding spec. The bar is frontier-lab quality and the pace is relentless - you're writing whatever the deliverable needs: environment code, task specs, scoring harnesses. Pulling the frontier into the platform. The best one-offs don't stay one-offs. You'd recognize when a custom build proves out a capability worth generalizing, and help fold it into the core product - so it compounds instead of sitting on a shelf.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
No Education Listed