Multi-Agent systems are becoming an increasingly important part of how AI is deployed, whether via fast small-model subagents inside a product, or large groups of agents solving very large problems. Training Claude to be maximally effective and safe within large groups is a challenging new area of reinforcement learning, and represents a new axis for scaling test time compute. We are looking for researchers who have experience training multi-agent systems at the largest scale and an appreciation for the incentives and mechanism design that come into play.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior