Production Engineering ensures CoreWeave’s cloud runs with world-class reliability, performance, and operational excellence. Herd is our newest innovation: an agentic AI platform that serves as CoreWeave’s intelligent SRE assistant - combining AI reasoning, data infrastructure, and observability into an autonomous operational intelligence layer for internal use. As a Production Engineer on Herd, you’ll define and build the systems that power a scalable agentic ecosystem. You’ll design distributed services and data pipelines that process, embed, and retrieve operational knowledge at scale, enabling LLM-powered agents to work alongside human engineers in production. This is a hands-on role at the intersection of AI operations, distributed systems, and data infrastructure.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
No Education Listed
Number of Employees
501-1,000 employees