TwelveLabs is building the intelligence layer for video data, enabling machines to understand video content across sight, sound, and motion. The company has raised over $210 million from prominent investors and partners. Headquartered in San Francisco with offices in Seoul, New York, and London, TwelveLabs values diversity and seeks individuals driven by challenging problems. This role focuses on the integration layer that makes TwelveLabs' video AI accessible to external developers and AI agents. The engineer will own the "Jockey" multimodal agent, which is built on a knowledge store, retrieval primitives, and an agent layer for planning and reasoning. The core responsibility is to design and operate the surfaces that allow AI agents and developers to connect to these capabilities, ensuring trustworthiness at an enterprise scale. This includes owning the MCP server end-to-end, handling capability discovery, invocation, tool interfaces, failure modes, and the evolution of the surface. The role also involves building the supporting infrastructure for authentication, metering, and rate limiting to ensure enterprise trust.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed