AI Research Intern

OpusClipMountain View, CA
$25 - $28Hybrid

About The Position

OpusClip is seeking an AI Research Intern to join their AI team. This role involves exploring cutting-edge research in multimodal AI, LLMs, computer vision, speech, and agent systems. The intern will work across OpusClip, AgentOpus, and next-generation AI products, collaborating with AI researchers and engineers to investigate emerging technologies, build research prototypes, and contribute to features used by millions of creators globally. OpusClip is a well-funded startup with significant traction in the AI video agent space, recognized by publications like Business Insider and The Information.

Requirements

  • Currently pursuing or recently completing a Master's degree in Computer Science, Artificial Intelligence, Mathematics, or a related field.
  • Solid understanding of Transformer architecture and Attention mechanisms.
  • Familiarity with mainstream generative model families (GANs, diffusion models, autoregressive models).
  • Familiarity with media processing fundamentals (video and/or audio — e.g., ffmpeg, codecs, signal processing basics).
  • Strong programming skills in Python.
  • Familiarity with Linux development environments, Git, and data structures.
  • Fluent in English with strong technical reading and writing skills.
  • Hands-on experience in one or more of the following: Computer Vision (especially low-level vision), Voice/speech (voice cleaning, voice cloning, voice generation), or LLM fine-tuning (SFT, RLHF/DPO, LoRA/PEFT, post-training of open-source models).

Nice To Haves

  • Experience building Agent Systems or LLM-powered product features with frontier-model APIs (e.g., ChatGPT, Claude, Gemini).
  • Familiarity with TypeScript.
  • Involvement in projects from inception to completion, with strong coding fundamentals; open-source contributions are a plus.
  • Academic background or interest in adjacent areas — video understanding and generation, multimodal systems, agents, and model evaluation/benchmarking.
  • Involvement or interest in academic research, with a focus on top-tier venues like CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, ICASSP, Interspeech, AAAI, MM, TIP, TPAMI, ACL, EMNLP etc.

Responsibilities

  • Research and develop deep learning models in areas such as computer vision (video enhancement, super-resolution, restoration), speech & audio (speech enhancement, voice cloning, voice generation), multimodal understanding and generation, and LLM post-training (SFT, RLHF, DPO).
  • Build AI-powered product features by integrating foundation models into production systems using prompt and context engineering, and Agent workflows (e.g., LangChain, RAG frameworks).
  • Collaborate with product and engineering teams to rapidly prototype and ship new AI capabilities.
  • Design scalable evaluation pipelines for multimodal AI systems.
  • Develop domain-specific benchmarks using automated evaluation methods and quality metrics.
  • Stay updated with the latest AI research and open-source developments.
  • Reproduce state-of-the-art research and translate new advances into production-ready systems.

Benefits

  • Flexible remote/on-site internship (3 days/week required, 4+ days/week preferred).
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service