We are building state-of-the-art context compression, aiming to become the 'Cloudflare for LLMs'. Our mission is to embed a compression layer into most LLM pipelines by default. We are a team of ex-EPFL MSc/PhDs who started by publishing papers, then joined YC and began generating revenue by assisting companies in reducing their LLM expenses. We operate our business like a research lab, forming hypotheses, discarding ineffective ones, and focusing on successful strategies. This internship offers competitive compensation, all necessary resources (GPUs, subscriptions, OpenAI/Anthropic credits), significant responsibility, and a fast-paced learning environment with technically proficient colleagues. There is a possibility of a full-time offer based on performance. However, we do not offer hands-on supervision; guidance is high-level, and interns are expected to own their work. Projects are not pre-defined due to our early-stage, customer-and-market-driven approach, requiring interns to navigate multiple directions. After a brief onboarding, interns will tackle challenging, customer-facing, and time-sensitive problems alongside the team. This is not a typical internship; it's a demanding environment designed for rapid growth and skill development.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Career Level
Intern
Education Level
Associate degree