Founding Engineer

Double Blind BioSan Francisco, CA

About The Position

As an AI-focused Software Engineer at Double Blind Bio, you will help build and operate intelligent systems that transform how clinical trials run: reducing administrative burden and accelerating the development of life-saving medicines. This is a high-impact, product-first role where your work directly contributes to improving patient outcomes and advancing scientific research. You will be a core contributor on a fast-moving, early-stage team, owning AI-powered features end-to-end and helping shape both the product and technical foundation. You are a software generalist who brings genuine depth in AI systems—not just in building them, but in measuring, evaluating, and improving them over time. You have strong instincts around when AI is working, when it isn't, and how to tell the difference. This role is ideal for engineers who thrive at the intersection of AI, healthcare, and product development, and who are energized by high ownership, rapid iteration, and close collaboration with users.

Requirements

  • Strong full-stack engineering experience, with the ability to build and ship production-ready applications
  • Genuine depth in AI/ML systems—you understand how models behave, how to evaluate outputs rigorously, and how to think about failure modes beyond "the model got it wrong"
  • Experience designing and running evals—you've built evaluation sets, defined metrics, and used data to drive decisions about AI system quality
  • Strong observability instincts—you know how to instrument pipelines, trace model calls, and build feedback loops that surface problems before users do
  • Experience building with LLMs in production (RAG, agents, prompt engineering, fine-tuning, or similar)
  • Product-oriented mindset with a focus on solving real user problems
  • Comfort working across disciplines and learning new domains (e.g., healthcare, life sciences, data systems)
  • Ability to work in a fast-paced, evolving environment with high ownership and autonomy
  • Strong communication skills and willingness to collaborate across teams and with customers

Nice To Haves

  • Experience with ML evaluation frameworks, LLM observability tooling (e.g., LangSmith, Braintrust, Weights & Biases, Honeyhive, or similar)
  • Familiarity with statistical thinking around model evaluation—confidence intervals, human-in-the-loop review, A/B testing of AI outputs
  • Experience building agentic workflows or multi-step AI pipelines
  • Familiarity with healthcare, life sciences, or clinical research workflows
  • Experience working in early-stage startups or founding teams
  • Deep expertise in at least one area of the stack (frontend, backend, infrastructure, or AI systems)

Responsibilities

  • Design, build, and deploy agentic AI systems that power web experiences and automate clinical workflows
  • Own the full eval lifecycle—define what "good" looks like for AI outputs, build evaluation datasets, author metrics, and run structured evals to measure model and pipeline quality
  • Instrument AI systems for observability—trace LLM calls, monitor output quality, detect regressions, and build tooling that gives the team visibility into how AI is behaving in production
  • Develop and maintain RAG pipelines, prompt chains, and AI-driven features with a systematic approach to improvement
  • Architect and optimize full-stack applications, including frontend, backend, infrastructure, and data pipelines
  • Own features end-to-end—from ideation and design through deployment, monitoring, and iteration
  • Collaborate with product, founders, and customers to understand needs and deliver impactful solutions
  • Contribute to engineering processes, architecture, and best practices as an early team member
  • Participate in customer conversations to better understand workflows and continuously improve the product
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service