Infrastructure Software Engineering Intern - Fall

Fab2San Francisco, CA
Onsite

About The Position

The fab2 team is responsible for designing and building all the necessary hardware and software for chip manufacturing, as well as the fabs, tools, and components themselves. They are seeking exceptional, hands-on individuals who can push the boundaries of what's possible. This role is for an Infrastructure and Site Reliability Intern for the fall term, starting in September with a preferred commitment of 4 to 8 months. The intern will design, build, deploy, and manage the on-prem backend infrastructure powering a semiconductor fab. This is a broad role covering all aspects of backend infrastructure and services, with a philosophy of minimal, understandable, on-site, and close-to-the-hardware systems. The environment emphasizes bare-metal Linux, systemd, and single-file binaries, with a focus on Rust and Go, and occasional Python. The ideal candidate has built real things, enjoys working close to the hardware, and demonstrates strong engineering excellence. This role is suitable for those excited by performance engineering, building complex features from scratch, and rapid learning. The company's philosophy prioritizes building and testing in days or weeks, not months. A portfolio (e.g., GitHub) showcasing software engineering excellence and curiosity is required.

Requirements

  • Pursuing a BS in Computer Science, Computer Engineering, or demonstrated exceptional skill in software engineering
  • Strong programming skills in Rust, Go, or other systems-oriented languages
  • Solid understanding of Linux systems, networking fundamentals, and distributed systems concepts
  • Experience building or maintaining non-trivial infrastructure, automation, or tooling projects (personal, academic, or professional)
  • Interest in one or more of the following areas: systems reliability, observability, infrastructure automation, performance optimization
  • A portfolio (e.g., GitHub) is required to apply

Nice To Haves

  • Experience with configuration management or fleet management at scale
  • Background in monitoring and observability tools (Prometheus, Grafana, ELK stack, Datadog)
  • Familiarity with performance profiling and capacity planning
  • Experience with database administration, backup strategies, or disaster recovery
  • Exposure to secret management systems (Vault, SOPS) or PKI

Responsibilities

  • Build and evolve infrastructure-as-code tooling to manage complex, multi-environment deployments
  • Manage and maintain the fleet of on-premises servers, ensuring high availability and optimal resource utilization
  • Monitor and maintain system performance and health across diverse production workloads
  • Design, implement, and refine alerting systems to catch issues before they impact users
  • Build automation to reduce toil and improve operational efficiency
  • Implement secure secret management
  • Debug production issues across the stack
  • Work closely with engineering teams to understand their infrastructure needs and deliver robust, scalable solutions

Benefits

  • Housing Stipend to help with first month expenses
  • Lunches daily
  • Dinners 3x per week
  • Stocked Office Kitchen with Snacks and Spindrifts
  • Weekly Learning & Development opportunities
  • Commuter Benefits including Parking and Late Night Uber rides from the office
  • Paid Time Off inclusive of Holidays and Sick Time
  • Visa Sponsorship
  • Medical, Dental, and Vision insurance
  • 401(k) retirement plan
  • Life and Disability Insurance
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service