About The Position

We develop next-generation AI infrastructure, cloud platforms, data infrastructure, developer platforms, and site reliability engineering (SRE) capabilities supporting engineering teams and manufacturing facilities worldwide. As TSMC accelerates AI adoption across the enterprise, we are building enterprise-scale AI platforms that enable secure, scalable, and reliable AI services across global regions. We are seeking a talented and experienced AI Infrastructure Engineer to define and build TSMC's enterprise AI Infrastructure Platform. You will design and lead enterprise-scale AI infrastructure that enables thousands of engineers across multiple global regions to securely develop, deploy, and operate AI applications using both internal and external foundation models. You will collaborate with engineering teams in North America and Taiwan to build a highly scalable AI platform supporting model serving, inference, AI gateways, GPU infrastructure, observability, governance, and platform automation. This role combines deep expertise in distributed systems, cloud infrastructure, AI platforms, and software architecture.

Requirements

  • BS/MS/PhD in Computer Science or related field.
  • 7+ years of software engineering experience.
  • 3+ years leading architecture for distributed systems or cloud infrastructure.
  • 2+ years hands-on experience designing enterprise AI platforms.
  • 1+ years’ experience as a Principal Engineer, Staff Engineer, Distinguished Engineer, or Technical Lead.
  • Strong technical communication skills for collaborating with global cross-functional teams.
  • AI Infrastructure, strong experience with: LLM inference, Enterprise AI Gateway, AI serving infrastructure, GPU scheduling, Multi-model AI architecture, Model lifecycle management.
  • Experience with technologies such as: NVIDIA GPU ecosystem, vLLM, TensorRT-LLM, Triton Inference Server, Ray, Kubeflow, MLflow.
  • Cloud Native Infrastructure Deep expertise with: Kubernetes, Docker, Helm, ArgoCD, Service Mesh, AWS or Azure or GCP.
  • Strong programming skills in one or more of Go, Python, JAVA.
  • Experience with: REST APIs, Microservices, SDKs.
  • Applicants must have legal authorization to work in the United States. We currently cannot provide sponsorship or take over sponsorship of an employment visa.
  • Employment at TSMC is contingent upon passing a background check and drug screening. In compliance with Washington state regulations, cannabis (marijuana) use will not be included in the drug screening.

Responsibilities

  • Define AI Platform Architecture. Lead the architecture and technical strategy for enterprise AI infrastructure. Drive long-term technical direction for: AI Platform, Enterprise AI Gateway, Multi-model inference platform, GPU infrastructure, AI governance, AI observability.
  • Build Enterprise AI Infrastructure: Design highly scalable platforms supporting LLM serving, GPU scheduling, Model routing, Model lifecycle management, RAG infrastructure, Vector databases, AI orchestration.
  • Design Distributed Systems: Lead architecture for high availability AI services, multi-region deployment, disaster recovery, service mesh, distributed caching, event-driven architecture, global load balancing.
  • AI Platform Engineering: Drive engineering best practices for: Kubernetes, Platform-as-a-Service, AI deployment automation.

Benefits

  • Market-competitive pay
  • Profit sharing and incentive bonuses
  • Tuition assistance
  • Medical, dental, and vision insurance
  • Life insurance
  • Access to a 401(k) plan with employer match
  • 12 holidays per year
  • Accrued paid time off annually
  • Onsite amenities include a fitness center, game room, physical therapist, and subsidized café.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service