Sr. Engineer, AWS Platform

Ayar LabsSan Jose, CA
$140,000 - $160,000Onsite

About The Position

Ayar Labs is seeking an experienced AWS engineer to build and operate the computing environment used by their engineers for chip design. The role will primarily focus on AWS, with some systems remaining on local hardware. The engineer will be responsible for migrating workloads to the cloud, ensuring the platform can support the demands of chip development, and taking projects from initial problem identification through implementation and into production. This position involves working with engineering and IT teams to make architecture decisions, resolve problems, and improve platform operations. The role also includes production support and on-call responsibilities, with support from existing systems administrators.

Requirements

  • 7+ years building and operating production infrastructure, including at least 3 years of hands-on responsibility for AWS environments.
  • A track record of taking a compute-intensive workload from design through deployment or migration on AWS and supporting it in production.
  • The ability to diagnose difficult problems in Linux-based environments and across cloud, network, and storage boundaries.
  • Experience implementing and maintaining production infrastructure through code, using Terraform or OpenTofu.
  • Sound judgment about how to build and operate a secure AWS environment, balancing performance, reliability, and cost.
  • The ability to turn an engineering need into a workable solution, carry it through implementation, and explain decisions clearly.
  • Leave documentation that helps others operate what you build.

Nice To Haves

  • Supporting chip-design or other engineering workloads, particularly where large compute jobs, shared data, and software licenses need to work together.
  • Building an AWS environment that integrates with local infrastructure, including secure remote access for engineers.

Responsibilities

  • Build and operate an AWS platform that gives engineers dependable access to the compute and data they need.
  • Make practical decisions about performance, security, cost, and how cloud and local systems work together.
  • Lead workload migrations from planning through production, understanding workload needs, testing approaches, and managing transitions without disrupting engineering work.
  • Investigate problems that slow down or interrupt design work, following issues across the cloud platform and connected systems, identifying causes, and implementing lasting fixes.
  • Automate provisioning and routine operations to ensure the environment is repeatable, changes can be tested, and the platform is straightforward to maintain.
  • Work with engineers to anticipate demand ahead of major design milestones, ensuring capacity, data access, and software licensing can support planned runs, and explaining constraints early.
  • Protect design data and keep the platform recoverable by establishing appropriate access, monitoring, and recovery practices, and taking responsibility for resolving production incidents.
  • Measure whether changes improve engineering turnaround time, reliability, and cost.
  • Document decisions and operating procedures so others can support and extend the platform.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service