DevOps Engineer II

D-WaveNew Haven, CT
$126,000 - $172,000Hybrid

About The Position

D-Wave is seeking a DevOps Engineer (Platform Reliability & Observability) to join our DevOps team in New Haven, reporting to the DevOps Engineering Manager. In this role, you will support and operate hybrid infrastructure platforms spanning on-premises environments, Kubernetes clusters, cloud services, and CI/CD pipelines. This role is remote-friendly, with a preference for candidates located within a few hours of our New Haven office to support occasional on-site collaboration. Your primary focus will be on improving system reliability, observability, and operational visibility across both on-prem and cloud environments. This includes supporting uptime, operational visibility, and service reliability objectives for D-Wave’s QCaaS platform and hardware-integrated systems spanning Kubernetes, AWS, and on-prem infrastructure. Working across infrastructure, cloud, and developer workflows, you will play a key role in troubleshooting issues, improving deployment processes, and ensuring systems are well-instrumented, scalable, and reliable. This role is ideal for an engineer who enjoys hands-on problem solving, working across systems, and continuously improving how platforms operate in practice.

Requirements

  • Bachelor’s degree in computer science, Engineering, or equivalent practical experience
  • 3+ years of experience in DevOps, infrastructure, systems engineering, or a related field
  • Strong Linux fundamentals and troubleshooting skills in distributed environments
  • Hands-on experience with AWS and cloud-based infrastructure
  • Experience working with CI/CD systems (e.g., GitHub Actions)
  • Experience with infrastructure-as-code tools such as Terraform
  • Familiarity with containerization using Docker and deploying containerized applications
  • Exposure to Kubernetes concepts and container orchestration environments
  • Experience with monitoring and observability tools (e.g., Zabbix, CloudWatch, OpenSearch, Grafana)
  • Familiarity with relational databases (e.g., PostgreSQL, MySQL/InnoDB, Amazon Aurora), including basic performance troubleshooting, connectivity, and operational considerations
  • Scripting experience (e.g., Python, Bash) to support automation and debugging
  • Solid understanding of networking fundamentals (e.g., DNS, routing, firewalls)
  • Strong problem-solving skills and ability to debug issues across multiple systems
  • Comfort working in a fast-moving environment with evolving systems and priorities
  • Desire to learn and grow in areas such as Kubernetes, observability, infrastructure automation, and platform reliability

Responsibilities

  • Improve observability across systems by developing and maintaining logging, metrics, dashboards, and alerting solutions
  • Work with tools such as Zabbix, OpenSearch, CloudWatch, and Grafana to improve system visibility, monitoring coverage, and operational reliability across hybrid cloud and hardware-integrated platforms
  • Investigate and resolve issues across infrastructure, Kubernetes workloads, CI/CD pipelines, and cloud services
  • Support the operation of Kubernetes platforms (on-prem and cloud), including deploying workloads, troubleshooting issues, and improving system reliability
  • Build, maintain, and troubleshoot CI/CD pipelines (e.g., GitHub Actions) to support reliable and repeatable deployments
  • Assist with troubleshooting and reliability improvements for stateful services and relational database platforms (e.g., PostgreSQL, MySQL/InnoDB, Amazon Aurora)
  • Assist in maintaining hybrid infrastructure across on-prem and AWS environments
  • Implement and maintain infrastructure-as-code using tools such as Terraform and Ansible
  • Support containerized applications, including building, deploying, and debugging Docker-based workloads
  • Help improve alert quality and reduce noise to support effective incident response
  • Support internal platform efforts, including development environments and shared infrastructure services
  • Collaborate with engineers across hardware, software, and infrastructure teams to resolve issues and improve system reliability
  • Contribute to automation efforts to reduce manual operational work and improve consistency across environments

Benefits

  • company ownership
  • competitive pay
  • a range of meaningful benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service