Epic Site Reliability Engineer II

Quest DiagnosticsSchaumburg, NJ
Hybrid

About The Position

This hybrid position requires 3 days on-site at a Quest location in Secaucus, NJ, or Schaumburg, IL. The role focuses on implementing and maintaining observability solutions, optimizing system performance, capacity planning, automation, documentation, and collaboration with development teams to ensure system reliability and efficiency. The engineer will also be involved in security and compliance, training, and continuous improvement initiatives.

Requirements

  • 4+ years of experience with multiple APM tools and extensive experience with Dynatrace
  • 3+ years SRE experience
  • Experience in software development, infrastructure, or operations roles
  • Certifications in relevant technologies (e.g. AWS, DevOps, Kubernetes, Dynatrace, Azure, etc.)
  • Working experience building CI/CD pipelines and version control systems
  • Working experience with scripting languages (e.g. Python, Bash, Go, etc.)
  • Excellent problem-solving and communication skills.
  • Ability to work collaboratively in a fast-paced, agile environment.
  • Bachelor's Degree in Computer Science, Engineering, or a related field (Required)

Nice To Haves

  • Working experience with Neoload, Jmeter or equivalent performance testing tool.
  • Experience executing software load and performance testing in an enterprise environment.
  • Experience testing applications hosted in the cloud.
  • Experience with infrastructure as code tools such as Terraform or CloudFormation.
  • Deep understanding of Linux systems administration and networking principles.
  • Experience with containerization and orchestration technologies such as Docker and Kubernetes.
  • Experience or familiarity with IIS, HTML, Java, Jboss.
  • Experience in Chaos Engineering
  • Programming experience using.NET, C, C++, Java, or other popular programming languages.
  • Perl/Python/JavaScript scripting experience may be considered equivalent.
  • Terraform and Ansible experience.
  • Exposure to Splunk tools.
  • Exposure to microservices.
  • Dynatrace Certifications
  • AWS/Azure/GCP Certifications
  • Chaos Engineering Certifications
  • Agile Certifications
  • Agile Certification (Project Management) (Preferred)

Responsibilities

  • Implement and maintain robust observability solutions to monitor system performance, identifying bottlenecks, and ensuring optimal operation.
  • Utilize tools to gather, analyze, and visualize key performance metrics.
  • Proactively identify and address performance bottlenecks through in-depth analysis and optimization strategies.
  • Work closely with development teams to implement performance improvements and enhance overall system efficiency.
  • Conduct capacity planning exercises based on observed patterns and future growth projections.
  • Collaborate with infrastructure and development teams to ensure adequate resources are available to meet system demands.
  • Develop and maintain automation scripts for routine tasks, enabling efficient monitoring and response procedures.
  • Implement automated processes for scaling and provisioning resources based on observed workload patterns.
  • Document system architecture, configurations, and observability best practices to facilitate knowledge transfer and onboarding for team members.
  • Keep documentation up-to-date to reflect changes in the system and its monitoring setup.
  • Work closely with software engineers to integrate observability tools into the development lifecycle.
  • Provide guidance on building observable systems and assist in instrumenting applications for effective monitoring.
  • Stay informed about industry best practices and emerging technologies related to observability and performance engineering.
  • Drive continuous improvement initiatives to enhance the reliability and performance of systems.
  • Collaborate with security teams to implement monitoring and observability measures that align with security requirements and compliance standards.
  • Participate in security incident response activities and contribute to ongoing security assessments.
  • Conduct training sessions for team members and other stakeholders on observability tools, best practices, and performance engineering concepts.
  • Foster a culture of knowledge sharing within the organization.
  • Perform other duties as assigned.

Benefits

  • Day 1 Medical
  • supplemental health
  • dental & vision for FT employees who work 30+ hours
  • Best-in-class well-being programs
  • Annual, no-cost health assessment program
  • Blueprint for Wellness healthyMINDS mental health program
  • Vacation and Health/Flex Time
  • 6 Holidays plus 1 MyDay off
  • FinFit financial coaching and services
  • 401(k) pre-tax and/or Roth IRA with company match up to 5% after 12 months of service
  • Employee stock purchase plan
  • Life and disability insurance, plus buy-up option
  • Flexible Spending Accounts
  • Annual incentive plans
  • Matching gifts program
  • Education assistance through MyQuest for Education
  • Career advancement opportunities
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service