Site Reliability Engineer II

Akamai•Cambridge, MA
•$138,237 - $171,000•Hybrid

About The Position

Perform site reliability engineering duties in furtherance of ensuring application and network performance, reliability and security for users of our platforms; deploy and operate scalable, highly available, and faulttolerant systems; analyze and improve the security, stability, speed, and capacity of our network; automate new cloud service deployments; monitor and troubleshoot cloud services to ensure SLA is met and adequate capacity is available; analyze logs and events from the cloud service and providing recommendations to ensure smooth service operation; work within large-scale cloud-based infrastructure; quickly identify and resolve infrastructure issues.

Requirements

  • Master’s degree or foreign equivalent in Computer Science, Software Engineering, Electrical Engineering, Electronics Engineering, Computer Engineering, Computer Networking, Information Technology, Information Systems, or Telecommunications Engineering
  • Two (2) years of experience in the job offered or closely related engineering role
  • Two (2) years of experience in Production support including interpreting system logs and collecting diagnostic data
  • Two (2) years of experience in root cause analysis
  • Two (2) years of experience in reviewing and implementing security vulnerability fixes
  • Two (2) years of experience in CI/CD pipe-lines
  • Two (2) years of experience in Agile
  • Two (2) years of experience in IaC
  • Two (2) years of experience in Ansible
  • Two (2) years of experience in Terraform
  • Two (2) years of experience in Salt
  • Two (2) years of experience in Python
  • Two (2) years of experience in Bash
  • Two (2) years of experience in GitLab
  • Two (2) years of experience in Linux
  • Two (2) years of experience in Jira
  • Two (2) years of experience in Confluence

Responsibilities

  • Perform site reliability engineering duties in furtherance of ensuring application and network performance, reliability and security for users of our platforms
  • Deploy and operate scalable, highly available, and faulttolerant systems
  • Analyze and improve the security, stability, speed, and capacity of our network
  • Automate new cloud service deployments
  • Monitor and troubleshoot cloud services to ensure SLA is met and adequate capacity is available
  • Analyze logs and events from the cloud service and providing recommendations to ensure smooth service operation
  • Work within large-scale cloud-based infrastructure
  • Quickly identify and resolve infrastructure issues

Benefits

  • Healthcare
  • 401K savings plan
  • Company holidays
  • Vacation (in the form of PTO)
  • Sick time
  • Family friendly benefits including parental leave
  • Employee assistance program including a focus on mental and financial wellness
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service