Linux Admin

WiproRedmond, WA
Remote

About The Position

The purpose of this role is to assist in the triage and documentation of priority incidents, support resolution of software and hardware service requests and ensure SLA adherence by escalating delays when required. The role involves working closely with the development team on maintaining the operational health of core compute services for API availability and low latency, managing and triaging tickets, driving prioritization and execution of work based on impact, and scaling systems sustainably through mechanisms such as easy-to-use tooling and automation. The individual will work in concert with service developers to evolve systems/products for better scalability, reliability, and development velocity, drive new runbooks to help reduce mean triage time of incidents, prioritize and automate high hit count runbooks, and practice sustainable incident response and drive root cause analysis.

Requirements

  • Understanding of Linux operating systems and Linux system administration (SysAdmin roles).
  • Good understanding of Linux/Unix commands (Strong User).
  • Experience automating tasks with scripting languages such as Python, Bash, and JavaScript.
  • Systematic problem-solving approach, strong communication skills, a sense of ownership and drive solutions.
  • Deep understand of service metrics and alarms through the development of dashboards, service KPIs, alarming systems.
  • Experience working in an operational environment with mission critical tier one services with associated pager duty.
  • Software development experience focused on Services and Operational tools.
  • CI/CD process and software knowledge.
  • Operational toil mitigation and reduction.
  • Linux Admin

Responsibilities

  • Perform initial triage and assist in response for priority incident tickets (P3 tickets) along with its documentation and following the defined SLA and escalation procedures.
  • Track and assist in resolution of software/hardware/network service requests, ensuring they are accurately logged, processed, and completed within defined timelines as per established procedures.
  • Track and monitor SLA timelines for entire lifecycle of priority incident tickets (P3 tickets). Escalate to higher levels for delays or potential SLA breaches.
  • Assist in creation of change requests basis types of incident tickets which are raised and resolved.
  • Work closely with development team on maintaining operational health of core compute services for API availability and low latency.
  • Managing and triaging tickets.
  • Driving prioritization and execution of work based on impact.
  • Scale systems sustainably through mechanisms such as easy-to-use tooling and automation.
  • Work in concert with service developers to evolve systems/products for better scalability, reliability and development velocity.
  • Drives new runbooks to help reduce mean triage time of incidents.
  • Prioritize and automate high hit count runbooks.
  • Practice sustainable incident response and drive root cause analysis.

Benefits

  • Full range of medical and dental benefits options
  • Disability insurance
  • Paid time off (inclusive of sick leave)
  • Other paid and unpaid leave options
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service