Production Support Engineer

Shree Narayani Networking SolutionsChandler, AZ
Hybrid

About The Position

This is a live incident production support seat, not a pure runbook follow role. The team monitors batch, feeds, and application health across Unix, SQL, Dynatrace, Splunk, and ServiceNow, and owns incidents end to end including root cause. Candidates need to be comfortable being handed a vague scenario, for example an application down two hours after a change went in, and walking an interviewer through live triage, not reciting a memorized process.

Requirements

  • 3-5+ years of hands-on application production support experience
  • Strong Autosys skills
  • Strong Unix and Shell scripting skills
  • Working SQL, Oracle, Hadoop or similar DBMS knowledge
  • Ability to juggle and prioritize multiple concurrent issues
  • Fast independent learner
  • Command level Unix fluency (specific syntax for commands like du, df, top, uptime, find)
  • Ability to distinguish failure types precisely (data issue vs. Unix file system issue vs. database connection pool issue)
  • SQL depth beyond basic querying (index behavior, query performance, truncate vs. delete, join types)
  • Autosys job state knowledge (inactive, on hold, on ice)
  • Experience using Splunk and Dynatrace together
  • Experience with noisy alert management and tuning alert thresholds
  • Fluency in incident severity levels (P1 through P4) and communication protocols for each
  • Basic API and auth troubleshooting exposure (JWT or token related login failures)
  • Real-world experience with shell scripting for automation (e.g., disk cleanup, log rotation)
  • Ability to confidently explain actual production support scenarios and demonstrate troubleshooting steps during technical interviews

Nice To Haves

  • Experience with ServiceNow

Responsibilities

  • Live incident production support
  • Monitor batch, feeds, and application health across Unix, SQL, Dynatrace, Splunk, and ServiceNow
  • Own incidents end to end including root cause analysis
  • Perform live triage of production issues
  • Troubleshoot and resolve production incidents
  • Manage and tune alerts
  • Automate repetitive tasks using shell scripting
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service