Monitoring/SRE Engineer (REMOTE)

Koniag Government Services, LLCWashington, DC
$120,000 - $133,000Remote

About The Position

Koniag IT Systems, LLC, a Koniag Government Services company, is seeking a Monitoring / SRE Engineer to support KITS and our government customer in Washington, DC. The position is remote. This position requires the candidate to be able to obtain a Public Trust. We offer competitive compensation and an extraordinary benefits package including health, dental and vision insurance, 401K with company matching, flexible spending accounts, paid holidays, three weeks paid time off, and more. Koniag IT Systems (KITS) is seeking a Monitoring / SRE Engineer with a minimum of 5 years of experience to support enterprise monitoring, alerting, and incident response for a federal civilian customer's hybrid infrastructure environment. This is a strong opportunity for an operations-minded engineer to build site reliability engineering skills while supporting a large-scale infrastructure modernization program.

Requirements

  • Associate's or Bachelor's degree in Information Technology or related field, or equivalent professional experience
  • Minimum of 5 years of experience in IT operations, monitoring, or a related support role
  • Familiarity with enterprise monitoring/observability tools (Splunk, SolarWinds, Grafana, Datadog, or similar)
  • Basic understanding of cloud infrastructure concepts (AWS preferred)
  • Strong attention to detail and ability to follow incident response procedures
  • Understanding of ITSM ticketing and escalation processes
  • Ability to obtain Public Trust clearance
  • Ability to work collaboratively in a fast-paced environment
  • Excellent communication skills and the ability to convey complex technical concepts to non-technical stakeholders
  • Ability to obtain public trust clearance

Nice To Haves

  • Basic scripting ability (Python, PowerShell, or Bash) is a plus
  • Experience working in a federal government IT environment
  • CompTIA Network+, Splunk Core Certified User, or AWS Certified Cloud Practitioner
  • Exposure to ServiceNow or similar ITSM/on-call platforms
  • Prior experience on a federal government IT support contract
  • Interest in growing toward a Site Reliability Engineering (SRE) career path

Responsibilities

  • Monitor enterprise infrastructure, application, and cloud dashboards (e.g., Splunk, SolarWinds, Grafana, or CloudWatch) to identify and triage issues.
  • Respond to system alerts, perform initial troubleshooting, and escalate incidents per established runbooks and ServiceNow processes.
  • Assist in building and maintaining dashboards, alert rules, and automated notifications for infrastructure and application health.
  • Support root cause analysis and after-action documentation for production incidents.
  • Assist with synthetic monitoring and validation checks supporting the customer's business disaster continuity and recovery (BDCR) testing.
  • Maintain accurate incident tickets, monitoring runbooks, and knowledge base articles.
  • Collaborate with server, cloud, storage, and database engineers to ensure systems are properly instrumented and monitored.
  • Participate in an on-call rotation to support after-hours incident response.

Benefits

  • health, dental and vision insurance
  • 401K with company matching
  • flexible spending accounts
  • paid holidays
  • three weeks paid time off
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service