About The Position

In this role, you will use your technical expertise and operational data to drive continuous improvement and an exceptional customer experience. You will ensure that engineering teams align their requirements, decisions, and quality metrics with a deep understanding of what creates the best possible customer experience. This is a high-visibility, technical leadership role with zero direct reports, requiring a blend of software engineering credibility, DevOps/SRE domain expertise, and organizational change-management skills. As the Staff Software Engineer, Reliability Engineering & SDLC Governance, you will establish accountability for software quality across the Access Edge organization and drive continuous quality improvement across our entire software delivery lifecycle. You will establish holistic ownership of our SDLC governance while keeping your primary focus on the critical interface between software development and live operations. You will drive increasing reliability as code moves from conception to production. You will architect operational feedback loops, observability strategies, and shift-left testing paradigms that ensure our full-stack, complex access edge hardware/software/networking environment remains highly reliable. This is a site‑based role, employees work 3+ days (60%+) per week from a Viasat office or work location within a standard five‑day workweek.

Requirements

  • 8+ years of experience software engineering —ideally operating at a Senior or Staff level.
  • Proven experience as a Software Engineer.
  • Must possess the technical depth to understand and optimize complex delivery lifecycles.
  • Demonstrated ability to anchor engineering initiatives in user advocacy, ensuring technical requirements are continuously mapped to user experience goals.
  • Demonstrated experience managing or establishing comprehensive SDLC governance frameworks across an engineering organization, with a strong emphasis on continuous deployment and automated release gates.
  • Experience working within complex, full-stack environments, preferably dealing with networking architectures, access edge technologies, or highly distributed systems.
  • Demonstrated track record of driving large-scale organizational change and holding engineering teams accountable to operational, quality, and continuous improvement standards.

Nice To Haves

  • Strong background in Site Reliability Engineering (SRE) and DevOps principles (SLIs/SLOs mapped to customer journeys, error budgets, blameless post-mortems).
  • Proven track record of designing and scaling continuous improvement programs across multi-team ecosystems to reduce technical debt and systemic fragility.
  • Experience implementing or steering AI-augmented software development and testing tools.
  • Advanced proficiency in designing observability and monitoring frameworks (e.g., Prometheus, Grafana, OpenTelemetry, Datadog).
  • Experience with building scaled agile processes leveraging AS9115 or similar frameworks.

Responsibilities

  • Establish, govern, and enforce a rigorous operational feedback loop.
  • Own the end-to-end post mortem and Incident Review process, ensuring root-cause analyses are completed on time and actionable remediation items are prioritized and completed.
  • Represent Access Edge in engineering-wide quality initiatives; work closely with other parts of the company, e.g. on postmortems, failure mode analysis, disaster recovery, and security.
  • Ensure engineering teams maintain an unwavering empathy for the customer through ensuring an excellent quality of service.
  • Guide teams to translate engineering decisions, feature requirements, and technical tradeoffs into their direct impact on the end-user experience.
  • Analyze post-mortem trends and live operational data to identify systemic process and technical execution weaknesses.
  • Lead post-mortem follow-ups to ensure engineering teams are actively iterating, learning, and upgrading their workflows based on real-world customer impacts.
  • Define, track, and report on core organizational quality and reliability metrics (e.g., MTTR, MTTD, defect escape rates, change failure rates) across the full lifecycle, ensuring these metrics directly correlate to improved customer satisfaction and system trust.
  • Define, institutionalize, and oversee the standards for the software development lifecycle (SDLC), ensuring that quality checkpoints, risk reviews, and reliability frameworks are embedded across all team deployments.
  • Partner with engineering leads to promote shift-left testing practices into the early stages of daily development workflows, moving quality verification closer to the point of code creation.
  • Advise and steer teams on implementing robust observability, telemetry, and monitoring frameworks to detect, isolate, and mitigate anomalies at the access edge before they degrade the customer experience.
  • Actively champion and accelerate the adoption of AI-augmented tools across the SDLC to level up code quality, test generation, and automated refactoring.
  • Influence and align multiple distributed engineering teams around shared quality, lifecycle, and operational maturity goals without having direct authority over them.
  • Act as the connective tissue between front-line operations, platform engineering, and product development to build a unified, customer-first culture of reliability.

Benefits

  • range of medical, financial, and/or other benefits, dependent on the position offered
  • comprehensive benefit offerings that are focused on your holistic health and wellness
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service