Sr Site Reliability Engineer, Platform

Blue River Technology
•$148,000 - $261,000•Remote

About The Position

Blue River Technology is looking for a Senior Site Reliability Engineer to join its Platform organization, which is dedicated to accelerating the company-wide adoption and scaling of automation and robotics. The Platform's product is a set of API services and infrastructure designed to overcome scaling hurdles, such as operational complexity and system exceptions, thereby enabling the rapid launch and scaling of new autonomy innovations and products at Blue River. In this role, you will drive architectural decisions, mentor engineers across teams, and shape the direction of our platform.

Requirements

  • Minimum of six years of experience in building and maintaining infrastructure for data-intensive, high-availability applications, including building and maintaining public cloud solutions
  • Use of cloud orchestration tools such as Kubernetes and Terraform in a production environment
  • Understanding of software design methodologies, information systems architecture, object-oriented design, and software design patterns
  • Experience securing cloud infrastructure (preferably AWS and Kubernetes) in a production environment
  • Experience in one or more of the following languages: Golang (preferred), Python, JavaScript, Rust
  • Production experience with CI/CD tooling (GitHub Actions, ArgoCD, ArgoCD Image Updater, Artifactory)

Nice To Haves

  • Are interested in robotic applications, and developing software that assists robots
  • Want to join a fun, fast-moving engineering team
  • Want to build a platform that dozens of autonomous product teams depend on daily
  • Are a self-starter with infectious enthusiasm, energy and problem-solving abilities

Responsibilities

  • Architect, scale, and take ownership of essential infrastructure
  • Build and maintain a Kubernetes based platform supporting multiple teams and services
  • Build backend services and internal tooling (Golang) to support autonomous systems.
  • Work with product teams for launch of new products on the platform
  • Grow our high availability infrastructure while maintaining key metrics such as uptime
  • Build tooling to support our platform and development teams
  • Perform end-to-end performance analysis, identify areas for improvement and implement with robust solutions.
  • Work with cloud vendors and external technical support for upgrades and rapid resolution of problems
  • Participate in on-call rotation, triaging and resolving production incidents with thorough root cause analysis and postmortem documentation
  • Design and maintain observability infrastructure — dashboards, alerts, and log aggregation — to ensure visibility into platform health and service performance
  • In collaboration with the security team, conduct regular risk assessments
  • Maintain risk register, develop and implement mitigation plans
  • Assess intrusion detection alerts. Improve systems and services that digest threat feeds
  • In collaboration with IT and purchasing teams, ensure payment process of each SaaS service is established and maintained

Benefits

  • annual performance bonus
  • competitive benefit package
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service