Sr. Site Reliability Engineer

Blackpoint Cyber
CA$131,000 - CA$164,250

About The Position

Blackpoint Cyber is the leading provider of world-class cybersecurity threat hunting, detection and remediation technology. Founded by former National Security Agency (NSA) cyber operations experts who applied their learnings to bring national security-grade technology solutions to commercial customers around the world, Blackpoint Cyber is in hyper-growth mode, fueled by a recent $190m series C round. We're hiring a Senior Site Reliability Engineer to design, implement, and maintain our cloud and on-premise infrastructure and CI/CD pipelines, with a focus on automation, scalability, and performance. You'll work across cloud platform administration, container orchestration, data streaming, observability, and incident response — partnering with engineering teams to keep our systems reliable, secure, and efficient, and helping foster a culture of continuous improvement.

Requirements

  • 5+ years of experience in a Senior Site Reliability Engineer role or equivalent, with substantial emphasis on cloud infrastructure management and automation.
  • Expertise in Infrastructure as Code (Terraform, Terragrunt) for enterprise-scale deployments.
  • Comprehensive knowledge of AWS, including designing, implementing, and maintaining secure, scalable, resilient cloud architectures.
  • Extensive hands-on experience with distributed data streaming (Confluent Cloud, Apache Kafka).
  • Proven experience with Redis for caching and Amazon RDS for relational database management.
  • Experience with enterprise search and analytics platforms (OpenSearch, Elasticsearch, ChaosSearch).
  • Proficiency designing and implementing monitoring/alerting infrastructure (Prometheus, Grafana, Alert Manager, OpsGenie/PagerDuty).
  • Practical experience with feature flag systems (LaunchDarkly/PostHog) for controlled release management.
  • Extensive experience administering production-grade Kubernetes (Helm, ArgoCD, Istio); working knowledge of Kustomize.
  • Strong problem-solving skills, with the ability to troubleshoot complex systems in production.
  • Strong communication and collaboration skills, with experience working in Agile environments.

Nice To Haves

  • Multi-cloud experience (Google Cloud Platform, Microsoft Azure).
  • Understanding of security frameworks and compliance standards for cloud-native/containerized environments.
  • Serverless computing and CI/CD pipeline experience (Jenkins, GitHub Actions).
  • Software development proficiency in Node.js, Python, and/or Go.

Responsibilities

  • Design, develop, and maintain highly scalable infrastructure using Infrastructure as Code (Terraform and Terragrunt) for automated cloud resource provisioning and orchestration.
  • Own and optimize our AWS cloud environment, ensuring cost efficiency, security best practices, and high-availability standards.
  • Manage and optimize Kubernetes cluster environments (Helm, ArgoCD, Istio, Kustomize) to support continuous delivery and infrastructure-as-code practices.
  • Administer and scale data streaming infrastructure (Confluent Cloud, Apache Kafka) to support enterprise-level data processing.
  • Deploy, configure, and maintain Redis for caching and real-time data processing.
  • Implement and maintain monitoring, alerting, and incident response frameworks (Prometheus, Grafana, Alert Manager, OpsGenie/PagerDuty) to ensure system reliability and performance.
  • Facilitate controlled feature deployments and progressive rollouts through LaunchDarkly/PostHog.
  • Partner with software development teams to ensure seamless integration of new services, applications, and features into existing infrastructure.
  • Diagnose and resolve complex system-level issues, implementing solutions that maintain high performance and maximize uptime.
  • Drive continuous improvement of automation tooling, operational processes, and engineering methodologies to enhance scalability, reliability, and maintainability.
  • Stay current on emerging SRE trends and tools, and help the team adopt relevant industry advancements and best practices.

Benefits

  • competitive Health, Vision, Dental, and Life Insurance plans
  • a robust 401k plan
  • Discretionary Time Off
  • equity participation
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service