Senior DevOps Engineer

Munich Re CareersNew York, NY
$147,000 - $170,000

About The Position

The Senior DevOps Engineer will design, build, and operate a secure and reliable internal platform for running business applications and machine learning workloads at scale. This role involves developing and maintaining self-service capabilities for infrastructure provisioning, workload deployment, and application lifecycle management through standardized APIs, templates, and automation. The engineer will build and evolve reusable platform primitives such as K8s clusters, ingress, service networking, secrets management, policy controls, and identity integration. They will maintain cloud and platform infrastructure using Infrastructure as Code (e.g., Terraform, Pulumi, Crossplane) in a secure, scalable, and reusable manner, including modular design, versioning, policy guardrails, automated validation, and safe rollout practices. Additionally, the role includes implementing and maintaining platform-level CI/CD patterns that support multiple teams while enforcing secure and compliant SDLC practices, providing opinionated reference architectures and reusable building blocks, and implementing guardrails for security and compliance while maintaining developer velocity. The engineer will own platform observability by establishing monitoring, logging, tracing, alerting, and SLO practices, and drive platform reliability and operational excellence through incident response, root cause analysis, postmortems, and continuous improvement to reduce toil. Support for daily operations, monitoring, and security functions of the platform stack is also required. The role involves partnering closely with application teams and stakeholders to understand friction points, prioritize the platform roadmap, and deliver measurable improvements in developer experience and time-to-production. Leadership through influence and consensus, providing technical guidance, reviews, and mentorship to peers and junior engineers is expected. The engineer will also own their professional development, continuously learning new technologies and domain context to operate as a subject matter expert (SME).

Requirements

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Hands-on experience designing and operating platform services using Infrastructure as Code (Terraform, Pulumi, Crossplane, or CSP-native tooling) with a modular, reusable approach.
  • Strong experience with K8s or OpenShift and core platform components, including networking, ingress, service discovery, storage, RBAC, admission control, and multi-tenancy.
  • Experience implementing and standardizing CI/CD for multi-team environments, ideally with GitHub Actions and Argo CD, including release strategies and deployment automation.
  • Experience with at least one major Cloud Service Provider such as Azure, GCP, or AWS.
  • Strong networking fundamentals, including DNS, TLS, load balancing, routing, and private connectivity.
  • Strong security foundations across identity, application, data, network, and supply-chain security, including scanning, signing, SBOM, secrets handling, and policy enforcement.
  • Strong software engineering fundamentals and experience building internal tooling and automation, ideally in Go, Python, or TypeScript.
  • Experience with containers and container tooling, including Docker, registries, and image build pipelines.
  • Excellent knowledge of Linux operating systems and operational troubleshooting.
  • Proven experience leading incident response, including triage, mitigation, coordination across teams, and driving post-incident improvements.
  • Deep knowledge of observability tooling and practices, such as Prometheus, Grafana, and Datadog.
  • Strong problem-solving and analytical abilities, with excellent communication and cross-functional collaboration skills.

Responsibilities

  • Design, build, and operate a secure and reliable internal platform for running business applications and machine learning workloads at scale.
  • Develop and maintain self-service capabilities (golden paths) for provisioning infrastructure, deploying workloads, and managing the application lifecycle through standardized APIs, templates, and automation.
  • Build and evolve reusable platform primitives such as K8s clusters, ingress, service networking, secrets management, policy controls, and identity integration.
  • Maintain cloud and platform infrastructure using Infrastructure as Code (e.g., Terraform, Pulumi, Crossplane) in a secure, scalable, and reusable manner, including modular design, versioning, policy guardrails, automated validation, and safe rollout practices.
  • Implement and maintain platform-level CI/CD patterns that support multiple teams while enforcing secure and compliant SDLC practices.
  • Provide opinionated reference architectures and reusable building blocks (e.g., Helm charts, Argo CD apps, Terraform modules, scaffolding tools) that enable consistent delivery.
  • Implement guardrails for security and compliance (policy-as-code, least privilege, workload identity, image provenance, runtime controls) while maintaining developer velocity.
  • Own platform observability by establishing monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable.
  • Drive platform reliability and operational excellence through incident response, root cause analysis, postmortems, and continuous improvement to reduce toil.
  • Support daily operations, monitoring, and security functions of the platform stack, including routine maintenance, access and identity workflows, vulnerability remediation, and operational support for internal platform services.
  • Partner closely with application teams and stakeholders to understand friction points, prioritize the platform roadmap, and deliver measurable improvements in developer experience and time-to-production.
  • Lead through influence and consensus, providing technical guidance, reviews, and mentorship to peers and junior engineers.
  • Own your professional development, continuously learning new technologies and domain context to operate as a subject matter expert (SME).

Benefits

  • Two options for your health insurance plan (PPO or High Deductible).
  • Prescription drug coverage (included in your health insurance plan).
  • Vision and dental insurance plans.
  • Basic life insurance equal to 1x annual salary and AD&D coverage that is equal to 1x annual salary.
  • Short and Long-Term Disability coverage.
  • Supplemental Life and AD&D plans that you can purchase for yourself and dependents (includes Spouse/domestic partner and children).
  • Voluntary Benefit plans that supplement your health and life insurance plans (Accident, Critical Illness and Hospital Indemnity).
  • A robust 401k plan with up to a 6% employer match
  • Paid time off that begins with 24 days each year, with more days added when you celebrate milestone service anniversaries.
  • Eligibility to receive a yearly bonus as a Munich Re employee.
  • A variety of health and wellness programs provided at no cost, including a gym fitness reimbursement.
  • Paid time off for eligible family care needs.
  • Tuition assistance and educational achievement bonuses.
  • A corporate matching gifts program that further enhances your charitable donation.
  • Paid time off to volunteer in your community.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service