We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles: autoscaling App Service and Container Apps workloads, VM Scale Sets, and Kafka-based event streaming form the backbone. We also have a smaller, well-run presence on AWS, and a Windows-based on-premises component that sits at the core of our product and integrates with our cloud environment. This role is primarily focused on Azure, where the biggest opportunity for impact lives, and you'll work across the full multi-cloud picture. We'd like to see the on-prem and cloud sides operate as a more unified, well-integrated system than they do today, and that integration work is part of what makes this role interesting. This is a high-impact role for someone who thinks in terms of systems rather than tickets, and who treats infrastructure like software. You'll have significant ownership over how our cloud infrastructure is architected, provisioned, secured, and operated going forward. If you get energized by taking a fast-growing environment and giving it real architectural rigor (consistent patterns, full infrastructure-as-code coverage, sane permission models, and cost discipline), this role was built for you. We're looking for someone who wants to build the operating model, establish the standards, and lead the transformation. You should bring genuine software engineering discipline to how that infrastructure work gets done: everything in git, everything reviewed, everything automated. No snowflakes, no manual changes made "just this once."
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed