This role requires a strong background in Site Reliability Engineering (SRE) or Production Engineering, with hands-on expertise in managing large-scale distributed messaging systems. The ideal candidate will have a deep understanding of system reliability, scalability, and high availability design, as well as messaging reliability patterns. Experience with observability tools, incident management, and scripting/automation is crucial. The role also involves managing Linux/Unix and Windows production environments and understanding event-driven architectures, messaging platform security, and vulnerability remediation.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed