About The Position

Medallia is seeking a hands-on Senior Platform Software Engineer to join their Platform Services team. This role involves building, operating, and continuously improving shared platforms that support the engineering organization. The team manages a large number of production and non-production instances across various technologies, ensuring reliable, scalable, and secure services for multiple engineering teams. The ideal candidate will have strong fundamentals in platform and distributed systems, experience operating complex systems in production, and a curiosity to learn diverse technologies.

Requirements

  • 5+ years of experience in software, systems, platform, DevOps, SRE, or related engineering roles.
  • 3+ years experience operating or troubleshooting distributed systems in production.
  • Experience building or operating distributed systems, high-availability architectures, data replication, and fault-tolerant platforms.
  • Hands-on experience with one or more distributed platforms such as Kafka, Redis, Elasticsearch, Spark, Airflow, or similar technologies.
  • Ability and willingness to learn and operate across a broad technology portfolio.

Nice To Haves

  • Demonstrated experience with debugging, incident response, and performance-tuning.
  • Experience with automation or software development using Java, Go, Python, or similar.
  • Experience managing and building services on cloud infrastructure and Kubernetes.
  • Experience collaborating with multiple engineering teams on architecture and production readiness.
  • Degree in Computer Science, Engineering, or a related field.

Responsibilities

  • Own and operate highly available distributed platforms, ensuring reliability, scalability, performance, security, and operational readiness.
  • Troubleshoot complex production issues, lead incident response and root-cause analysis, and drive systemic improvements.
  • Design and validate failover, disaster recovery, backup/restore, upgrade, and capacity strategies.
  • Build automation and tooling to reduce operational toil and improve provisioning, deployments, upgrades, and recovery.
  • Develop observability through metrics, logging, dashboards, alerts, and actionable operational signals.
  • Review architectures and production-readiness of services consuming our platforms.
  • Partner with engineering teams to establish reliable and scalable platform patterns.
  • Create technical documentation, runbooks, and knowledge-sharing practices that improve the team's operational maturity.
  • Participate in a periodic on-call rotation supporting 24/7 reliability.

Benefits

  • Competitive health and wellness benefits, including medical, dental, vision.
  • 401(k).
  • Short-term and long-term disability.
  • Life and AD&D insurance.
  • Statutory leaves.
  • Paid parental leave.
  • Paid holidays.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service