Principal Software Engineer, Core Infrastructure

OracleNashville, TN
$114,600 - $234,600

About The Position

Lead the development of scalable, elastic, and highly available distributed systems for hyperscale cloud environments. Architect critical components and drive performance, reliability, operational excellence, security, and automation. Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. With AI embedded across its products and services, Oracle helps customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

Requirements

  • Experience leading the development of scalable, elastic, and highly available distributed systems for hyperscale cloud environments.
  • Experience architecting critical components.
  • Experience driving performance, reliability, operational excellence, security, and automation.
  • Experience designing and developing scalable, elastic distributed systems supporting horizontal and vertical scaling.
  • Experience defining scalability requirements and optimizing code and data paths for high-throughput, hyperscale workloads.
  • Experience leveraging distributed state management and data-plane platforms for large-scale retrieval, storage, and processing.
  • Experience building fault-tolerant, highly available systems using redundancy, replication, failover, and appropriate consistency and availability tradeoffs.
  • Experience implementing load shedding, throttling, and rate limiting while meeting defined SLOs.
  • Experience defining KPIs, telemetry, dashboards, and alerts to proactively monitor system health and performance.
  • Experience designing performance, load, fault-injection, and brownout testing to validate scalability, correctness, and resilience.
  • Experience implementing replication and synchronization mechanisms to ensure data integrity, durability, and availability.
  • Experience proactively diagnosing and resolving complex production issues and ensuring operational readiness.
  • Experience supporting incident response, root cause investigations, and operational support rotations.
  • Experience designing systems for in-service maintenance and upgrades with minimal customer impact.
  • Experience mentoring engineers on troubleshooting and operational best practices.
  • Experience implementing robust security controls for multi-tenant cloud environments, including encryption and access controls.
  • Experience remediating security gaps and ensuring compliance with applicable standards and requirements.
  • Experience maintaining accurate security and compliance documentation.
  • Experience developing Infrastructure as Code (IaC) and automation for managing cloud infrastructure.
  • Experience enabling safe, repeatable patching, upgrades, deployments, and rollbacks through effective change-management practices.

Responsibilities

  • Design and develop scalable, elastic distributed systems supporting horizontal and vertical scaling.
  • Define scalability requirements and optimize code and data paths for high-throughput, hyperscale workloads.
  • Leverage distributed state management and data-plane platforms for large-scale retrieval, storage, and processing.
  • Build fault-tolerant, highly available systems using redundancy, replication, failover, and appropriate consistency and availability tradeoffs.
  • Implement load shedding, throttling, and rate limiting while meeting defined SLOs.
  • Define KPIs, telemetry, dashboards, and alerts to proactively monitor system health and performance.
  • Design performance, load, fault-injection, and brownout testing to validate scalability, correctness, and resilience.
  • Implement replication and synchronization mechanisms to ensure data integrity, durability, and availability.
  • Proactively diagnose and resolve complex production issues and ensure operational readiness.
  • Support incident response, root cause investigations, and operational support rotations.
  • Design systems for in-service maintenance and upgrades with minimal customer impact.
  • Mentor engineers on troubleshooting and operational best practices.
  • Implement robust security controls for multi-tenant cloud environments, including encryption and access controls.
  • Remediate security gaps and ensure compliance with applicable standards and requirements.
  • Maintain accurate security and compliance documentation.
  • Develop Infrastructure as Code (IaC) and automation for managing cloud infrastructure.
  • Enable safe, repeatable patching, upgrades, deployments, and rollbacks through effective change-management practices.

Benefits

  • Flexible medical
  • Life insurance
  • Retirement options
  • Volunteer programs
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service