Site Reliability Developer 4

OracleSanta Clara, CA
Remote

About The Position

Oracle America, Inc. is seeking a Site Reliability Developer 4 to solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence. The role involves designing, writing, and deploying software to improve the availability, scalability, and efficiency of Oracle products and services. Additionally, the position requires designing and developing designs, architectures, standards, and methods for large-scale distributed systems, as well as facilitating service capacity planning and demand forecasting, software performance analysis, and system tuning. This position may telecommute.

Requirements

  • Linux/Unix systems and system internals
  • Networking basics (TCP/IP, DNS, HTTP/HTTPS, load balancing, and routing)
  • Scripting (Python and Bash) and programming (Go and Java)
  • Cloud platforms (OCI, AWS, GCP, and Azure) and compute, networking, storage, and IAM
  • Infrastructure as Code: Terraform, Ansible, CloudFormation, and OCI Resource Manager
  • Containerization and orchestration: Docker, Kubernetes, Helm, and operators
  • Reliability engineering
  • Monitoring and metrics (Prometheus, Grafana, and OCI Monitoring)
  • Logging and tracing (ELK/EFK, OpenTelemetry, Jaeger/Zipkin, and OCI Logging)
  • Compliance awareness (SOC 2 and ISO 27001), change management, and audit trails
  • Databases (SQL and NoSQL), replication, backups and restoration, and PITR
  • Caching (Redis and Memcached) and message queues (Kafka and RabbitMQ)
  • Knowledge of server hardware and software configuration
  • Knowledge of networking
  • Knowledge of standard internet services
  • Knowledge of scripting languages
  • Knowledge of cloud computing patterns
  • Knowledge of technology security and compliance
  • Experience running large scale customer facing web services
  • Understanding of load balancing technologies
  • Experience with development in programming languages, databases and big data stores, and container technologies
  • Defining and documenting technical architecture of complex and highly scalable products

Responsibilities

  • Solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence.
  • Design, write, and deploy software to improve the availability, scalability, and efficiency of Oracle products and services.
  • Design and develop designs, architectures, standards, and methods for large-scale distributed systems.
  • Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuning.

Benefits

  • Flexible medical
  • Life insurance
  • Retirement options
  • Volunteer programs
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service