Platform Software Engineer

OracleUnited States,
$92,500 - $209,500

About The Position

We’re hiring a Platform Software Engineer to help build the execution layer behind OCI’s AI and GPU growth motion. You’ll work across Oracle data platforms, GPU scheduling, orchestration pipelines, and MLOps agentic tooling — turning customer demand into production-grade systems that scale across our cluster fleet. This is a builder role on a small, high-leverage team. Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing [email protected] or by calling 1-888-404-2494 in the United States. Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Requirements

  • Builder role on a small, high-leverage team
  • Work across Oracle data platforms, GPU scheduling, orchestration pipelines, and MLOps agentic tooling
  • Turn customer demand into production-grade systems that scale across our cluster fleet
  • Integrate with Oracle data platforms (OCI Streaming, 23ai, Object Storage, Data Flow)
  • Move training and inference data reliably across customer and internal pipelines
  • Build and extend GPU schedulers and capacity-aware placement logic for A100, H100, H200, and Blackwell fleets
  • Develop orchestration pipelines for training, fine-tuning, and inference workloads using Kubernetes, Argo, and Slurm where appropriate
  • Ship MLOps agentic tooling — observability, automated triage, cost and SLO agents
  • Reduce operator load on large GPU deployments
  • Partner with PMs, SAs, and customer-facing teams
  • Convert field requirements into reusable components, not one-off scripts

Responsibilities

  • Integrate with Oracle data platforms (OCI Streaming, 23ai, Object Storage, Data Flow) to move training and inference data reliably across customer and internal pipelines.
  • Build and extend GPU schedulers and capacity-aware placement logic for A100, H100, H200, and Blackwell fleets.
  • Develop orchestration pipelines for training, fine-tuning, and inference workloads using Kubernetes, Argo, and Slurm where appropriate.
  • Ship MLOps agentic tooling — observability, automated triage, cost and SLO agents — that reduces operator load on large GPU deployments.
  • Partner with PMs, SAs, and customer-facing teams to convert field requirements into reusable components, not one-off scripts

Benefits

  • flexible medical
  • life insurance
  • retirement options
  • volunteer programs
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service