AI Platform Engineer

Axos BankSan Diego, CA
Onsite

About The Position

Axos Bank is hiring an AI Platform Engineer to engineer, operate, and help govern the platforms and infrastructure that power the bank’s AI and automation capabilities. This is a hands-on, senior individual-contributor role spanning platform engineering, the supporting data and compute tier, reliability, and platform governance — applied across a growing portfolio of enterprise AI and automation platforms rather than any single product. The successful candidate is platform- and infrastructure-focused: equally comfortable deploying and hardening an enterprise platform, tuning a PostgreSQL cluster or Redis tier, and establishing the engineering standards and controls the platform estate runs under. The role ensures that the bank’s AI and automation platforms are reliable, performant, secure, and well-governed as adoption scales across lines of business. This is the hands-on engineering seat behind our AI and automation platforms, the person who deploys, hardens, scales, and governs the platforms and the infrastructure they run on. The role offers deep technical ownership of a fast-growing platform footprint, close partnership with senior architects in Infrastructure, Security, and Identity, and a central part in scaling the bank’s Automation Center of Excellence.

Requirements

  • 5+ years of platform or infrastructure engineering experience in an enterprise environment, with a strong hands-on (not primarily managerial) track record
  • Hands-on PostgreSQL administration — replication, high availability, backup/restore, point-in-time recovery, connection pooling, and performance/query tuning
  • Hands-on Redis operational experience — Sentinel or Cluster, persistence configuration, eviction policy, and queue/pub-sub workloads
  • Strong Linux administration plus Docker and containerized application operations (Kubernetes a plus)
  • Experience deploying, operating, and hardening enterprise platforms — such as workflow automation, integration/iPaaS, data, or other application platforms — across multiple environments
  • Strong scripting/coding proficiency in Python or JavaScript/TypeScript
  • Experience integrating platforms with Entra ID (or Azure AD) for SSO/OIDC, RBAC, and group-based access
  • Experience with CI/CD and infrastructure-as-code (Terraform or equivalent)
  • Experience with platform governance, security operations, and change management in a regulated environment (financial services, healthcare, or similar), including audit support
  • Strong written communication for runbooks, design documents, and standard operating procedures

Responsibilities

  • Deploy, configure, upgrade, and operate the bank’s portfolio of enterprise AI and automation platforms across Dev/QA/UAT/Prod — including high-availability topology, scaling, version and patch management, and full platform lifecycle
  • Tune platforms for performance, capacity, and cost as adoption grows; plan and execute upgrades and migrations with minimal service impact
  • Integrate platforms with enterprise identity (Entra ID/OIDC), secrets management, and source control; manage platform-level RBAC and credential lifecycle
  • Evaluate and onboard new AI and automation platform capabilities into the supported estate
  • Engineer and operate the PostgreSQL tier supporting the platform estate — replication, high availability and automated failover, point-in-time recovery, connection pooling, and performance/query tuning
  • Engineer and operate the Redis tier — persistence configuration, memory and eviction policy, queue durability, and high availability (Sentinel or Cluster)
  • Engineer the Linux hosts and Docker/Kubernetes containers running the platforms — hardening, patching, capacity, and performance tuning, in partnership with Infrastructure
  • Design, implement, and validate HA/DR for the platform data tier against defined RTO/RPO — failover behavior, backup validation, and recurring recovery exercises
  • Instrument the platforms and their infrastructure for service health, replication/failover events, performance, and capacity using the enterprise observability stack
  • Build and maintain dashboards, alerts, and runbooks; lead incident response and root-cause analysis for platform and data-tier events
  • Build and maintain CI/CD pipelines and promotion workflows for platform artifacts and configuration across environments
  • Manage infrastructure-as-code (Terraform or equivalent) for the platform data and compute tier, and automate routine platform operations
  • Define, document, and enforce platform engineering standards, hardening baselines, RBAC models, and configuration governance across the platform estate
  • Partner with InfoSec, Infrastructure, Network, and Identity on platform security operations — access reviews, key/secret rotation via PAM, and control attestation
  • Operate within and contribute to the Automation CoE governance framework; support internal audit, vendor risk, and regulatory examination for the platforms in scope
  • Document designs, runbooks, and standard operating procedures, and mentor junior engineers

Benefits

  • Medical, Dental, Vision, and Life Insurance
  • Paid Sick Leave, 3 weeks’ Vacation, and Holidays (about 11 a year)
  • HSA or FSA account and other voluntary benefits
  • 401(k) Retirement Saving Plan with Employer Match Program and 529 Savings Plan
  • Employee Mortgage Loan Program and free access to an Axos Bank Account with Self-Directed Trading
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service