Machine Learning Operations Engineer

University of California, IrvineIrvine, CA
Onsite

About The Position

Under the direction of the Principal Investigator, the Machine Learning Operations Engineer supports the research efforts of several regional projects. Key responsibilities include duties related to setup, administration, system hardening, and orchestration of the laboratory’s on-premises and cloud-based Linux servers. Development and testing of machine learning (ML) analyses and statistical models, including pipelines incorporating large language models and dependent on CUDA and similar technologies, in the Python and R programming languages. Additional responsibilities include packaging of analyses and software into publishable research products, including Jupyter notebooks, Python/R packages, and Docker containers. Deployment of informatics pipelines utilizing ML models into production environments, where they can support both research and operational activities at UCI Health. Development of testing and monitoring harnesses for these pipelines, which may include regression tests and dashboards for key performance metrics. Maintenance and debugging of ML pipelines across the full stack, from the hardware and virtualization layers up through user interface components. Coordination and engagement in collaborative scientific and technical writing, including creating training and educational documents, data collection tools, recruitment or protocol scripts, IRB materials, scientific manuscripts, and components of grant proposals. Supervision and support of lab members and students in the course of their work on the laboratory’s server infrastructure.

Requirements

  • Clear and professional communication skills; verbal and written
  • Ability to analyze a problem from inception to completion and provide suggested solutions.
  • Effective and professional interpersonal skills
  • Highly attentive to proper handling of confidential information and documents
  • Ability to maintain accurate database files
  • Ability to function well in a team environment
  • Experience deploying, administering, and securing Linux-based servers in on-premises and/or cloud environments.
  • Bachelor's degree in related area and / or equivalent experience / training
  • Minimum of 2-3 years of experience.
  • Prior research experience with demonstrated independent responsibilities and activities
  • Must be able to provide proof of work authorization

Nice To Haves

  • Experience working in a laboratory

Responsibilities

  • Setup, administration, system hardening, and orchestration of on-premises and cloud-based Linux servers.
  • Development and testing of machine learning (ML) analyses and statistical models, including pipelines incorporating large language models and dependent on CUDA and similar technologies, in Python and R.
  • Packaging of analyses and software into publishable research products (Jupyter notebooks, Python/R packages, Docker containers).
  • Deployment of informatics pipelines utilizing ML models into production environments.
  • Development of testing and monitoring harnesses for ML pipelines, including regression tests and dashboards for key performance metrics.
  • Maintenance and debugging of ML pipelines across the full stack.
  • Coordination and engagement in collaborative scientific and technical writing.
  • Supervision and support of lab members and students on server infrastructure.

Benefits

  • medical insurance
  • sick and vacation time
  • retirement savings plans
  • access to a number of discounts and perks
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service