SRE / Infrastructure Engineer (LABGEN)

ComtronGreat Neck, NY
Onsite

About The Position

Comtron, part of the MEDFAR group, is seeking a hands-on Site Reliability Engineer / Infrastructure Engineer to own the reliability, security, and continuity of Comtron’s Windows- and Linux-based production environment for the Labgen platform. The role involves managing production infrastructure, availability, and security, overseeing backups and disaster recovery, and handling production deployments and hosting environments. Additionally, the engineer will support infrastructure modernization, technology upgrades, and potential transitions to public or hybrid-cloud environments. This position reports to the Software Development Manager and collaborates with various teams including IT, Software Development, Quality Assurance, Security, Support, and Finance.

Requirements

  • 5+ years of experience in Site Reliability Engineering, Infrastructure, DevOps, or a similar role, ideally within a SaaS or regulated environment.
  • Strong hands-on Linux administration experience, including RHEL, CentOS, Ubuntu, or similar environments.
  • Experience administering Windows Server and application-hosting environments.
  • Strong Apache web server administration experience.
  • Experience designing or managing high-availability Linux production environments.
  • Strong knowledge of TCP/IP, DNS, VPNs, firewalls, certificates, and network segmentation.
  • Experience with Azure, AWS, GCP, or hybrid-cloud environments.
  • Experience with enterprise backups, disaster recovery, and production incident management.
  • Strong working knowledge of SQL and relational databases, including query troubleshooting, performance analysis, migrations, upgrades, and connectivity issues.
  • Familiarity with .NET and JavaScript-based application environments.
  • Experience with version control, build and release processes, deployment automation, and CI/CD tools such as Git, CVS, and Jenkins.
  • Experience with access controls, system hardening, firewall management, vulnerability remediation, and infrastructure security.
  • Familiarity with ISO 27001 controls and evidence requirements.
  • Familiarity with HIPAA, HITECH, ONC Health IT Certification requirements, or other U.S. healthcare privacy and security obligations is an asset.
  • Strong incident ownership, documentation, and cross-functional communication skills.
  • Strong written and verbal communication skills in English.

Nice To Haves

  • Experience working in regulated healthcare environments.
  • Familiarity with laboratory information systems, electronic medical records, or clinical information systems.
  • ISO 27001 Lead Implementer or Auditor certification.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, ELK, Datadog, or SentinelOne.
  • Python or Bash scripting experience.
  • French language proficiency.

Responsibilities

  • Manage the lifecycle of Windows and Linux production servers, including provisioning, configuration, patching, hardening, and decommissioning.
  • Design and maintain high-availability configurations to reduce single points of failure and ensure continuous access to the Labgen platform.
  • Monitor uptime, performance, capacity, resource utilization, and infrastructure costs.
  • Maintain production access controls, firewall rules, certificates, networking, storage, and database connectivity.
  • Maintain network segmentation between production, development, and corporate environments.
  • Coordinate vulnerability assessments, penetration testing, and remediation activities.
  • Manage privately hosted infrastructure and relationships with hardware, software, connectivity, and data center vendors.
  • Support infrastructure purchasing, contract renewals, licensing, and capacity planning.
  • Evaluate infrastructure tools and public or hybrid-cloud solutions, including Azure, AWS, or GCP.
  • Own production backup policies, including scope, frequency, retention, monitoring, and integrity validation.
  • Ensure backups are completed successfully and recovery procedures are regularly tested and documented.
  • Establish and maintain recovery time and recovery point objectives.
  • Maintain and regularly test disaster recovery procedures for hosted systems.
  • Lead improvements following disaster recovery exercises and production incidents.
  • Manage staging, UAT, and production hosting environments.
  • Execute production deployments and rollback procedures in coordination with the Software Development team.
  • Coordinate deployment windows, infrastructure changes, and change-management activities.
  • Manage application hosting infrastructure, including Windows Server, Apache, application services, scheduled processes, networking, storage, certificates, and database connectivity.
  • Support CI/CD tooling, deployment automation, and pipeline reliability.
  • Plan and coordinate operating system, database, framework, library, and application dependency upgrades.
  • Assess cloud migration options, application dependencies, risks, and phased modernization strategies.
  • Support proof-of-concept initiatives and ensure proposed solutions meet security, reliability, scalability, compliance, and disaster recovery requirements.

Benefits

  • Generous health, vision, and dental group insurance coverage (after probation)
  • 2 weeks of paid time off
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service