Staff Data Engineer

Clarium
•$170,000 - $210,000•Remote

About The Position

Clarium is seeking a Staff Data Engineer to set technical direction for their data platform. This role involves deciding how data is modeled, moved, and trusted across the company, improving both systems and the engineers around them. The engineer will work with transactional databases and analytical warehouses, own the pipelines connecting them, and collaborate with analytics, product, and engineering teams. A significant part of the role involves handling supply chain data from hospital ERP systems, which often have varying schemas and poor documentation, requiring strategic decision-making about what to build.

Requirements

  • 8+ years of data or software engineering experience, including significant time owning systems end to end in production
  • Expert-level SQL: you write complex analytical queries as a matter of course and can read a query plan and explain why something is slow
  • Strong Python for production data work, including pipeline code, transformation logic, testing, and tooling
  • Deep experience with relational databases, including schema design, normalization tradeoffs, transactions, and performance tuning (Postgres and Snowflake strongly preferred)
  • Hands-on experience with data pipeline and orchestration tooling such as Airflow, Dagster, Prefect, dbt, Fivetran, Spark, or Kafka; we care more about depth and judgment than an exact stack match
  • Production experience running data workloads on AWS (e.g., S3, RDS, Lambda, ECS), with an understanding of the cost, security, and networking implications of how you build
  • A track record of leading multi-quarter, cross-team initiatives that depended on teams you don't manage, and comfort with the ambiguity of deciding what should be built

Nice To Haves

  • Experience with healthcare data (claims, EHR/EMR, HL7 or FHIR, ICD-10, CPT) or working under HIPAA with PHI, de-identification, and audit requirements
  • Familiarity with supply chain data from ERP systems such as Oracle, Workday, or Lawson, including where it tends to be unreliable
  • Infrastructure-as-code and deployment automation (Terraform, CDK, CI/CD for data infrastructure)
  • Streaming and event-driven architectures, or data quality and observability tooling

Responsibilities

  • Own the architecture and technical roadmap for core data infrastructure on AWS, spanning ingestion, transformation, storage, and serving layers
  • Design, build, and operate reliable batch and near-real-time pipelines with clear SLAs and the observability to back them up
  • Model supply chain data from external ERP systems into a coherent, reusable warehouse model, and lead the migration of legacy assets toward it
  • Tune performance and cost across Postgres and Snowflake, including query plans, indexing, partitioning, warehouse sizing, and storage strategy
  • Establish engineering standards for data work (testing, code review, CI/CD, data quality checks, documentation) and mentor engineers through design reviews, pairing, and honest technical feedback
  • Partner with analytics, product, and engineering teams to turn ambiguous questions into durable data models instead of ad hoc extracts
  • Lead governance for sensitive data, including lineage, access controls, retention, auditability, and de-identification where required

Benefits

  • Incentive Stock Options proportionate to your salary
  • Fully remote, with a NYC co-working space available; distributed team across multiple time zones with opportunities for in-person time
  • Unlimited PTO
  • Top-tier health, vision, and dental benefits
  • 401K
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service