Staff Data Engineer

Twenty•New York, NY
•$192,000 - $455,000•Onsite

About The Position

You'll be based in Twenty's New York City office, building and operating the data integrations that feed Twenty's mission-critical platform. Twenty builds software for cyber operations, and the applications this role feeds are used by operators and analysts working against hard national security problems. This is a hands-on engineering role with real production ownership: you'll build the pipelines that retrieve, process, and integrate new data sources into our applications, monitor the health of the data flowing through the systems you can reach, and debug production issues when they surface. Where your work runs in environments you can't access, you'll partner with Twenty's forward deployed engineers, who operate on-site at customer facilities, to get it deployed and keep it healthy. You'll join the data platform team, reporting to its engineering manager. The data integration tooling you'll build on is young and under active development. You'll be among its first production users, and what you learn in operation will directly shape how it evolves. If you're an experienced data engineer who wants your work in the hands of operators the same week you build it, and you'd rather own a production system end to end than ship features into a backlog, this role is for you.

Requirements

  • 5+ years of experience in data engineering, or software engineering with a substantial data infrastructure focus.
  • Expert Python skills, with a track record of designing and maintaining production-grade codebases that other engineers extend and depend on.
  • Hands-on experience building and operating ETL/ELT pipelines with Spark (PySpark), ideally as AWS Glue jobs.
  • Strong SQL and schema design skills, including the ability to analyze access patterns and make sound partitioning and indexing decisions.
  • Experience with column-oriented analytical databases (example technologies: ClickHouse, Redshift, BigQuery).
  • Experience debugging production data systems: reading logs, tracing failures across components, and root-causing issues under time pressure.
  • Demonstrated ability to work independently on production systems, including sound judgment about when to decide and when to escalate.
  • Ability to work on-site full-time at Twenty's New York City office, with an initial onboarding period at our Arlington, VA office and occasional travel there afterward.
  • U.S. citizenship required.
  • No active clearance required to start.

Nice To Haves

  • Experience with data lake and large-scale query technologies (example technologies: Apache Iceberg, Delta Lake, Trino, Presto, Athena).
  • Experience with streaming and message queue technologies (example technologies: Kafka, NATS, Kinesis, RabbitMQ).
  • Hands-on experience with observability tooling (Grafana, Datadog, Splunk, or similar).
  • Experience deploying and operating software in air-gapped, classified, or otherwise restricted environments.
  • Familiarity with containerized environments (Docker and container orchestration) and CI/CD concepts.

Responsibilities

  • Design and build data pipelines that retrieve, parse, transform, and load new data sources into Twenty's applications, primarily as AWS Glue jobs written in PySpark.
  • Investigate unfamiliar source systems and data formats, working with Twenty's engineers to understand semantics, access patterns, and constraints before writing code.
  • Design schemas and data models that fit the platform's performance characteristics and the shape of the data, including analytical models in ClickHouse.
  • Write code that is testable, debuggable, and maintainable by engineers operating it in environments you can't access. Tests carry unusual weight here: they're how the field trusts a change you'll never see run.
  • Write integration code that meets the security standards of classified environments: no hardcoded secrets, disciplined credential handling, and audit-ready practices throughout, so your work moves through customer approval processes without friction.
  • Deploy, operate, and iterate on integrations against the systems reachable from the office, end to end.
  • Package pipelines and integrations destined for restricted environments so forward deployed engineers and customer deployment teams can install, verify, and operate them without you in the room.
  • Support field deployments remotely: reproduce issues, cut fixes, and move them through the pipeline quickly when an on-site engineer is blocked.
  • Operate and improve pipelines built elsewhere on the team, and feed operational learnings back into their design.
  • Monitor the health of data pipelines and platform services using the LGTM stack (Grafana, Loki, Tempo, Mimir).
  • Track data quality, throughput, and freshness; identify anomalies and degradations before they become incidents.
  • Build and improve the dashboards and alerts that make pipeline health visible to the team and to the forward deployed engineers who depend on it.
  • Diagnose and resolve production issues in the data path: failed ingests, malformed source data, pipeline stalls, performance degradation.
  • Own resolution of incidents within your sphere of responsibility, escalating with a clear picture of what you found, not just what triggered the alert.
  • Drive root cause analysis and post-incident reviews, and turn findings into runbooks, fixes, or upstream design changes.
  • Participate in on-call rotation using PagerDuty, with clear escalation paths across the engineering team.
  • Work closely with forward deployed engineers on-site at customer facilities, sharing context in both directions so field reality and platform design stay connected.
  • Turn operational observations, data source opportunities, and recurring patterns from the field into platform capabilities rather than one-off fixes.

Benefits

  • Medical, dental, and vision plan options.
  • Life / AD&D, disability coverage options.
  • Paid parental leave for eligible full-time employees. 12 weeks for birthing parents, 4 for non-birthing parents, 6 weeks for adoptive, foster, or intended parents through surrogacy.
  • Paid holidays and flexible PTO.
  • 401(k) with pre-tax and Roth options.
  • HSA/FSA options, dependent care FSA.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service