Mid-Level Data Engineer, Cloud Data Platforms

Kentro•UNAVAILABLE, UNAVAILABLE
•Remote

About The Position

Kentro is hiring for a Mid-Level Data Engineer to support secure cloud data and applied AI initiatives for government customers. You will build governed data pipelines, automate cloud and operational tasks, solve cross-layer technical problems, and turn evolving requirements into reliable, well-documented solutions. This role is well suited to an engineer who brings strong fundamentals and, above all, the drive to take ownership, learn quickly, communicate clearly, and see important work through to a verified outcome. You will contribute independently within a defined scope while collaborating with technical leaders and growing toward broader platform ownership. This position can be performed remotely within the United States and will support Eastern Time working hours.

Requirements

  • Three to five years of relevant experience in data engineering, software engineering, cloud engineering, analytics engineering, or a closely related discipline.
  • Hands-on programming experience with Python or a comparable language, including reading unfamiliar code and debugging systematically.
  • Practical SQL experience with joins, aggregations, transformations, and row-level and aggregate validation.
  • Experience developing or supporting data pipelines, schemas, APIs, databases, cloud storage, or distributed data-processing workflows.
  • Working knowledge of Git, including commits, branches, pull requests, code review, and ordinary conflict resolution.
  • Experience testing work, retaining evidence, and writing documentation or runbooks that another engineer can follow.
  • Demonstrated ownership, learning agility, persistence, responsiveness, and follow-through when solving unfamiliar or ambiguous problems.
  • Ability to communicate technical status, risks, assumptions, and blockers clearly to both technical and nontechnical stakeholders.
  • Commitment to security, least privilege, responsible data handling, and compliance with customer requirements.
  • Ability to work effectively in a remote environment and collaborate during Eastern Time working hours.
  • US Citizen or Lawful Permanent Resident (Green Card)
  • Willing and able to obtain and maintain Public Trust Clearance or higher

Nice To Haves

  • Experience with Azure, Azure Government, AWS, or another major cloud platform.
  • Experience with Databricks, Apache Spark, PySpark, Delta Lake, Lakeflow, or a comparable data-processing platform.
  • Experience with Terraform or another Infrastructure-as-Code tool.
  • Familiarity with Azure Data Lake Storage, Data Factory, Key Vault, Entra ID, managed identities, RBAC, private endpoints, DNS, or virtual networks.
  • Knowledge of data governance, catalogs, lineage, stewardship, audit trails, observability, and reproducible processing.
  • Experience with REST APIs, JSON/JSONL, CSV, Parquet, large-file processing, third-party ingestion, CI/CD, containers, or structured logging.
  • Exposure to AI/LLM-enabled applications, model-serving endpoints, evaluation, or automated regression testing.
  • Experience in a government, healthcare, or other regulated or compliance-sensitive environment.
  • Bachelor's degree in computer science, data science, information systems, engineering, or a related field, or equivalent relevant experience.

Responsibilities

  • Develop, test, deploy, and maintain batch and API-based data ingestion, transformation, and publishing workflows.
  • Write clear, maintainable Python, PySpark, and SQL for data processing, validation, reconciliation, and automation.
  • Build and support Databricks workflows using notebooks, jobs, compute, Delta tables, catalogs, schemas, permissions, and governed data products.
  • Implement layered data designs, including landing/raw, Bronze, Silver, and curated or presentation outputs.
  • Add data-quality checks, lineage, audit metadata, checksums, retry and recovery behavior, logging, and reproducible tests to pipelines.
  • Support Azure data-platform services, cloud storage, secrets, workload identities, role-based access, monitoring, networking, and private connectivity.
  • Develop and review Infrastructure as Code, primarily Terraform, using reusable modules and environment-specific configuration.
  • Deliver traceable changes through Git branches, pull requests, code reviews, issue tracking, and CI/CD workflows.
  • Troubleshoot data, code, access, authentication, permissions, deployment, networking, and runtime issues using logs, tests, queries, and documented evidence.
  • Support AI-enabled data workflows, including model endpoints, response validation, evaluation, regression testing, and responsible-use controls.
  • Create and maintain architecture diagrams, runbooks, implementation notes, decision records, status updates, and handoff documentation.
  • Communicate progress, risks, availability, and blockers promptly; ask for help early enough to protect delivery and close the loop on commitments.
  • Participate in stand-ups, design reviews, demonstrations, and stakeholder discussions, translating technical findings for the intended audience.

Benefits

  • Paid time off
  • Healthcare benefits
  • Supplemental benefits
  • 401k including an employer match
  • Discount perks
  • Rewards
  • Education reimbursement for certifications, degrees, or professional development
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service