Engineer I, Data Engineering

Samsung•Plano, TX

About The Position

This role involves understanding and documenting scope and requirements, designing and developing scalable custom ETL pipelines using big data technologies on a Cloud platform, and creating code to ingest data from various sources like relational databases, APIs, and files. The position requires developing complex SQL transformations on BigQuery, optimizing code, and adhering to engineering best practices for source control (GitHub), release management, and deployment. Responsibilities also include providing production support, job scheduling/monitoring, and ensuring ETL data quality through reporting. The engineer will contribute to software development and business management by solving business-related software problems through data modeling, analysis, and prediction, and by establishing and supporting software and infrastructure for big data analysis and integrity. This includes implementing software data analysis strategies that align with business strategy and establishing systems for data security and integrity.

Requirements

  • Bachelor’s degree in Computer Science, Applied Computer Science, Computer Applications, Computer Engineering, Information Technology, a related field, or a foreign equivalent plus 3 years post-baccalaureate experience in job offered or any engineering/IT related job titles.
  • 3 years of experience in statistical analysis and predictive modeling using Python.
  • 3 years of experience in ETL development using python and SQL.
  • 3 years of experience with GitHub, development IDEs including VS code.
  • 3 years of experience in data manipulation using SQL.
  • 3 years of experience in design, execution, and measurement of A/B and multivariate tests.
  • 3 years of experience with GCP services including Bigquery, Kubernetes and Composer and Apache Airflow.
  • 3 years of experience in performing data analysis on data visualization dashboards Superset, Jupyterlab, Tableau or PowerBI.
  • 3 years of experience with software tools including Jira and Confluence.
  • 3 years of experience working in the big data domain including providing production support, job scheduling/monitoring, ETL data quality, and data quality reporting.

Responsibilities

  • Understand and document scope and requirements through interactions with analysts and stakeholders.
  • Design and develop scalable code based custom ETL pipelines using big data technologies on Cloud platform.
  • Design and develop code to ingest data from a variety of data sources such as Relational databases, APIs and files.
  • Develop Complex SQL transformations on Bigquery.
  • Code and query Optimization.
  • Follow and contribute to engineering best practices for source control using Github, release management, deployment etc.
  • Provide production support, job scheduling/monitoring, ETL data quality, data quality reporting.
  • Contribute to software development and business management by effectively solving business-related software problems through data modeling/analysis/prediction.
  • Establish/support software and infrastructure that continuously grows to strategically analyze big data.
  • Understand and implement software data analysis strategies that meet business strategy.
  • Establish and operate software and infrastructure that can collect/analyze/predict big data (structured/unstructured).
  • Design and operate software structure that can maintain data integrity.
  • Proceed with software data model for business-related decision-making and evaluate data model's consistency.
  • Establish a software security system for data and protect from unauthorized users.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service