Engineer I, Data Engineering

Samsung Electronics•Plano, TX

About The Position

This role involves understanding and documenting scope and requirements through interactions with analysts and stakeholders. The engineer will design and develop scalable code for custom ETL pipelines using big data technologies on a Cloud platform. Responsibilities include designing and developing code to ingest data from various sources like relational databases, APIs, and files, as well as developing complex SQL transformations on BigQuery and optimizing code and queries. The role also requires adherence to and contribution to engineering best practices for source control using Github, release management, and deployment. Additionally, the engineer will provide production support, job scheduling/monitoring, ETL data quality, and data quality reporting. The position contributes to software development and business management by solving business-related software problems through data modeling, analysis, and prediction, and by establishing/supporting software and infrastructure that continuously grows to strategically analyze big data. The engineer will understand and implement software data analysis strategies that meet business strategy and establish and operate software and infrastructure capable of collecting, analyzing, and predicting big data (structured/unstructured). Designing and operating software structures that maintain data integrity, proceeding with software data models for business-related decision-making, evaluating data model consistency, and establishing a software security system for data to protect from unauthorized users are also key aspects of this role.

Requirements

  • Bachelor’s degree in Computer Science, Applied Computer Science, Computer Applications, Computer Engineering, Information Technology, a related field, or a foreign equivalent plus 3 years post-baccalaureate experience in job offered or any engineering/IT related job titles.
  • 3 years of experience in statistical analysis and predictive modeling using Python.
  • 3 years of experience in ETL development using python and SQL.
  • 3 years of experience with GitHub, development IDEs including VS code.
  • 3 years of experience in data manipulation using SQL.
  • 3 years of experience in design, execution, and measurement of A/B and multivariate tests.
  • 3 years of experience with GCP services including Bigquery, Kubernetes and Composer and Apache Airflow.
  • 3 years of experience performing data analysis on data visualization dashboards Superset, Jupyterlab, Tableau or PowerBI.
  • 3 years of experience with software tools including Jira and Confluence.
  • 3 years of experience working in the big data domain including providing production support, job scheduling/monitoring, ETL data quality, and data quality reporting.

Responsibilities

  • Understand and document scope and requirements through interactions with analysts and stakeholders.
  • Design and develop scalable code based custom ETL pipelines using big data technologies on Cloud platform.
  • Design and develop code to ingest data from a variety of data sources such as Relational databases, APIs and files.
  • Develop Complex SQL transformations on Bigquery.
  • Code and query Optimization.
  • Follow and contribute to engineering best practices for source control using Github, release management, deployment etc.
  • Provide production support, job scheduling/monitoring, ETL data quality, data quality reporting.
  • Contribute to software development and business management by effectively solving business-related software problems through data modeling/analysis/prediction.
  • Establish/support software and infrastructure that continuously grows to strategically analyze big data.
  • Understand and implement software data analysis strategies that meet business strategy.
  • Establish and operate software and infrastructure that can collect/analyze/predict big data (structured/unstructured).
  • Design and operate software structure that can maintain data integrity.
  • Proceed with software data model for business-related decision-making and evaluate data model's consistency.
  • Establish a software security system for data and protect from unauthorized users.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service