Data Engineer III

WalmartBentonville, AR
$90,000 - $180,000Onsite

About The Position

This role focuses on problem formulation, applying business acumen, and supporting data governance and strategy. The Data Engineer III will be responsible for extracting data from identified databases, creating data pipelines, and transforming data using appropriate techniques. They will develop knowledge of current data science and analytics trends, identify suitable data sources, and perform initial data quality checks. The position involves analyzing complex data elements, designing conceptual, physical, and logical data models (including data warehouse and data mart designs), defining relational tables, keys, and stored procedures, and evaluating existing data models. Additionally, the role requires developing efficient data flows, analyzing data-related system integration challenges, creating training documentation, and training end-users on data modeling. The Data Engineer III will also oversee the tasks of less experienced programmers and provide system troubleshooting support. Code development includes writing code for solutions and application features, creating test cases, developing proofs of concept, testing code, deploying software to production, contributing code documentation, maintaining playbooks, and providing progress updates. The role requires demonstrating up-to-date expertise and applying it to action plans, providing expert advice, aligning efforts to meet business needs, and building commitment for perspectives.

Requirements

  • Big Data infrastructure design using Cloud VMs and VPCs, Cloud GPUs, Cloud load balancer, Cloud DNS and CDN, and Serverless
  • Design and develop Orchestration pipelines using Apache Airflow, Automic, Crontab, Control-M, Google Cloud composer and Azure Data factory (ADF)
  • Design MLOps pipelines using Vertex AI platform, Google AutoML, MLFlow, Dialogflow, Gemini Code Assist, BigQuery ML & Azure Databricks
  • Design and develop data ingestion batch and streaming pipelines using Spark Connect, Sqoop, Google Cloud Pub/Sub, Google Cloud Dataflow, Apache NiFi, Azure Data Factory, Azure Event hubs, Azure DataBricks and Azure Stream analytics
  • Design pipelines for processing streaming data using Apache Kafka core, Confluent Kafka connect, Confluent KsqlDB, Apache Spark Streaming, Apache Spark Structured Streaming, Google Cloud functions, Google Data Stream, Google Dataproc and Apache Flink
  • Develop Data models using ER(Erwin) Studio and Microsoft Visio
  • Perform SQL operations using Hive, Presto (Trino), Apache Drill, Spark SQL, Hudi with Spark & hive, and Apache Iceberg with Spark
  • Build ETL Pipelines using Apache Spark, Scala, PySpark, Azure-databricks and Google-databricks
  • Design and develop Cloud Data Warehousing platforms using Google BigQuery, Google BigLake, Azure Synapse, Azure Data Lake, Snowflake, Apache Hudi and Databricks
  • Build dashboards and reports using Looker, Looker Studio, Tableau, PowerBI, Plotly, Ggplot2, & Uber H3
  • Perform data crunching/mining using Tableau Prep, Alteryx, DBT, Dataiku, RapidMiner, KNIME and Weka
  • Employer will accept any amount of experience with the required skills.

Responsibilities

  • Identifies possible options to address business problems through analytics, big data analytics, and automation.
  • Supports the development of business cases and recommendations.
  • Owns delivery of project activity and tasks assigned by others.
  • Supports process updates and changes.
  • Solves business issues.
  • Supports the documentation of data governance processes.
  • Supports the implementation of data governance practices.
  • Understands, articulates, and applies principles of the defined strategy to routine business problems that involve a single function.
  • Extracts data from identified databases.
  • Creates data pipelines and transforms data to a structure that is relevant to the problem by selecting appropriate techniques.
  • Develops knowledge of current data science and analytics trends.
  • Supports the understanding of the priority order of requirements and service level agreements.
  • Helps identify the most suitable source for data that is fit for purpose.
  • Performs initial data quality checks on extracted data.
  • Analyzes complex data elements, systems, data flows, dependencies, and relationships to contribute to conceptual, physical, and logical data models.
  • Develops the Logical Data Model and Physical Data Models including data warehouse and data mart designs.
  • Defines relational tables, primary and foreign keys, and stored procedures to create a data model structure.
  • Evaluates existing data models and physical databases for variances and discrepancies.
  • Develops efficient data flows.
  • Analyzes data-related system integration challenges and proposes appropriate solutions.
  • Creates training documentation and trains end-users on data modeling.
  • Oversees the tasks of less experienced programmers and stipulates system troubleshooting supports.
  • Writes code to develop the required solution and application features by determining the appropriate programming language and leveraging business, technical, and data requirements.
  • Creates test cases to review and validate the proposed solution design.
  • Creates proofs of concept.
  • Tests the code using the appropriate testing approach.
  • Deploys software to production servers.
  • Contributes code documentation, maintains playbooks, and provides timely progress updates.
  • Demonstrates up-to-date expertise and applies this to the development, execution, and improvement of action plans by providing expert advice and guidance to others in the application of information and best practices; supporting and aligning efforts to meet customer and business needs; and building commitment for perspectives and rationales.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service