Cloudera Developer - Onsite

USDCCharlotte, NC
Onsite

About The Position

Genesis10 is seeking a Cloudera Developer for a contract position with a Global Financial Institution. This role involves developing and maintaining data solutions using the Cloudera platform, working closely with cross-functional teams to understand data requirements and deliver robust, scalable data solutions. The primary focus will be on designing, developing, and implementing data processing pipelines and data ingestion frameworks using PySpark, ezflow, and other methods for moving data within a Hadoop environment.

Requirements

  • 5-7 years of experience with distributed data/computing tools: Hadoop, Hive, MySQL, etc
  • Strong problem-solving skills with an emphasis on product development
  • Experience working with and creating data architectures
  • Excellent written and verbal communication skills for coordinating across teams
  • Experience creating clear and effective data visualizations and dashboards
  • Knowledge of ETL processes, data integration, and data quality best practices
  • Proficient in Python programming for data processing, automation, and analytical solutions
  • Experience using pandas to cleanse, transform, and analyze large datasets
  • Experience using PySpark for large-scale data processing and distributed computing environments
  • Strong SQL skills with experience developing complex queries, joins, aggregations, and performance optimization
  • Experience working with AutoSys
  • Ability to analyze data, identify trends and anomalies, and deliver actionable business insights
  • A drive to learn and master new technologies and techniques

Responsibilities

  • Utilize multiple architectural components in the design and development of client requirements
  • Maintain, improve, clean, and manipulate data for the operational and/or analytics data systems
  • Design solutions by finding better ways of solving technical problems and challenging the status quo
  • Document and communicate required information for deployment, maintenance, support, and business functionality
  • Adhere to team delivery/release process and cadence pertaining to code deployment and release
  • Design, develop, and maintain data processing pipelines using Cloudera technologies such as Apache Hadoop, Apache Spark, Apache Hive, and Python
  • Collaborate with data engineers and data scientists to understand data requirements and translate them into technical specifications
  • Develop and maintain data ingestion frameworks for efficiently extracting, transforming, and loading data from various sources into the Cloudera platform
  • Optimize and tune data processing jobs to ensure high performance and scalability
  • Implement data governance and security policies to ensure data integrity and compliance
  • Monitor and troubleshoot data processing jobs to identify and resolve issues in a timely manner
  • Perform unit testing and debugging of data solutions to ensure high quality and reliability
  • Document technical specifications, data flows, and data architecture diagrams
  • Stay updated with the latest advancements and best practices in Cloudera technologies and big data analytics

Benefits

  • Behavioral Health Platform
  • Medical, Dental, Vision
  • Health Savings Account
  • Voluntary Hospital Indemnity (Critical Illness & Accident)
  • Voluntary Term Life Insurance
  • 401K
  • Sick Pay (for applicable states/municipalities)
  • Commuter Benefits (Dallas, NYC, SF, and Illinois)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service