Senior Data Engineer – Big Data & Cloudera

Sabenza IT
Johannesburg, Gauteng
Full-time
Posted 1 day ago
Full-time IT & Software
Apply for this job

Responsibilities

  • Design, develop and maintain scalable Big Data and ETL data pipelines.
  • Work extensively with the Cloudera Data Platform (CDP) and Hadoop ecosystem.
  • Develop and optimize data processing solutions using Apache Spark and PySpark.
  • Build and manage data ingestion pipelines using Apache NiFi and Sqoop.
  • Work with HDFS, Hive and Impala for large-scale data storage, processing and querying.
  • Develop complex and optimized SQL queries for data extraction, transformation and analysis.
  • Develop data engineering solutions using Python and Shell scripting.
  • Integrate and process data from enterprise data sources, including Oracle.
  • Develop, maintain and optimize ETL processes to support business and analytical requirements.
  • Monitor data pipelines and scheduled workloads using Control-M.
  • Perform troubleshooting, performance tuning and root-cause analysis across data processing environments.
  • Work within Linux/Unix environments to administer, troubleshoot and automate data engineering processes.
  • Support data quality, data integrity and data availability across enterprise data platforms.
  • Collaborate with Data Analysts, Developers, Architects, Business Analysts and other technology teams.
  • Contribute to the continuous improvement of data engineering standards, processes and platforms.

Requirements

  • 7–8 years of solid hands-on experience as a platform and data engineer (intermediate to senior level).
  • Design, develop and maintain scalable Big Data and ETL data pipelines.
  • Work extensively with the Cloudera Data Platform (CDP) and Hadoop ecosystem.
  • Develop and optimize data processing solutions using Apache Spark and PySpark.
  • Build and manage data ingestion pipelines using Apache NiFi and Sqoop.
  • Work with HDFS, Hive and Impala for large-scale data storage, processing and querying.
  • Develop complex and optimized SQL queries for data extraction, transformation and analysis.
  • Develop data engineering solutions using Python and Shell scripting.
  • Integrate and process data from enterprise data sources, including Oracle.
  • Develop, maintain and optimize ETL processes to support business and analytical requirements.
  • Monitor data pipelines and scheduled workloads using Control-M.
  • Perform troubleshooting, performance tuning and root-cause analysis across data processing environments.
  • Work within Linux/Unix environments to administer, troubleshoot and automate data engineering processes.
  • Support data quality, data integrity and data availability across enterprise data platforms.
  • Collaborate with Data Analysts, Developers, Architects, Business Analysts and other technology teams.
  • Contribute to the continuous improvement of data engineering standards, processes and platforms.

Interested in this role?

You'll apply on Sabenza IT's own site before 12 October 2026. Put your most relevant experience at the top of your CV first.

Apply Now

Real employers never ask you to pay to apply. How to spot a job scam

Similar jobs you might like
Business Analyst
Mindworx Consulting · Johannesburg
Business Analyst
HR Genie ·
Business Analyst
HR Genie ·
Mid - Senior OutSystems Developer - Centurion, Hybrid
Datafin IT Recruitment · Centurion
Apply for this job
Job details
Job TypeFull-time
LocationJohannesburg, Gauteng
CategoryIT & Software
Experience7+ years
Posted28 September 2026
Closing Date12 October 2026
Sabenza IT
Johannesburg, Gauteng

New to applying?

Free guides on CVs, cover letters, interviews and the Z83 form.

How to apply guides
Apply for this job