Senior Data Engineer – Cloudera CDP & Big Data Pipelines
Summary
Design, develop, and optimize enterprise-scale data solutions on Cloudera CDP and Hadoop, building ETL pipelines with Spark, PySpark, NiFi, Sqoop, SQL, Python, and shell scripting in a Linux environment.
Sabenza IT & Recruitment seeks a data engineer to join a high-performing data engineering environment in Johannesburg, South Africa. You will design, develop and optimize enterprise-scale data solutions, focusing on CDP and Hadoop ecosystems, Spark, PySpark, NiFi, Sqoop, and SQL-driven analytics.
You will build robust ETL pipelines, process large datasets, and support analytics initiatives with Python and shell scripting in a Linux/Unix environment.