Data Engineer Spark & PySpark - ETL & Big Data
Summary
Senior data role at SII Ouest in Le Mans, France: migrating existing processing to Apache Spark, building PySpark ETL pipelines in a Big Data environment, optimizing SQL queries, and working with Cloudera/Hadoop, Hive and Impala, while contributing to testing and technical documentation.
SII Ouest recherche unanalyste Data Senior pour migrer des traitements vers Apache Spark et développer des pipelines PySpark dans un cadre Big Data. Vous contribuerez à optimiser les requêtes SQL et à intervenir sur Cloudera/Hadoop, Hive et Impala, tout en participant aux tests et à la documentation technique.
Le candidat idéal possède au moins 5 ans d’expérience Data, maîtrise Spark/PySpark et SQL, et connaît Hadoop/Cloudera/CDP.