Data Engineer
Summary
Build and maintain big-data pipelines and warehouses using Hive, Impala, PySpark, and ETL tools; deploy and virtualize data with Cloudera, Denodo, and Linux.
- 1-5 years of experience in data engineering or a related field.
- Good understanding and completion of projects using waterfall/Agile methodology.
- Hands‑on experience in DevOps deployment and data virtualisation tools like Denodo will be preferred.
- Understanding of reporting or visualization tool like SAP BO and Tableau is important.
- Track record in implementing systems using Hive, Impala and Cloudera Data Platform will be preferred.
- Hands‑on experience in big data engineering jobs using Python, Pyspark, Linux, and ETL tools like Informatica.
- Strong SQL and data modelling and data analysis skills.
- Good understanding of analytics and data warehouse implementations.
- Ability to troubleshoot complex issues ranging from system resource to application stack traces.