freehire launches on Product Hunt on 26 August.

Follow →

Senior Executive, Data Engineer

Summary

Build and maintain scalable data pipelines and analytics infrastructure for a global palm oil company, integrating ERP, IoT, and geospatial data to support audit, reporting, and advanced analytics like ML and RPA.

We value our people and encourage everyone to grow professionally.

Responsibilities

  • Strategic Alignment & Architecture: Assist the Head, GCA IT & Data Analytics in defining and executing the department’s data engineering vision, roadmap, and target architecture, ensuring alignment with organisational objectives, enterprise policies, and technology standards.
  • Data Pipeline & Development: Design, develop, and maintain robust Extract, Transfer & Load (ETL) pipelines to ingest, process, and transform data from diverse sources, including Enterprise Resource Planning (SAP ECC6 and S/4HANA) and operational systems, IoT devices, geospatial platforms, and external data providers.
  • Scalable Analytics Infrastructure: Ensure that data pipelines supporting the GIGA analytics infrastructure are scalable, secure, cost-efficient, and highly reliable to meet growing business and operational demands.
  • Performance & Cost Optimisation: Monitor, troubleshoot, and continuously optimise data workflows to improve performance, resilience, and cost efficiency.
  • Data Storage & Management: Build, manage, and maintain GIGA’s Data Mart, data lake, and related data storage solutions to support department’s reporting, analytics, and advanced use cases.
  • Data Modelling & Architecture Collaboration: Develop and maintain data models that support analytics, operational reporting, and decision-making tailored for audit and investigation activities. Collaborate closely with Group IT and Group Digital teams on data architecture decisions to support current and future data requirements.
  • Data Quality, Reliability & Governance: Implement data validation, testing, monitoring and observability practices to ensure data accuracy, completeness and consistency. Proactively identify and resolve data quality issues, pipeline failures and anomalies.
  • Stakeholder and Cross-Functional Collaboration: Work closely with GCA auditors, data analytics specialists, data scientists, and business stakeholders across plantation’s upstream, downstream and support operations to understand data requirements related to audit and controls.
  • Self-service Analytics Enablement: Enable self-service analytics by delivering clean, trusted, and well-documented datasets. Support reporting, dashboards, and other GIGA data consumers through coaching, knowledge sharing and technical guidance.
  • Advanced Analytics & Emerging Technologies: Explore, develop, and support the implementation of advanced analytics solutions, including but not limited to Robotic Process Automation (RPA), Machine Learning, and AI, leveraging large structured and unstructured datasets (i.e. voice, video, text, geospatial images, and sensor data) to improve audit and investigation efficiencies.
  • Innovation & Continuous Improvement: Stay informed of emerging trends, tools, and best practices in data engineering and analytics, and proactively recommend improvements to enhance data capabilities and operational efficiency.
  • Documentation & Standards: Establish and maintain comprehensive documentation for data pipelines, data models and engineering standards for GIGA. Actively participate in analytics innovation initiatives and continuous improvement programmes.

Requirements

  • Relevant tertiary education (minimum bachelor’s degree), primarily in Computer Science / Computer Engineering, Software Engineering, Data Science, Statistics, Mathematics or equivalent.
  • Professional certification or technical qualification (e.g. Certified Data Analyst, Certified Data Scientist, Certified Data Engineer, Microsoft Certified: Azure Fundamentals) is a plus.
  • Minimum 3 years of experience in data engineering, software engineering or a related role.
  • Proficiency in SQL and at least one programming language (e.g. Python, Java, Scala)
  • Minimum 3 years of experience with ETL / ELT tools and orchestration framework.
  • Familiarity with cloud data platforms and data warehouses.
  • Understanding of data modelling concepts and performance optimisation.

SD Guthrie Berhad (SD Guthrie) is one of the world’s leading producers of Certified Sustainable Palm Oil (CSPO), representing approximately 12% of the global market share (as of 31 December 2022). We are listed on Bursa Malaysia. Supported by a large institutional base, we are a strategic company of Permodalan Nasional Berhad, Malaysia’s largest unit trust company and our major shareholder. With over two centuries of heritage, SD Guthrie has evolved into an integrated company operating throughout the entire spectrum of the palm oil value chain. Our dedicated workforce of more than 84,000 employees across 12 countries serve customers in over 90 countries worldwide. We operate 234 plantation estates in Malaysia, Indonesia, Papua New Guinea, and the Solomon Islands, supported by 11 refineries globally with a total refining capacity of 4 million metric tonnes per year. Our downstream businesses engage in trading, manufacturing, and sales of a diverse range of palm oil derivatives, including oleochemicals, biodiesel, and nutraceuticals. We are equipped with 5 R&D Centres, 3 Innovation Centres, and 1 Genetic Testing Facility to support our R&D efforts spearheaded by lab scientists, data engineers, and tech innovators. Our focus is on next-gen robotics and tech-driven solutions for the palm oil and agri-business sector, alongside renewable energy initiatives. We are the world’s first palm oil company to have our net-zero GHG emissions reduction targets approved by the Science Based Targets initiative (SBTi). Join us today, as we continue to Unlock Nature's Superpowers.

See also