freehire launches on Product Hunt on 26 August.

Follow →

PySpark Data Engineer

Roles and Responsibilities:

  • Responsible for developing and maintaining applications with PySpark
  • Contribute to the overall design and architecture of the application developed and deployed.
  • Performance Tuning wrt to executor sizing and other environmental parameters, code optimization, partitions tuning, etc
  • Interact with business users to understand requirements and troubleshoot issues.
  • Implement Projects based on functional specifications.

Must-Have Skills:

  • Relevant Experience: 3-6 Years
  • SQL - Mandatory
  • Python - Mandatory
  • SparkSQL - Mandatory
  • PySpark - Mandatory
  • Hive - Mandatory
  • HDFS and Spark - Mandatory
  • Scala - Advantage
  • Apache Airflow - Advantage


Requirements

3-6 Years of Experience
Must Have: PySpark/Spark, Python, SQL, Knowledge on Hadoop ecosystem
Good to have: Airflow, Scala

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available