Point your AI agent at freehire and let it find you a job.

Get the CLI →

Marktine Technology Solutions Pvt Ltd

Open 48d

PySpark Data Engineer

Posted Updated 3 views
Discussion

Summary

Builds and optimizes data pipelines using PySpark, SQL, and the Hadoop ecosystem to process large datasets and support business analytics.

Roles and Responsibilities:

  • Responsible for developing and maintaining applications with PySpark
  • Contribute to the overall design and architecture of the application developed and deployed.
  • Performance Tuning wrt to executor sizing and other environmental parameters, code optimization, partitions tuning, etc
  • Interact with business users to understand requirements and troubleshoot issues.
  • Implement Projects based on functional specifications.

Must-Have Skills:

  • Relevant Experience: 3-6 Years
  • SQL - Mandatory
  • Python - Mandatory
  • SparkSQL - Mandatory
  • PySpark - Mandatory
  • Hive - Mandatory
  • HDFS and Spark - Mandatory
  • Scala - Advantage
  • Apache Airflow - Advantage


Requirements

3-6 Years of Experience
Must Have: PySpark/Spark, Python, SQL, Knowledge on Hadoop ecosystem
Good to have: Airflow, Scala

Skills

Apply

See also

Data Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available