freehire launches on Product Hunt on 26 August.

Follow →

PySpark Developer / Senior Data Engineer

Summary

Designs, builds, and maintains scalable data pipelines using PySpark and Apache Spark; works with large datasets, optimizes Spark jobs, and collaborates with teams on data engineering tasks.

Do you love a career where you Experience, Grow & Contribute at the same time, while earning at least 10% above the market? If so, we are excited to have bumped onto you.

Learn how we are redefining the meaning of work , and be a part of the team raved by Clients, Job-seekers and Employees.

  • Jobseeker Video Testimonials
  • Employee Glassdoor Reviews

If you are a PySpark Developer / Senior Data Engineer looking for excitement, challenge and stability in your work, then you would be glad to come across this page.

We are an IT Solutions Integrator/Consulting Firm helping our clients hire the right professional for an exciting long-term project. Here are a few details.

Check if you are up for maximizing your earning/growth potential, leveraging our Disruptive Talent Solution.

Role: PySpark Developer / Senior Data Engineer
Location:
HYDERABAD | BANGALORE | PUNE | CHENNAI
Experience: 6+ Years
Employment Type: Contract to hire
Notice Period:0-30 days(If you have negotiable notice period or buyout option please apply)




Requirements

We are looking for an experienced PySpark Developer with strong hands-on expertise in big data processing, distributed computing, and data engineering. The ideal candidate will have deep experience building scalable data pipelines, transforming large datasets, and working with Spark-based ecosystems in production environments.

Key Responsibilities

  • Design, build, and maintain scalable data pipelines using PySpark and Apache Spark
  • Develop efficient ETL/ELT workflows for batch and near-real-time processing
  • Optimize Spark jobs for performance, reliability, and cost efficiency
  • Work with large structured and unstructured datasets
  • Integrate data from multiple sources such as databases, APIs, files, and cloud storage
  • Write reusable, modular, and maintainable PySpark code
  • Troubleshoot job failures, data quality issues, and performance bottlenecks
  • Collaborate with data architects, analysts, platform teams, and business stakeholders
  • Implement data validation, monitoring, and logging frameworks
  • Support deployment, scheduling, and orchestration of data pipelines
  • Participate in design reviews, code reviews, and technical discussions
  • Mentor junior engineers and contribute to team best practices

Required Skills

  • Strong hands-on experience with PySpark and Apache Spark
  • Deep understanding of Spark concepts such as RDDs, DataFrames, datasets, partitioning, caching, shuffling, joins, and window functions
  • Strong Python programming skills
  • Experience with SQL and relational databases
  • Knowledge of big data concepts and distributed data processing
  • Hands-on experience with ETL/ELT pipeline development
  • Good understanding of performance tuning and optimization techniques in Spark
  • Experience with version control tools like Git
  • Familiarity with Linux/Unix environments
  • Strong debugging and analytical skills


Benefits

Visit us at . Alignity Solutions is an Equal Opportunity Employer, M/F/V/D.
CEO Message: Click Here
Clients Testimonial: Click Here



See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available