Point your AI agent at freehire and let it find you a job.

Get the CLI →

Tata Consultancy Services

NewBe an early applicant

Hadoop, Spark, Scala Data Engineer

Posted Updated
Discussion

Summary

Data engineer with 5-8 years' experience designing and maintaining large-scale ETL/data pipelines using Hadoop, Spark, Kafka, and Hive, programming in Scala/Python/Java, and working with cloud platforms (AWS/Azure/GCP) and warehouses like Redshift, BigQuery, and Snowflake. Onsite role in Bangalore, Chennai, or Hyderabad.

Desired Experience Range: 5-8 years

Location: Bangalore, Chennai and Hyderabad

Roles Responsibilities-

  • Proficiency with Big Data, Spark, HDFS, Kafka, Scala, AWS, Azure, GCP, ETL, Hadoop, Hive etc.
  • Strong experience in programming languages such as Python, Java, or Scala for data manipulation and engineering tasks.
  • Expertise in SQL and NoSQL databases
  • Hands-on experience with big data technologies like Hadoop, Spark, Kafka, and Hive to handle large-scale data processing and real-time data streams.
  • In-depth knowledge of data warehousing solutions such as Amazon Redshift, Google BigQuery, and Snowflake for building and managing data warehouses.
  • Proficiency in designing, developing, and maintaining ETL (Extract, Transform, Load) processes using tools like Apache NiFi, Talend, or Informatica.
  • Familiarity with cloud platforms like AWS, Azure, or Google Cloud for deploying, managing, and scaling data infrastructure and services.
  • Design, develop, and maintain scalable data pipelines for extracting, transforming, and loading data from various sources to ensure seamless data flow and accessibility.
  • Collaborate with cross-functional teams to integrate data from multiple disparate sources, ensuring consistency, accuracy, and reliability of data.
  • Optimize data processing workflows and storage solutions for performance, scalability, and cost-efficiency
  • Implement data quality checks and validation processes to ensure the accuracy, completeness, and consistency of data throughout the data lifecycle.
  • Monitor the performance of data pipelines and infrastructure, identifying and resolving issues to maintain system stability and reliability.


Skills

See also

Data Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available