Senior Big Data Engineer (Databricks + AWS) @ SoftServe
Summary
Senior data engineer designing and maintaining scalable batch and streaming data pipelines on AWS using Databricks, PySpark, and SQL. Works across the full project lifecycle from PoC to production, with Spark, Delta Lake, Airflow/MWAA orchestration, and Kafka/MSK/Kinesis streaming.
IIn this role, you will contribute to developing scalable cloud-based data solutions on AWS using Databricks and modern big data technologies. You will work on batch and streaming data pipelines, support optimisation and modernisation initiatives, and collaborate with cross-functional teams to deliver reliable and efficient data platforms across different stages of the project lifecycle.
- 4+ years of experience in Big Data or Data Engineering
- Strong proficiency in Python (PySpark) and SQL
- Hands-on experience with AWS cloud services for data engineering solutions
- Practical experience with Databricks on AWS, including building and managing data pipelines
- Strong knowledge of Apache Spark and large-scale data processing
- Experience with batch and streaming data processing
- Familiarity with Delta Lake and Databricks components such as Workflows and Jobs
- Experience with orchestration tools such as Apache Airflow or MWAA
- Knowledge of streaming technologies such as Apache Kafka, Amazon MSK, or Kinesis
- Upper-intermediate or higher level of English
IIn this role, you will contribute to developing scalable cloud-based data solutions on AWS using Databricks and modern big data technologies. You will work on batch and streaming data pipelines, support optimisation and modernisation initiatives, and collaborate with cross-functional teams to deliver reliable and efficient data platforms across different stages of the project lifecycle.