Junior Data Engineer (Azure Databricks)
Summary
The Junior Data Engineer will develop and maintain modern data pipelines and Lakehouse architectures using Azure Databricks, PySpark, and Azure Data Factory. The role involves integrating data from multiple sources, optimizing data transformations, and supporting SQL Server-based data warehouse environments.
We are looking for a Junior Data Engineer specialized in Azure Databricks to join our data platform team.
The candidate will develop and support modern data pipelines and Lakehouse architectures, leveraging Azure Databricks, Spark, and Azure Data Factory, while integrating with existing SQL Server-based data warehouse environments. The role offers an opportunity to learn and grow within a modern cloud data ecosystem while contributing to analytics and business intelligence initiatives.
Key responsibilities:
- Support and evolve existing SQL Server data platforms and ETL solutions
- Develop and maintain data pipelines using Azure Databricks
- Build and optimize data transformations using PySpark and SQL in Databricks
- Develop ETL/ELT pipelines orchestrated through Azure Data Factory
- Integrate data from multiple sources into the data platform and analytical layers
- Maintain data models and data warehouse structures for analytics
- Ensure data quality, scalability, and performance of large-scale data processing pipelines
- Collaborate with BI teams to support Power BI and reporting platforms
- Participate in the implementation of modern cloud-based data architectures under guidance from senior team members
- 2+ years of experience in Data Engineering, Data Warehouse development, or a related field
- Experience with Azure Databricks
- Experience developing data pipelines using PySpark and Spark SQL
- Understanding of distributed data processing and big data concepts
- Good SQL skills and experience with SQL Server relational databases
- Experience building data pipelines using Azure Data Factory
- Exposure to data processing and performance optimization concepts
Nice to have
- Knowledge of Spark optimization techniques (partitioning, caching, cluster tuning)
- Familiarity with Power BI and Databricks Structured Streaming
- Experience in migrating from traditional ETL process to cloud architectures
- Familiarity with Delta Lake and Lakehouse architectures
Soft Skills
- Strong analytical and problem-solving skills
- Good communication and collaboration skills
- Team player with a proactive mindset
- Eager to learn and continuously develop technical expertise
- Adaptable and comfortable working in dynamic environments
We offer:
- Culture of relentless performance: join an unstoppable technology development team with a 99% project success rate and more than 30% year-over-year revenue growth.
- Competitive pay and benefits: enjoy a comprehensive compensation and benefits package, including health insurance, language courses, and a relocation program.
- ForeverRemote work culture: make the most of the flexibility that comes with remote work.
- Growth mindset: reap the benefits of a range of professional development opportunities, including certification programs, mentorship and talent investment programs, internal mobility and internship opportunities.
- Global impact: collaborate on impactful projects for top global clients and shape the future of industries.
- Welcoming multicultural environment: be a part of a dynamic, global team and thrive in an inclusive and supportive work environment with open communication and regular team-building company social events.
- Social sustainability values: join our sustainable business practices focused on five pillars, including IT education, community empowerment, fair operating practices, environmental sustainability, and gender equality.
* Miratech is an equal opportunity employer and does not discriminate against any employee or applicant for employment on the basis of race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, gender identity, or any other protected status under applicable law.