Junior Data Engineer
NewBe an early applicantSummary
A junior data engineer role in Bangalore (on-site, full-time) building and supporting Azure cloud data pipelines: ingestion, ETL/ELT, and transformations with Azure Databricks, PySpark, SQL, and Python, using Delta Lake, Medallion Architecture, ADLS Gen2, and CI/CD via Azure DevOps. Requires 2–3 years of hands-on project experience.
Role: Junior Data Engineer
Location: Bangalore
Experience: 2–3 Years
Employment Type: Full-time
Job Summary
We are looking for a Junior Data Engineer with 2–3 years of hands-on/project experience in Azure Databricks, PySpark/Spark, SQL, Python, ETL/ELT, Delta Lake, ADLS Gen2, Azure SQL, and Azure DevOps. The candidate should have a good understanding of modern cloud data engineering, Medallion Architecture, data quality, security, and pipeline development.
Key Responsibilities
- Develop and support data ingestion and ETL/ELT pipelines.
- Perform data transformations using PySpark/Spark, SQL, and Python.
- Work with Bronze, Silver, and Gold layers using Medallion Architecture.
- Integrate data from ADLS Gen2 and Azure SQL.
- Develop and support Databricks Notebooks, Jobs, and Workflows.
- Work with Delta Lake, Unity Catalog, and basic streaming.
- Perform data cleansing, validation, and basic data-quality checks.
- Monitor pipelines, execute routine jobs, and troubleshoot basic issues.
- Assist with Azure DevOps, CI/CD, and deployment activities.
- Apply basic RBAC and security practices.
- Support basic data modeling and cloud data engineering activities.
- Document pipelines and processes and work with senior team members to understand requirements.
Required Skills
- Azure Databricks
- PySpark / Apache Spark
- SQL
- Python
- ETL / ELT
- Delta Lake
- Medallion Architecture – Bronze / Silver / Gold
- ADLS Gen2
- Azure SQL
- Databricks Notebooks / Jobs / Workflows
- Unity Catalog
- Basic streaming
- Azure DevOps & basic CI/CD
- RBAC / security fundamentals
- Basic data cleansing, validation, data quality, and data modeling
Preferred / Good-to-Have
- Academic/project experience with PySpark and SQL
- Previous Databricks projects or internships
- AZ-900 certification/fundamentals
- Familiarity with Power BI
- Exposure to other cloud platforms
- Basic understanding of data modeling