Data Engineer — Databricks (2-5 Years Experience)
Summary
Data engineer (2-5 yrs experience) at Sublime Data, a Databricks partner consulting firm in Ahmedabad, India. You'll design, build, and maintain client data pipelines using Databricks (PySpark, Delta Lake, Delta Live Tables, Workflows), support cloud data migrations (Azure primary), and work directly with clients in this in-office, client-facing role.
About Us
Sublime Data is a boutique data engineering and AI consulting firm, Databricks partner, building specialized expertise in data platform modernization and agentic AI implementation for clients.
What You'll Do
- Design, build, and maintain data pipelines using Databricks (PySpark, Delta Lake, Delta Live Tables, Workflows)
- Work directly with client stakeholders to understand requirements, provide updates, and troubleshoot issues — this is a client-facing role, not purely back-office delivery
- Participate in technical discovery and client interviews as part of our presales and onboarding process
- Collaborate with our architecture team on solution design for new client engagements
- Support data migration and modernization projects across cloud platforms (Azure primary, AWS/GCP as needed)
- Travel to client locations as required for project kickoffs, workshops, or on-site delivery phases
What We're Looking For
- 2-5 years of total data engineering experience, with at least 2 years of hands-on Databricks experience (PySpark, SQL, Delta Lake)
- Databricks certification (Data Engineer Associate/Professional) is a strong plus, not mandatory
- Strong SQL and Python fundamentals
- Exposure to at least one other data platform — Snowflake, Microsoft Fabric, or native AWS/GCP data services — is a genuine advantage
- Comfortable communicating directly with clients — confident in explaining technical work to non-technical or semi-technical stakeholders
- Ability to join within 15 days or immediately
- Based in or willing to relocate to Ahmedabad — this is an in-office role, not remote
- Willingness to travel for client engagements as needed
Good to Have
- Experience with Unity Catalog, MLflow, or Databricks Workflows specifically
- Prior client-facing consulting or services company experience (vs. pure product company background)
- Exposure to any GenAI/LLM tooling (LangChain, RAG pipelines) — increasingly relevant to our engagements
- Experience in BFSI, healthcare/pharma, or energy sector data projects