freehire launches on Product Hunt on 26 August.

Follow →

Data Engineer

Open 35d

Konovo is a global healthcare intelligence company on a mission to transform research through technology- enabling faster, better, connected insights.

Konovo provides healthcare organisations with access to over 2 million healthcare professionals—the largest network of its kind globally. With a workforce of over 200 employees across 5 countries: India, Bosnia and Herzegovina, the United Kingdom, Mexico, and the United States, we collaborate to support some of the most prominent names in healthcare. Our customers include over 300 global pharmaceutical companies, medical device manufacturers, research agencies, and consultancy firms.

As we transition from a service-oriented model to a product-driven platform, we are expanding our hybrid Bengaluru team. We are looking for an experienced data engineer to contribute to our mission by deep-diving into business problems, building out our data lakehouse and contributing to our advanced analytics. As a Konovo engineer, you will unlock the power of our data, learn a breadth of technologies and aspects of data engineering, help us optimize and solve some of our greatest challenges!

How You'll Make an Impact

  • Build and own production-grade batch ELT pipelines in Databricks/Spark (15-minute cadence during working hours; hourly off-hours)
  • Design, scale, and operate our lakehouse using medallion patterns (Bronze/Silver/Gold), including incremental loads, backfills, and late-arriving data handling
  • Create curated, well-documented data products that are reusable across teams and trusted for decision-making
  • Improve reliability through automated testing, monitoring/alerting, and clear SLAs for critical pipelines and tables
  • Partner with engineering and business stakeholders to translate needs into durable data models (not one-off extracts)
  • Ensure data governance and security expectations are met through disciplined implementation (access controls, lineage, and quality standards)

What We're Looking For

  • 5+ years building and operating production data pipelines (data engineering—not primarily BI/analytics or data science)
  • Strong SQL plus strong data modeling skills (dimensional and/or lakehouse modeling)
  • Hands-on Databricks + Spark experience in production (debugging, performance tuning, cost awareness)
  • Experience building batch ELT at frequent cadence (e.g., 15-minute schedules), including idempotency, backfills, and late-arriving data patterns
  • Experience with orchestration and transformation tooling (e.g., Airflow + dbt) and modern development practices (version control, CI/CD)
  • Strong data quality discipline: automated tests/expectations, monitoring/alerting, and clear SLAs for critical datasets
  • Ownership mindset: you build it, you run it (triage, incident response, continuous improvement)
  • Clear written and verbal communication for engineering work (requirements clarification, design docs, tradeoffs)
  • Comfortable working hybrid in Bengaluru (Whitefield) and overlapping 3–4 hours with EST

Bonus Points:

  • Deep Delta Lake experience (schema evolution, OPTIMIZE/Z-ORDER, compaction, partitioning strategy)
  • CDC ingestion patterns (e.g., Debezium/Fivetran/HVR or custom CDC) and handling late/out-of-order events
  • Streaming or near-real-time pipelines (Structured Streaming, Kafka, Auto Loader) even if the core role is batch
  • Strong observability practices for data systems (metrics, lineage, data contracts, incident postmortems)
  • Cost and performance optimization in Databricks (cluster sizing, job tuning, Photon, caching strategies)
  • Experience building governed data products for multi-tenant consumption (RBAC, PII handling, auditability)
  • Exposure to healthcare, life sciences, or market research data and related compliance considerations

Why Join Konovo?

  • Be part of a mission-driven organization that is shaping the future of healthcare decision-making.
  • Join a fast-growing global team with opportunities for professional growth and advancement.
  • Enjoy a collaborative and hybrid work environment that fosters innovation and flexibility.
  • Experience a workplace that puts employees first, offering a workplace designed for growth, well-being, and balance.
  • Become a part of an organization that prioritizes your well-being with comprehensive benefits, including group medical coverage, accident insurance, and a robust leave policy.
  • Make a real-world impact by helping healthcare organizations innovate faster.

If you’re eager to build a meaningful career in healthcare market research — and you’re smart, nice, and driven to make an impact — we’d love to hear from you!

Apply now to be part of our journey.

What this application asks

greenhouse

First Name, Last Name, Email, Phone, Resume/CV, Cover Letter

  • Preferred First Name optional
  • LinkedIn Profile optional
  • Website optional
  • How did you hear about this job?
  • Are you comfortable working in a hybrid mode, 3 days a week, from our Whitefield, Bengaluru office? * choose one
  • Are you willing to work in a UK shift (2 PM IST to 11 PM IST)? choose one
  • How many years of experience do you have building and operating production data pipelines (not primarily BI or data science)? choose one
  • How would you describe your hands-on Databricks/Spark experience? choose one
  • Which best describes your experience with frequent-cadence batch ELT (e.g., 15-minute schedules)? choose one
  • Select all the data engineering technologies you have hands-on experience using. choose any
  • Describe a production data pipeline you owned end to end. What do you think worked well in the architecture, and what could have worked better? written answer
  • What is your current CTC?
  • What is your expected CTC?
  • What is your notice period (in days)?
  • Processing of Personal Data* choose any
  • Are you currently legally authorized to work in the country in which this job is based (e.g. you are a citizen, you have a visa, etc.)? choose one