freehire launches on Product Hunt on 26 August.

Follow →

Lead Data Engineer (Contract, Full-Time) [HR208]

About Smart Working

At Smart Working, we believe your job should not only look right on paper but also feel right every day. This isn’t just another remote opportunity - it’s about finding where you truly belong, no matter where you are. From day one, you’re welcomed into a genuine community that values your growth and well-being. Our mission is simple: to break down geographic barriers and connect skilled professionals with outstanding global teams and products for full-time, long-term roles. We help you discover meaningful work with teams that invest in your success, where you’re empowered to grow personally and professionally.

Join one of the highest-rated workplaces on Glassdoor and experience what it means to thrive in a truly remote-first world.

About the Role

As a Lead Data Engineer, you will build and lead the data infrastructure powering an intelligent AI assistant that automates operational workloads for property managers, letting agents and build-to-rent teams. The platform manages communications, compliance, maintenance coordination and scheduling, helping property businesses operate with the efficiency of modern digital platforms.

You will architect and scale the systems behind the AI products, spanning real-time data pipelines, analytics infrastructure, vector databases and machine learning data workflows. Working closely with AI engineers, backend engineers, product teams, founders and product leadership, you will ensure the platform can process large volumes of operational data reliably and intelligently.

As the first senior data hire, you will define the data architecture, tooling, engineering standards, culture and hiring bar, building the foundations of the future data team.

Responsibilities

  • Architect and build scalable data pipelines and infrastructure to support AI and product systems.
  • Design and maintain data ingestion, transformation and storage architectures for operational and AI workloads.
  • Develop and manage batch and real-time data pipelines.
  • Build and optimise systems for vector search, retrieval and machine learning data pipelines.
  • Ensure data reliability, security and governance across the platform.
  • Collaborate with AI and backend engineering teams to support training, inference and product features.
  • Implement monitoring, observability and data quality frameworks.
  • Optimise the performance of large-scale datasets and query systems.
  • Contribute to technical architecture decisions and long-term data strategy.
  • Act as the founding data hire, defining culture, standards and the hiring bar for the data function as it scales.
  • Partner directly with founders and product leadership to translate data capabilities into product decisions.

Requirements

  • 7+ years of professional experience, with the majority of that experience in dedicated data engineering roles.
  • Strong experience designing and building data pipelines and distributed data systems.
  • Experience working with relational databases, with PostgreSQL preferred, although MySQL or similar is acceptable.
  • Experience working with NoSQL databases.
  • Experience with vector databases used in modern AI systems.
  • Strong programming experience in Python.
  • Demonstrated ability to make and justify architectural decisions, rather than only implementing them.
  • Experience building scalable backend systems.
  • Experience designing data models and storage architectures.
  • Strong understanding of data processing performance and optimisation.
  • Experience with some of the following data frameworks and infrastructure technologies is highly desirable: Apache Spark, Apache Airflow, Kafka, and Elasticsearch or OpenSearch.
  • Experience with relevant database technologies is highly desirable, including PostgreSQL, MongoDB, and vector databases such as Qdrant, Milvus or pgvector.
  • Experience with Python data-processing libraries such as Pandas or Polars is highly desirable.

Nice to Have

  • Experience working on AI or machine learning platforms.
  • Familiarity with stream processing and event-driven architectures.
  • Experience with cloud infrastructure such as GCP, AWS or Azure.
  • Experience working in high-growth startups or early-stage companies.

What this application asks

lever

In which location did you find the job?, Resume/CV, Full name, Email, Phone, Current location, Current company, LinkedIn URL, Twitter URL, GitHub URL, Portfolio URL

  • Please respond truthfully. How many years of professional experience do you have, with the majority of that experience in dedicated data engineering roles? choose one
  • Please respond truthfully. What level of professional experience do you have designing and building scalable data pipelines and distributed data systems? choose one
  • Please respond truthfully. What level of professional experience do you have designing data ingestion, transformation and storage architectures for operational and AI workloads? choose one
  • Please respond truthfully. What level of professional experience do you have developing and managing both batch and real-time data pipelines? choose one
  • Please respond truthfully. What level of professional experience do you have building and optimising vector search, retrieval and machine learning data pipeline systems? choose one
  • Please respond truthfully. What level of professional experience do you have using Python for data engineering? choose one
  • Please respond truthfully. What level of professional experience do you have with relational databases such as PostgreSQL, MySQL or similar, including designing data models and storage architectures? choose one
  • Please respond truthfully. What level of professional experience do you have with NoSQL databases such as MongoDB? choose one
  • Please respond truthfully. What level of professional experience do you have ensuring data reliability, security and governance, including implementing monitoring, observability and data quality frameworks? choose one
  • Please respond truthfully. What level of professional experience do you have making and justifying data architecture decisions and contributing to long-term data strategy? choose one
  • Please respond truthfully. What is your level of spoken and written English? choose one
  • Are you comfortable working the fixed shift hours of 12 PM – 9:30 PM IST (Summer) and 1 PM – 10:30 PM IST (Winter), Monday to Friday? choose one
  • We are looking for candidates who can start within the next 30 days. ⚠️ Providing false or misleading information will result in disqualification. Candidates progressing to the next stage will be required to upload proof of their notice period or last working day before meeting the client. When is your earliest realistic start date? choose one
  • What is your current salary in LPA? We will cross-check references later in the process. choose one · optional
  • What is your minimum expected salary (LPA, CTC)? choose one · optional
  • Are you open to negotiation? choose one
  • Do you acknowledge that submitting more than one application for the same role, or providing any false or misleading information, will automatically and permanently disqualify you from the recruitment process? choose one

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available