freehire launches on Product Hunt on 26 August.

Follow →

Senior AI Data Engineer

Summary

Senior AI Data Engineer builds and maintains AI-ready knowledge bases and pipelines for enterprise AI solutions, copilots, and RAG systems using Python, Postgres + pgvector, and unstructured data sources.

More about us

Billennium is a global technology company with over 20 years of experience, committed to innovation and empowering businesses. As an employer, we offer a supportive, growth-focused environment where collaboration and creativity thrive. Join us to shape the future of technology together!

About the Role

We are looking for an experienced Senior AI Data Engineer who will be supporting the development of enterprise AI solutions, copilots, and RAG-based systems. In this role, you will transform complex and unstructured business data into high-quality, AI-ready knowledge assets that power trustworthy, scalable, and efficient AI applications.

What you will do

  • Lead enterprise data discovery, cleanup, and standardization initiatives across SharePoint, document repositories, databases, and data lakes.

  • Design and build robust ingestion pipelines for unstructured and semi-structured data, including metadata extraction, normalization, and enrichment.

  • Create and maintain AI-ready knowledge bases optimized for retrieval, embedding, and RAG use cases using Postgres + pgvector.

  • Define and implement data quality frameworks, evaluation processes, and feedback loops to continuously improve retrieval accuracy and trust.

  • Ensure governance, traceability, and privacy compliance through data lineage, auditing, and PII detection/masking practices.

What we are looking for

  • 5+ years of experience in Data Engineering, Applied Data Science, or Analytics Engineering with ownership of production-grade data pipelines.

  • Strong Python skills and hands-on experience with data processing, parsing, cleaning, normalization, and ETL/ELT workflows.

  • Proven experience working with unstructured enterprise content such as PDFs, Office documents, knowledge bases, wikis, and SharePoint repositories.

  • Practical experience building data pipelines for NLP, search, retrieval, or RAG-related use cases, including metadata management and corpus preparation.

  • Excellent stakeholder management skills with the ability to collaborate across business, architecture, and engineering teams to define data quality standards and AI readiness

Perks and benefits (our offer)

  • Comprehensive benefits - enjoy Udemy for Business, private medical care, Multisport card, veterinary package, language lessons, and shopping vouchers.

  • Flexibility - adaptable working hours and remote/hybrid work options to suit your lifestyle & location.

  • Career growth - access opportunities for professional development and learning, including perks related to our official partnerships with global IT giants: Microsoft, AWS, Snowflake, Salesforce & more.

  • Global collaboration - work with a diverse, international team.

  • Innovative environment - be part of a forward-thinking and growth-oriented workplace.

  • Engaging community – Work with passionate professionals and participate in team-building events, hackathons, and CSR initiatives to make an impact beyond work.

  • Team-building events including our company tradition (annual company event in Mazury).

  • A pleasant surprise to start your journey with us in the form of a welcome pack.

Recruitment process

  • HR call

  • Technical Interview

  • Interview with the dedicated Client

  • Final decision / Feedback

Sounds interesting? Click "Apply" and have a chance to hear more!

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available