AI/Data Engineer for Unstructured Data & RAG Pipelines

Summary

Virtusa is hiring an AI/ML-focused data engineer in Dubai to build intelligent data pipelines for unstructured content using PySpark and Python, covering document classification, cleansing, and quality metrics, and integrating LLMs, vector databases, and RAG frameworks to enable AI-first applications.

Virtusa is seeking an AI/ML-focused Data Engineer to build intelligent data pipelines for unstructured content and integrate with modern ML ecosystems. The role emphasizes PySpark and Python, with a focus on document classification, cleansing, quality metrics, and working with LLMs, vector databases, and RAG frameworks.

You will bridge data engineering and machine learning to enable AI-first applications, collaborating with AI architects and platform teams to design end-to-end AI data readiness

See also

Data Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available