Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Design and build scalable Azure data pipelines and Lakehouse architectures using ADF, Synapse, PySpark, and Delta Lake, ensuring secure, governed data flows for enterprise clients.
Build and scale a modern data platform using Databricks and Azure Data Factory to power advanced analytics, ML models, and near-real-time decision systems.
Builds and optimizes PySpark pipelines in Python/SQL to power scalable data workflows that enhance the e-commerce personal-shopping experience.
Builds scalable data pipelines and modern analytics platforms using Databricks, Spark/PySpark, and cloud tech to process and analyze large datasets.
Data Engineer building ETL pipelines, data models, and dashboards for a multinational insurer using SQL, Python, PySpark, and Power BI.
Builds and maintains data pipelines on Azure and Microsoft Fabric, ingesting batch/streaming data with PySpark/Databricks and orchestrating workflows via Data Factory and ADX.
Build and optimize scalable data pipelines on Databricks using Spark/PySpark and Medallion architecture for modern cloud-based data platforms.
Build and maintain data ingestion and transformation pipelines using AWS, S3, Redshift, PySpark, and SQL to process structured sources and files.
Build and maintain cloud data platforms on AWS, using Spark/PySpark, Python, and SQL to transform raw data into insights for clients.
Design and optimize scalable data pipelines using Apache Spark and AWS for batch and near-real-time processing, creating reliable datasets for analytics and ML.
Designs and builds scalable Azure-based data pipelines using PySpark, Delta Lake, and Databricks to move and transform data efficiently across cloud systems.
Designs and builds scalable Python-based ETL pipelines using PySpark and AWS Glue to move and transform large datasets.
Build and maintain Python-based ETL pipelines using PySpark and AWS Glue to process large datasets, with Docker and CI/CD for reliability.
Lead end-to-end ML/AI model development and MLOps at a construction and infrastructure firm, setting standards for reproducibility, validation, and responsible AI while mentoring junior scientists.
Build and maintain scalable data pipelines using PySpark, Python, and Azure Data Factory to ingest and transform data for reliable analytics and reporting.
Design and optimize SQL-based ETL pipelines for financial data using SSIS, AWS Glue, and Python, ensuring compliance and performance in a regulated environment.
Build and maintain Databricks pipelines, deploy ML models with MLflow, and create Power BI dashboards to support analytics and reporting in an AI-focused environment.
Designs and builds scalable Azure-based data pipelines and lakehouse solutions using Databricks, PySpark, and Azure Data Factory to power analytics and reporting for enterprise clients.
Build and optimize ETL pipelines using Azure Synapse Analytics and PySpark to move and transform data for clients.
Build and optimize PySpark-based ETL pipelines on Databricks, refactoring legacy code and ensuring scalable, high-quality data flows for enterprise clients.
We couldn't check your fit for this role — add a CV to your profile to see it next time.