Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Design and scale enterprise-grade Databricks Lakehouse platforms using Spark, Delta Lake, and cloud ecosystems. Drive architecture strategy and data modernization initiatives.
Чем предстоит заниматься: Выполнение задач полного жизненного цикла (от сбора требований до выкатки на установку в прод) Разработка/доработка витрин/решений и их автоматизация через оркестраторы Изменение алгоритма…
• Strong hands-on experience with PySpark and Python for data engineering. • Experience developing ETL pipelines using AWS Glue. • Proficiency with AWS Step Functions for workflow orchestration. • Experience building…
A Data Science Engineer role for a fresher, working remotely on data preparation, exploration, feature engineering, machine learning modeling, and automation using Python3, Numpy, Pandas, Scikit-learn, SQL, and PySpark.
Builds and maintains big data pipelines, processes large-scale data using Hadoop/Spark, and develops/optimizes ETL workflows with Python/SQL. Supports existing systems and implements new requirements in a remote Agile environment.
Builds and optimizes data pipelines using PySpark, SQL, and the Hadoop ecosystem to process large datasets and support business analytics.
Develops data pipelines and assets using Python, PySpark, and SQL to drive supply chain analytics and manage digital transformation projects.
Data Scientist III at Banco Bradesco, working in the BI and Corporate Security area to analyze large datasets and develop advanced machine learning models for fraud and financial crime prevention.
As Data Scientist III - Estratégia e Gestão de Crédito at Banco Bradesco, you analyze credit portfolio risk indicators, develop data-driven studies on market/regulatory impacts, and support strategic decisions using SQL, Python/PySpark, and Databricks.
Develops and validates AI/Generative AI pipelines in Databricks, ensuring data quality, governance, and performance for risk-modeling solutions in a fintech environment.
Design and validate data pipelines and GenAI solutions on Databricks using Python, Spark, and SQL. Manage model risk and data governance, implementing MLOps/LLMOps practices in a hybrid environment.
Principal Software Engineer at JPMorganChase building and scaling a global KYC and risk data platform, leading AI-driven engineering practices and agentic systems in a regulated financial environment.
Build and maintain data pipelines, ETL/ELT processes, and cloud-based data lakes/warehouses using Azure and Databricks to deliver clean, secure data for analytics and AI.
Design, develop, and maintain scalable data pipelines and workflows on Azure Databricks using PySpark, Azure Data Factory, and Apache Spark for advanced analytics and data engineering.
Junior Data Engineer helping build and maintain scalable ETL/ELT pipelines and database infrastructure to support AI and advanced analytics solutions, primarily using Python and SQL.
The Data Engineer will design and maintain API-driven data pipelines, data lakes, and warehouses to support AI initiatives and business intelligence. The role involves integrating enterprise systems like Salesforce and Oracle using AWS services, Python, and Spark.
Hands-on data engineer building production data pipelines and agentic AI components (RAG, tool-calling agents) for clinical and non-clinical data at Lilly, using Python, Spark/PySpark, Databricks, and LLM frameworks in a regulated pharma environment.
Senior Data Engineer leading data engineering delivery for data lakes, platform modernization, and reporting solutions at Manulife in Hong Kong, using Azure Cloud, Databricks/Spark, and AliCloud big data tools.
Senior Data Engineer builds and maintains scalable ETL/ELT pipelines and lakehouse datasets to power analytics, ML, and business applications at Hertz.
We couldn't check your fit for this role — add a CV to your profile to see it next time.