freehire launches on Product Hunt on 26 August.

Follow →

Associate Data Engineer

Summary

Build and maintain cloud-based data pipelines on AWS and Databricks, ensuring high uptime and performance while optimizing costs and security.

Requirements

  • Degree in Computer Science or Computer Engineering
  • Minimum 6 years working experience in system operations compliance and management areas
  • Project hands-on experience specifically with AWS platform (primary requirement)
  • project experience in cloud operations or cloud architecture
  • Must be cloud certified (AWS)

Core Technical Skills

  • proficiency in Databricks platform, including workspace management, cluster configuration, and job orchestration
  • Strong expertise in Apache Spark within Databricks environment, including Spark SQL, DataFrames, and RDDs
  • Good in-depth understanding of data warehouse concepts, data profiling, data verification and advanced analytics techniques
  • Strong knowledge of monitoring, incident management, and cloud cost control
  • AWS cloud services and architecture
  • IDMC (Informatica Data Management Cloud)
  • ML Ops practices within Databricks environment
  • STATA for statistical analysis is advantage
  • Amazon SageMaker integration with Databricks
  • AWS certification (Associate or Professional level) - highly preferred
  • Exposure to hospital information/clinical systems is an added advantage
  • Understanding of DevOps practices and CI/CD pipelines for Databricks-based data engineering projects

Responsibilities

Operational Responsibilities:

  • Monitor and maintain production data pipelines to ensure 99.9% uptime and optimal performance
  • Implement comprehensive logging, alerting, and monitoring systems using Application monitoring tools
  • Perform regular health checks performance, job execution times, and resource utilization to identify and resolve bottlenecks proactively
  • Manage incident response procedures for pipeline failures, including root cause analysis, resolution, and post-incident reviews
  • Establish and maintain disaster recovery procedures and backup strategies for critical data assets within the Databricks environment
  • Conduct regular performance tuning of Spark jobs and Databricks cluster configurations to optimize cost and execution efficiency
  • Maintain comprehensive documentation for operational procedures, runbooks, and troubleshooting guides
  • Coordinate scheduled maintenance windows and system upgrades with minimal business impact
  • Manage user access controls, workspace configurations, and security policies within Application environments.

Singapore -Malaysia- India - Thailand – Japan

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available