JR-193248 Data Engineer Senior
Summary
Senior Data Engineer who builds and maintains code-based ETL pipelines for clinical/patient-level data, monitors CI/CD processes, and delivers outputs to AWS S3 for client access. Core stack: Python, SQL/Postgres, AWS (Redshift, Lambda, S3, Glue), and Git-based collaboration.
Project Description:
Job Responsibilities:
Responsibilities
- Create ETL pipelines in code and link them up to internal tooling
- Monitor CI/CD processes to ensure ETL pipelines continue to run smoothly
- Coordinate with clinical data analysts to queue patient-based data extraction
- Collaborate with clinical data analysts to ensure the data meets quality standards
- Handle ad-hoc requests from CSS folks about the patient-level data
- Package and deliver the outputs of ETL pipelines to AWS S3 for client access
- Context switch between multiple ETL pipelines and CSS stakeholders
- Make recommendations on code and process efficiency improvements
- Collaborate effectively with other engineers through the use of Git and merge requests (or similar version control tools)
Requirements:
Qualifications
- Bachelor's degree or equivalent experience in Computer Science or related field
- Development experience with programming languages
- Experience with SQL database or any relational database skills, Postgres preferred
- 4+ years of experience with Python Experience CICD, Redshift and AWS environment (Lambda, S3 and Glue)
- Experience with at least one ETL tool
- Experience with data to debug ETL pipelines
- Experience communicating effectively with non-technical stakeholders
- Advanced English Level (all interviews will be conducted in English)
Additional Comments:
