Junior Data Engineer
Summary
Build and maintain data pipelines, databases, and APIs to support analytics and machine learning models for an HR-tech platform.
Early regularization: get regularized as early as your 3rd month!
Paid trainings and certifications
Paid leaves upon hire
Fun and innovative culture - we love getting things done while also having fun!
DIRECTLY REPORTS TO:
Head of Data Science
MAIN AREA OF RESPONSIBILITY:
The primary responsibility of the Junior Data Engineer is to transform data into a format that can be easily analyzed. He/She will be responsible for designing and developing Sprouts data system, processing and extracting data features and deploying the data science teams machine learning models. He/She will support our software engineers, architects and data scientists on data initiatives and will ensure optimal data delivery architecture is consistent throughout ongoing projects.
- Build, optimize and maintain conceptual, logical and physical database models
- Develop database solutions to store and retrieve information
- Assemble datasets that meet functional/non-functional business requirements
- Monitor data integrity and adopt appropriate tools
- Improve system performance
- Suggest optimizations or give recommendations on data architecture to support Sprouts next generation of products and data initiatives
- Design, develop, test and deploy web service APIs
- Work with Data Scientists to identify future needs and requirements
- Deploy models and algorithms developed by the Data Science team
- Perform other duties as assigned by the company
QUALIFICATIONS | COMPETENCIES:
- Knowledge on databases (SQL and/or NoSQL) and data engineering best practices
- SQL and other programming languages(e.g. Python, Java, Scala, shell scripting etc.)
- Experience with data modeling (data warehouse, data lake) and designing data storage schemes
- Familiarity with data engineering and ETL software tools, hadoop, spark, talend, SSAS, etc. is also helpful
- Experience building and optimizing data pipelines, architecture and datasets
- Experience with Azure
- A successful history of manipulating, processing and extracting value from large disconnected datasets is a plus
- Familiarity with data visualization tools (e.g. PowerBI or Google Data Studio)
- Familiarity with agile development as a project management methodology is a plus
- Strong problem-solving and analytical skills
- Must be self-motivated and comfortable supporting the data needs of multiple teams, systems and products
- A good team player and willingness to learn
- Strong innate desire and proven track record of continuous self-improvement (in learning, job expansion, extracurricular activities, etc.)
Sprout Solutions provides equal Opportunity Employment and Welcomes applications from all sectors of the society. Discrimination on the basis of race, religion, age, nationality, ethnicity, gender, citizenship, civil partnership status, or any other grounds as protected by law.