Data Engineer
Summary
Design and develop complex data structures and ETL pipelines using Java, Hadoop, and SQL/NoSQL databases to ensure data quality and scalability.
Job Description: Design and develops complex software that processes, stores and serves data for use by others. Designs and develops complex and large-scale data structures and pipelines to organize, collect and standardize data to generate insights and addresses reporting needs.
Writes complex ETL (Extract / Transform / Load) processes, designs database systems and develops tools for real-time and offline analytic processing. Ensures that data pipelines are scalable, repeatable and secure. Improves data consistency and integrity.
Integrates data from a variety of sources, assuring that they adhere to data quality and accessibility standards. Has knowledge of large-scale search applications and building high-volume data pipelines. Knowledge of Java, Hadoop, Hive, Cassandra, Pig, MySQL or NoSQL or similar.
Complexity & Problem Solving
- Learns routine assignments of limited scope and complexity.
- Follows practices and procedures to solve standard or routine problems.
Autonomy & Supervision
- Receives general instructions on routine work and detailed guidance from more senior members on all new tasks.
- Work is typically reviewed in detail at frequent intervals for accuracy.
Communication & Influence
- Builds stable internal working relationships.
- Communicates and seeks guidance/feedback regularly from more senior members of the team.
- Primarily interacts with supervisors, project leads, mentors, or other professionals in the same discipline.
- Explains facts, policies, and practices related to discipline.
Knowledge & Experience
- Typically requires a college degree (or equivalent) with up to one year of experience but may not have any.
- Has conceptual knowledge of theories, principles, and practices within discipline and industry.