Software Engineer [Multiple Positions Available]
Summary
This role involves designing and implementing large-scale data pipelines and ETL processes using Python, Spark, and Talend to support enterprise data systems. The engineer will work within the SDLC to manage data integration, cloud migration to AWS, and database optimization across various platforms.
DESCRIPTION:
Duties: Design, develop and implement software solutions. Solve business problems through innovation and engineering practices. Involved in all aspects of the Software Development Lifecycle (SDLC) including analyzing requirements, incorporating architectural standards into application design specifications, documenting application specifications, translating technical requirements into programmed application modules, and developing or enhancing software application modules. Identify or troubleshoot application code- related issues. Take active role in code reviews to ensure solutions are aligned to pre- defined architectural specifications. Assist with design reviews by recommending ways to incorporate requirements into designs and information or data flows. Participate in project planning sessions with project managers, business analysts, and team members to analyze business requirements and outline proposed solutions.
QUALIFICATIONS:
Minimum education and experience required: Master's degree in Information Systems, Computer Engineering, Computer Science, Information Technology, or related field of study plus 2 years of experience in the job offered or as Software Engineer, Senior Software Engineer, or related occupation. The employer will alternatively accept a Bachelor's degree in Information Systems, Computer Engineering, Computer Science, Information Technology, or related field of study plus 4 years of experience in the job offered or as Software Engineer, Senior Software Engineer, or related occupation.
Skills Required: This position requires two (2) years of experience with the following: Designing and implementing large data pipelines for ETL; Utilizing Python to design ETL pipelines that ingest data from multiple sources, applying business-rule transformations, and loading data into Salesforce to integrate databases, APIs, and cloud platforms to build unit tests and validation scripts; Utilizing Shell scripts to automate file operations, monitor system resources, log ETL job statuses, and send alerts in case of failures; configuring environments, managing permissions, and handling pre- and post-processing steps in ETL workflows; Utilizing Talend Studio for maintaining, auditing, and accurately migrating legacy ETL pipelines into PySpark-based architectures ensuring data integrity and consistency during the decommissioning and modernization of enterprise-scale data systems; Utilizing Spark framework to architect scalable, distributed data architectures that execute high-performance transformations and aggregations on massive datasets across both real-time streaming and batch processing environments; Designing data models (Dimensional Modeling, Relational Modeling), data integration, data security, and metadata management to support analytics and business intelligence; Working on data loads and extracts from Salesforce objects using SOQL and ETL processes to ensure effective data handling; Utilizing JIRA to track and manage software development tasks; Ensuring software quality through comprehensive testing including functional, manual, performance, regression, smoke, system integration, unit, and user acceptance testing; using Waterfall and Agile SDLC methodologies to provide a structured approach to project execution; Developing and optimizing SQL and PySpark queries across multiple data sources; utilizing advanced window functions, recursive CTEs, and multi-stage transformations to efficiently process large-scale datasets and ensure high- performance data retrieval for business-critical intelligence; implementing database solutions using SQL on MS SQL Server, Oracle 11g, PostgreSQL, and Cassandra to enhance data management and access; Building robust applications and automation scripts using Python to improve software functionality and operational efficiency including extracting and analyzing data from REST APIs; maintaining data warehousing solutions and data marts for efficient storage and retrieval using Erwin Data Modeler, Talend Open Studio, Python, and Cassandra; Incorporating data governance principles including data quality, security, and compliance into design specifications using Python Libraries; Leveraging command line interfaces including IntelliJ CLI and Putty for automated development tasks, efficient cloud management, and secure Linux server access. This position requires any amount of experience with the following: Automating data ingestion and processing pipelines in cloud; Cloud services including AWS for migrating on-premise applications to AWS cloud; Leveraging AWS for scalable data storage as a processing solution, and managing large datasets; Leveraging command line interfaces including AWS CLI for automated development tasks, efficient cloud management, and secure Linux server access; handling formats including Parquet and JSON; Managing source code using GIT for version control and CI/CD practices; maintaining data warehousing solutions and data marts for efficient storage and retrieval using Erwin Data Modeler, Talend Open Studio, Python, and Cassandra; Leveraging AWS for scalable data storage as a processing solution, and managing large datasets.
Job Location: 575 Washington Boulevard, Jersey City, NJ 07310.
We offer a competitive total rewards package including base salary determined based on the role, experience, skill set, and location. For those in eligible roles, discretionary incentive compensation which may be awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. In addition, please visit: https://careers.jpmorgan.com/us/en/about-us.
Full-Time. Salary: $167,500 - $215,000 per year.