Infrastructure Engineer II
With your expertise in delivering infrastructure solutions, you are a top-performer in your field. Come on board as a highly appreciated member of a winning team.
As an Infrastructure Engineer II at JPMorganChase within the Corporate Sector – Infrastructure Platforms, you develop knowledge of software, applications, and technical processes within the infrastructure engineering discipline. Through this work you begin to apply your proficiency in a single application or technical methodology.
Job responsibilities
Applies technical knowledge to assignments with a defined scope, including testing infrastructure performance, verifying requirements were met, and collecting and analyzing monitoring data in test and production through completion
Uses enterprise-authorized AI capabilities within the work environment to accelerate infrastructure analysis and documentation (e.g., summarizing monitoring findings), validating outputs and handling data according to sensitivity and security requirements.
Help improve and evolve our web infrastructure platforms (Apache, Tomcat, IIS, WebSphere/IHS) by applying strong infrastructure engineering fundamentals, SRE practices, observability, and AI-enabled automation to deliver stable, secure, and scalable services.
Working in a highly motivated, Agile team, you will collaborate with customers and partner teams to deliver platform improvements, modernize technology processes, and evolve web infrastructure products in alignment with broader platform strategy and reliability goals across JPMorganChase's global technology network
Demonstrate an AI-first mindset by building hands-on agentic automation for operational workflows (e.g., triage, runbook execution, incident summarization, change validation) with appropriate controls and monitoring.
Demonstrate and champion site reliability culture and practices within your team; contribute to initiatives to improve reliability and stability using data-driven analytics to improve service levels.
Collaborate with team members, technical experts, key stakeholders, and customers to identify service level indicators and help establish reasonable service level objectives and error budgets.
Diagnose and implement changes to resolve issues, modernize technology processes, and ensure end-to-end monitoring of applications.
Collect and analyze monitoring/telemetry data in test and production environments; design and implement dashboards; ensure system reliability, performance, and security.
Escalate issues appropriately with detailed technical write-ups; partner with application and infrastructure teams to identify and remediate capacity risks; understand platform interdependencies and limitations.
Required qualifications, capabilities, and skills
Formal training or certification on infrastructure engineering concepts and 2+ years applied experience
Working knowledge of using enterprise-authorized AI capabilities within the work environment to support infrastructure engineering workflows with strong validation habits and awareness of data sensitivity.
Ability to review and validate AI-assisted recommendations before use, escalating when uncertain and following security and data handling expectations.
Proficiency in scripting/automation and infrastructure-as-code (Python, PowerShell, Ansible, Terraform) and fluency in at least one programming language (e.g., Python, Go, Shell Scripting, .NET, etc.), including hands-on experience applying AI-assisted automation and agentic patterns to engineering/operational workflows.
Proven hands-on ability to leverage GitHub Copilot and coding assistants to accelerate skill development and build AI agents/workflows that support specific infrastructure operations tasks (e.g., triage, automation, runbook execution).
Knowledge of infrastructure engineering areas such as operating systems (Linux/Windows), networking terminology and protocols, databases, deployment practices, and automation.
Familiarity with Web products like Apache, Tomcat, IIS, WebSphere/IHS.
Proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and similar best practices, with ability to implement these practices within a platform.
Experience in observability, monitoring, alerting, and telemetry collection using tools such as Grafana, Dynatrace, Datadog, Prometheus, CloudWatch, Splunk, etc., including designing and implementing monitoring dashboards using Splunk or Dynatrace with effective production support.
Preferred qualifications, capabilities, and skills
Implementation of CI/CD pipelines, code reviews using GitHub, and process automation with Python and scripting.
Hands-on experience and certifications in AWS, Azure, GCP, or other cloud environments; AWS/Azure exposure with understanding of resiliency, scalability, observability, and monitoring.
Experience utilizing Terraform or other infrastructure-as-code technologies for cloud resource management.
Experience supporting complex and mission critical applications involving a multitude of components of varying technical generations; Familiarity with modern front-end technologies and experience in designing, developing, and implementing software solutions.
Strong drive to expand infrastructure engineering knowledge across emerging technologies and domains; drive to self-educate and evaluate new technology.