Point your AI agent at freehire and let it find you a job.

Get the CLI →

JPM Unternehmensgruppe

NewBe an early applicant

Senior Lead Software Engineer - Performance and Resiliency Engineering

Posted Updated
Discussion

Summary

Performance & Resiliency Engineering Lead (SVP/ED) for Merchant Services at JPMorgan: drives latency reduction, throughput/TPS gains, capacity planning, and peak-event readiness on high-volume payment platforms. Core stack is AWS with Kubernetes/EKS, observability tooling, and resiliency/progressive-delivery patterns.

Performance and Resiliency Engineering Lead, Merchant Services (Senior Vice President or Executive Director)

Merchant Services is hiring an SVP/ED Performance & Resiliency Engineer to improve the performance and stability of critical, high-volume platforms. The role focuses on latency reduction, throughput/TPS improvements, capacity planning, and peak-event readiness, while also strengthening resiliency and release safety.

You will partner closely with application development teams and infrastructure/platform engineering to identify bottlenecks across the stack (application/runtime, database, network, compute, and platform), implement durable fixes, and raise engineering standards through technical leadership, mentorship, and strong cross-team collaboration.

Key responsibilities

  • Lead performance engineering efforts: load/stress/soak testing, capacity modeling, performance tuning, and regression prevention (KPIs, guardrails, and acceptance criteria).
  • Improve production readiness: observability (metrics/logs/traces/APM), actionable alerting, incident triage, and root-cause analysis leading to durable remediation.
  • Strengthen resiliency patterns: timeouts/retries, circuit breakers, backpressure/rate limiting, graceful degradation, and failover readiness.
  • Drive safer releases via canary/progressive delivery and automated rollback patterns.
  • Optimize containerized workloads across EKS (primary) and ECS/other compute where applicable; drive autoscaling strategy and right-sizing.

Qualifications

  • Senior experience in performance engineering for distributed systems and/or SRE-style reliability engineering in production.
  • Strong cloud/container background (AWS + Kubernetes/EKS; ECS exposure beneficial).
  • Experience with modern observability tooling (e.g., Datadog, Dynatrace, Grafana, OpenTelemetry, CloudWatch or equivalent).
  • KEDA and/or Karpenter: large plus.
  • Akamai: strongly preferred.
  • Demonstrated ability to lead through influence, mentor engineers, and work effectively across teams.

What they ask for

Required

  • Senior experience in performance engineering for distributed systems and/or SRE-style reliability engineering in production
  • Strong cloud/container background: AWS + Kubernetes/EKS
  • Experience with modern observability tooling (Datadog, Dynatrace, Grafana, OpenTelemetry, CloudWatch or equivalent)
  • Demonstrated ability to lead through influence, mentor engineers, and work effectively across teams

Preferred

  • ECS exposure
  • KEDA and/or Karpenter experience
  • Akamai experience

Skills

See also

Software Engineering jobs by country — openings, pay and top skills →

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available