Senior Vice President – Data Platform Lead - Platform Reliability Engineering

Open 16d reposted 2× · 2 open copies

Job Title: Senior Vice President – Platform Reliability Engineering (Data Platform PREM)

Location: Pune

Overview

We are seeking a highly experienced and hands-on Senior Vice President – Platform Reliability Engineering to lead reliability, stability, and operational excellence for Jefferies’ Data Platform (PREM). This role is pivotal in ensuring resiliency across front-to-back technology systems supporting post-trade processing, operations, and regulatory functions.

The ideal candidate will combine strong technical depth with strategic leadership, driving platform reliability, automation, and observability while partnering closely with global engineering and business teams.

Key Responsibilities

  • Lead and own platform reliability, stability, and resilience across middle-office, operations, and data platform applications.
  • Drive incident management, major incident command, root cause analysis, and problem management governance globally.
  • Partner with architecture, engineering, and regional teams to define and execute reliability and scalability initiatives.
  • Champion SRE principles including error budgets, SLAs/SLOs, and continuous improvement practices.
  • Identify opportunities to eliminate manual processes through automation and tooling, improving operational efficiency.
  • Define and implement enterprise-grade monitoring, observability, and alerting frameworks (AppDynamics, OpenTelemetry, Grafana stack, etc.).
  • Collaborate with engineering teams on system design, performance optimization, capacity planning, and resiliency improvements.
  • Provide oversight on production support, including troubleshooting complex, cross-stack issues.
  • Lead and mentor distributed teams, fostering a culture of ownership, accountability, and engineering excellence.
  • Work closely with stakeholders to align platform capabilities with regulatory, operational, and business requirements.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
  • 15+ years of experience in software engineering, production support, or platform reliability roles, preferably in financial services.
  • Strong programming expertise in one or more languages such as Python, Go, C/C++, or C#.
  • Deep understanding of distributed systems, scalability, and high-availability architecture.
  • Proven experience implementing SRE frameworks, including SLAs/SLOs, observability, and automation.
  • Strong expertise in Linux/Unix and Windows environments.
  • Hands-on experience with databases, data platforms, and troubleshooting data access/performance issues.
  • Familiarity with open-source and data ecosystem tools such as Kafka, Redis, MongoDB, and Elasticsearch.
  • Experience with observability and monitoring platforms (Grafana, Prometheus, Jaeger, Loki, AppDynamics).
  • Exposure to DevOps and CI/CD tools such as Git, Jenkins, Ansible.
  • Excellent leadership, communication, and stakeholder management skills, with the ability to engage at executive levels.
  • Strong ownership mindset, with a focus on operational excellence and continuous improvement.