VP - SRE - Platform Engineering

Posted 2 Days Ago
Be an Early Applicant
2 Locations
In-Office
Senior level
Financial Services
The Role
Lead platform and site reliability engineering across mission-critical financial services systems. The role owns reliability strategy, SLOs, observability, incident and problem management, automation, infrastructure modernization, resilience, and operational risk reduction. It provides senior technical leadership, mentors global SRE and DevOps teams, drives cloud-native and Infrastructure as Code practices, and supports trading, post-trade, settlement, risk, and regulatory reporting platforms.
Summary Generated by Built In

Vice President, Platform Reliability Engineer (SRE)

Location: Pune

Role Overview

We are seeking an experienced Vice President, Platform Reliability Engineer (SRE) to lead reliability engineering initiatives across critical front-to-back trading, post-trade, and operations platforms. The role combines hands-on technical expertise with strategic leadership, driving platform stability, operational excellence, observability, automation, and resilience across the technology estate.

The successful candidate will partner with Engineering, Infrastructure, Architecture, Operations, and Business stakeholders globally to define reliability standards, drive platform modernization, reduce operational risk, and improve service availability.

 

Key Responsibilities

Reliability & Platform Engineering Leadership

  • Provide technical leadership for platform reliability, stability, scalability, and resilience across business-critical production systems.
  • Define and drive the strategic roadmap for Platform Reliability Engineering and Site Reliability Engineering practices.
  • Establish reliability objectives, service level indicators (SLIs), service level objectives (SLOs), and operational excellence standards across supported platforms.
  • Act as a senior escalation point during major incidents, driving resolution, recovery, stakeholder communication, and post-incident reviews.
  • Lead root cause analysis initiatives and ensure corrective and preventative actions are implemented effectively.

Operational Excellence

  • Drive reduction of operational toil through automation, self-healing capabilities, and process simplification.
  • Establish best practices for incident management, problem management, change management, and release governance.
  • Identify reliability risks and proactively implement mitigation strategies to improve platform resilience.
  • Define operational KPIs and reliability metrics, leveraging data-driven insights to drive continuous improvement.

Observability & Monitoring

  • Own and enhance enterprise observability capabilities across applications, infrastructure, middleware, and cloud environments.
  • Drive adoption of modern observability frameworks leveraging Datadog, OpenTelemetry, Grafana, Prometheus, Loki, and Jaeger.
  • Ensure effective monitoring, alerting, logging, tracing, and capacity planning practices are implemented across platforms.

Engineering & Automation

  • Partner with development teams to embed reliability principles throughout the software development lifecycle.
  • Lead engineering efforts focused on infrastructure automation, deployment automation, and platform modernization.
  • Champion Infrastructure as Code (IaC), CI/CD, and DevOps best practices.
  • Drive automation initiatives using Python, Terraform, Ansible, Jenkins, Kubernetes, and cloud-native technologies.

Stakeholder & Team Leadership

  • Collaborate closely with senior technology leaders, application owners, infrastructure teams, cybersecurity teams, and business stakeholders.
  • Provide technical mentorship and guidance to SRE, PRE, DevOps, and Production Support engineers.
  • Influence technology strategy and architectural decisions with reliability, scalability, and operational sustainability in mind.
  • Lead cross-functional initiatives spanning multiple regions and technology teams.
  • Represent Platform Reliability Engineering in governance forums, technology reviews, and operational risk discussions.

Financial Services Platform Reliability

  • Ensure operational stability and support of platforms that underpin trading, post-trade processing, settlements, risk management, and regulatory reporting.
  • Maintain high service availability and minimize disruption to revenue-generating and business-critical workflows.
  • Drive regulatory, audit, and operational risk compliance within supported environments.
 

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Technology, or related discipline.
  • 8+ years of experience in Platform Engineering, Site Reliability Engineering (SRE), DevOps, Production Engineering, or Application Support.
  • Proven experience supporting and operating large-scale, mission-critical production platforms.
  • Strong programming experience in Python, Go, Java, C#, or similar languages.
  • Extensive experience with Linux/Unix environments and distributed systems.
  • Strong understanding of SRE principles, operational excellence frameworks, and reliability engineering practices.
  • Experience leading major incident management and problem management processes.
  • Strong understanding of databases, messaging systems, and middleware technologies.
  • Excellent troubleshooting and analytical problem-solving skills across application, infrastructure, and data layers.
  • Experience working in globally distributed teams and managing senior stakeholder relationships.
  • Strong verbal and written communication skills with both technical and business audiences.
 

Preferred Qualifications

Observability

  • Datadog
  • OpenTelemetry
  • Grafana
  • Prometheus
  • Loki
  • Jaeger

DevOps & Automation

  • Git
  • Jenkins
  • GitHub Actions
  • Ansible
  • Terraform
  • CI/CD Frameworks

Cloud & Containers

  • Kubernetes
  • Docker
  • OpenShift
  • AWS / Azure / GCP

Data & Messaging Platforms

  • Kafka
  • Redis
  • MongoDB
  • Elasticsearch
  • PostgreSQL
  • SQL Server

Financial Services Experience

  • Investment Banking
  • Capital Markets
  • Equities
  • Fixed Income
  • Prime Brokerage
  • Post-Trade Processing
  • Operations Technology

About Us

Jefferies is a leading global, full-service investment banking and capital markets firm that provides advisory, sales and trading, research, and wealth and asset management services. With more than 40 offices around the world, we offer insights and expertise to investors, companies, and governments.

At Jefferies, we believe that diversity fosters creativity, innovation and thought leadership through the infusion of new ideas and perspectives. We have made a commitment to building a culture that provides opportunities for all employees regardless of our differences and supports a workforce that is reflective of the communities where we work and live. As a result, we are able to pool our collective insights and intelligence to provide fresh and innovative thinking for our clients.

Jefferies is an equal employment opportunity employer, and takes affirmative action to ensure that all qualified applicants will receive consideration for employment without regard to race, creed, color, national origin, ancestry, religion, gender, pregnancy, age, physical or mental disability, marital status, sexual orientation, gender identity or expression, veteran or military status, genetic information, reproductive health decisions, or any other factor protected by applicable law. We are committed to hiring the most qualified applicants and complying with all federal, state, and local equal employment opportunity laws. As part of this commitment, Jefferies will extend reasonable accommodations to individuals with disabilities, as required by applicable law.

Skills Required

  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline
  • 8+ years of experience in Platform Engineering, Site Reliability Engineering, DevOps, Production Engineering, or Application Support
  • Experience supporting and operating large-scale, mission-critical production platforms
  • Strong programming experience in Python, Go, Java, C#, or similar languages
  • Extensive experience with Linux or Unix environments and distributed systems
  • Strong understanding of SRE principles, operational excellence frameworks, and reliability engineering practices
  • Experience leading major incident management and problem management processes
  • Strong understanding of databases, messaging systems, and middleware technologies
  • Strong troubleshooting and analytical problem-solving skills across application, infrastructure, and data layers
  • Experience working in globally distributed teams and managing senior stakeholder relationships
  • Strong verbal and written communication skills with technical and business audiences
  • Experience with Datadog, OpenTelemetry, Grafana, Prometheus, Loki, or Jaeger
  • Experience with Git, Jenkins, GitHub Actions, Ansible, Terraform, and CI/CD frameworks
  • Experience with Kubernetes, Docker, OpenShift, AWS, Azure, or GCP
  • Experience with Kafka, Redis, MongoDB, Elasticsearch, PostgreSQL, or SQL Server
  • Financial services experience in investment banking, capital markets, equities, fixed income, prime brokerage, post-trade processing, or operations technology

Jefferies Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Jefferies and has not been reviewed or approved by Jefferies.

  • Strong & Reliable Incentives Compensation is positioned as competitive with meaningful upside, with potential for significant earnings depending on group performance and strong years. A formulaic, performance-linked bonus model ties payouts to individual fees or P&L, reinforcing a pay-for-results dynamic.
  • Parental & Family Support Family-building support is broad, including primary and non-primary caregiver leave, adoption assistance, subsidized emergency child and eldercare, and a $25,000 stipend for qualified surrogacy, adoption, or fertility support. Added resources like family support programs and parental-leave coaching further strengthen caregiver benefits.
  • Wellbeing & Lifestyle Benefits Lifestyle perks extend beyond standard coverage, including discount programs, commuter benefits, legal plans, charitable matching, and education assistance such as tuition support and scholarships for employees’ family members. Location-specific perks like gym stipends and free cafeteria lunch are also described as available in some offices.

Jefferies Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
6,435 Employees
Year Founded: 1962

What We Do

Jefferies, the global investment banking firm, has served companies and investors for 60 years. Headquartered in New York, with offices in over 30 cities around the world, the firm provides clients with capital markets and financial advisory services, institutional brokerage and securities research, as well as asset and wealth management. The firm provides research and execution services in equity, fixed income, and foreign exchange markets, as well as a full range of investment banking services including underwriting, mergers and acquisitions, restructuring and recapitalization, and other advisory services, with all businesses operating in the Americas, Europe and Asia. Jefferies Group LLC is a wholly-owned subsidiary of Jefferies Financial Group (NYSE: JEF), a diversified financial services company. More about our company can be found at www.jefferies.com.

Similar Jobs

Remote or Hybrid
2 Locations
289097 Employees

Mastercard Logo Mastercard

Director, Software Engineering

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Mastercard Logo Mastercard

Senior Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Navi Mumbai, Thane, Maharashtra, IND
38800 Employees

Mastercard Logo Mastercard

Software Engineer

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Pune, Maharashtra, IND
38800 Employees

Similar Companies Hiring

Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account