SRE Manager

Posted 2 Days Ago
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Expert/Leader
Cloud • Fintech • Information Technology • Consulting
The Role
Leads the SRE function and team responsible for reliability, scalability, automation, observability, incident management, and operational excellence across cloud, networking, data center, and telecom-scale infrastructure. The role manages distributed systems, microservices, Kubernetes and Docker environments, infrastructure automation, CI/CD, monitoring, major incident response, postmortems, and reliability improvements while partnering with architecture, DevOps, networking, NOC, and cloud teams.
Summary Generated by Built In
Job description 
Job Title: SRE Manager
Location: Bangalore
Department: IT-Dev 
Experience - 15+ Years Industry Experience | 8+ Years Relevant Experience 
Client – JIO 
 
Role Overview: 
We are looking for an experienced SRE Manager to lead the reliability engineering function 
responsible for operating and scaling mission-critical infrastructure across cloud, networking, 
and data centre environments. 
You will lead a team responsible for building highly reliable, scalable, and automated 
production platforms supporting enterprise and telecom-scale workloads. This role requires 
deep expertise in distributed systems, cloud infrastructure, networking, and observability 
along with strong leadership capabilities. 
Key Responsibilities: 
Reliability Engineering 
Build and lead the Site Reliability Engineering (SRE) function focused on reliability, 
scalability, and operational excellence. 
Define and manage SLIs, SLOs, and SLAs for critical production systems. 
Drive improvements in system resilience, fault tolerance, and performance 
optimization. 
Production Infrastructure & Platform Operations 
Ensure stability of large-scale production environments across cloud and data centers. 
Manage reliability of distributed platforms, microservices environments, and 
containerized systems. 
Support architecture teams in designing highly scalable infrastructure platforms. 
Incident Management & Operational Excellence 
Lead major incident response and outage management across infrastructure and 
platform services. 
Establish incident management frameworks, escalation processes, and postmortem 
practices. 
Drive root cause analysis and reliability improvements. 
Automation & DevOps 
Drive automation initiatives to reduce operational overhead. 
Implement Infrastructure as Code (IaC) and automated infrastructure provisioning. 
Improve CI/CD pipelines and operational workflows. 








 




Observability & Monitoring 
Design and maintain observability platforms including monitoring, logging, and tracing 
systems. 
Establish real-time operational visibility across infrastructure, applications, and 
networks. 
Build dashboards and analytics to measure system performance and reliability. 
Team Leadership 
Lead and mentor SRE engineers, platform engineers, and reliability teams. 
Build a culture of automation, engineering excellence, and reliability-first mindset. 
Collaborate with DevOps, network engineering, NOC, cloud, and architecture teams. 
Technical Expertise 
Strong experience with cloud platforms such as AWS, Microsoft Azure, or Google Cloud. 
Deep understanding of distributed systems and microservices architecture. 
Experience with container orchestration platforms such as Kubernetes and Docker. 
Knowledge of core networking technologies including BGP, VXLAN, EVPN, and SD-WAN. 
Experience with observability and monitoring platforms such as Prometheus, Grafana, 
ELK, Datadog, or similar tools. 
Familiarity with infrastructure automation tools such as Terraform, Ansible, or similar 
frameworks. 
Understanding of ISP/telecom infrastructure, network operations, and large-scale 
traffic environments. 
Experience with data centre infrastructure, virtualization, and private cloud 
environments. 

Skills Required

  • 15+ years of industry experience
  • 8+ years of relevant experience
  • Experience leading Site Reliability Engineering, platform engineering, or reliability teams
  • Strong experience with AWS, Microsoft Azure, or Google Cloud
  • Deep understanding of distributed systems and microservices architecture
  • Experience with Kubernetes and Docker
  • Knowledge of BGP, VXLAN, EVPN, and SD-WAN networking technologies
  • Experience with Prometheus, Grafana, ELK, Datadog, or similar observability platforms
  • Familiarity with Terraform, Ansible, or similar infrastructure automation frameworks
  • Understanding of ISP or telecom infrastructure, network operations, and large-scale traffic environments
  • Experience with data center infrastructure, virtualization, and private cloud environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
79 Employees
Year Founded: 2017

What We Do

ZYBISYS CONSULTING SERVICES LLP is an India-based IT services and consulting firm focused on helping fintech and other businesses modernize technology. It provides public, private, and hybrid cloud hosting, managed IT services, DevOps, data-center management, infrastructure support, cybersecurity, and technology consulting. Its offerings also include software development, automation, monitoring, and cloud-native solutions designed to improve scalability, operational efficiency, security, and business continuity.

Similar Jobs

In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
15967 Employees
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
345 Employees

Cleo Logo Cleo

Support Engineer

Cloud • eCommerce • Information Technology • Professional Services • Software
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
500 Employees

CSC Logo CSC

Database Administrator

Fintech • Legal Tech • Software • Financial Services • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
8500 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account