Senior SRE Engineer

Reposted 12 Hours Ago
Be an Early Applicant
Hyderabad, Telangana, IND
In-Office
Senior level
Information Technology • Consulting
The Role
Lead site reliability efforts to define SLIs/SLOs, run incident response and RCAs, build observability and runbooks, automate recovery and alerting, collaborate on resilience and capacity planning to improve availability and reduce MTTR.
Summary Generated by Built In
Position Title: SR SRE
Experience: 6 - 8 years
Job Location: Hyderabad/Ahmedabad
Work Mode: Hybrid Mode
Time Zone/Shift: Starts at 12:30 PM IST

Requirements:
6–8 years of experience in SRE or Infrastructure Engineering within cloud-native environments, with strong expertise in GCP (GKE, Load Balancing, VPN, IAM), Kubernetes, and Docker. Hands-on experience with observability tools such as Prometheus, Grafana, ELK, and Datadog, along with Terraform and Helm for infrastructure automation. Strong background in incident management, RCA, on-call support, SLIs/SLOs, error budgets, and reducing MTTR. Experience in designing highly available, scalable, and resilient platforms, with a focus on automation, performance optimization, capacity planning, and operational excellence.

Roles and Responsibilities:
  • Define and measure Service Level Indicators (SLIs), Service Level Objectives (SLOs), and manage error budgets across services.
  • Lead incident management for critical production issues – drive root cause analysis (RCA) and postmortems.
  • Create and maintain runbooks and standard operating procedures for high availability services.
  • Design and implement observability frameworks using ELK, Prometheus, and Grafana; drive telemetry adoption.
  • Coordinate cross-functional war-room sessions during major incidents and maintain response logs.
  • Develop and improve automated system recovery, alert suppression, and escalation logic.
  • Use GCP tools like GKE, Cloud Monitoring, and Cloud Armor to improve performance and security posture.
  • Collaborate with DevOps and Infrastructure teams to build highly available and scalable systems.
  • Analyze performance metrics and conduct regular reliability reviews with engineering leads.
  • Participate in capacity planning, failover testing, and resilience architecture reviews.
Who you'll be working with
  • SRE, DevOs, Infrastructure, and Engineering Teams
  • Engineering Leads and Technical Architects
  • Cloud and Platform Engineering Teams
  • Cross-functional teams during critical production incidents
  • Security and Operations teams
  • Product and Technology stakeholders
  • Teams focused on observability, automation, and platform resilience
Required Skills & Qualifications
  • 6–8 years of experience in SRE or Infrastructure Engineering.
  • Strong hands-on experience with GCP, GKE, Kubernetes, Docker, Terraform, and Helm.
  • Proficiency in Prometheus, Grafana, ELK, and Datadog for observability and monitoring.
  • Strong understanding of SLIs, SLOs, error budgets, incident management, RCA, and on-call operations.
  • Experience in designing and maintaining highly available, scalable, and resilient cloud-native systems.
  • Strong problem-solving, troubleshooting, and communication skills.
  • Experience with PagerDuty/OpsGenie and automated incident response.
Preferred / Nice-to-Have Skills
    • Familiarity with programming languages
    • Define product strategy by connecting the dots
    • Make product decisions without ambiguity in collaboration with the team
    • Manage important stakeholders across the organization while leading initiatives
    • Decision-making and accountability/ownership is critical
    • One point of executive customer engagement for a large portfolio
    • Contribute effectively outside your comfort zone

About Us:
We are a global, cloud-native organization with a strong presence across North America and India, delivering innovative digital transformation solutions to clients across diverse industries such as Financial Services, Healthcare, Retail & E-commerce, Manufacturing, and Technology. Our strong client base includes Fortune 500 enterprises as well as high-growth mid-market and startup organizations, giving our teams exposure to a wide variety of business challenges and cutting-edge solutions.
Our technology practices are built around modern, future-ready capabilities including Cloud Engineering, Data & Analytics, AI/ML, Digital Experience Platforms, Application Modernization, and Enterprise Solutions such as SAP and other leading platforms. We follow a design thinking-led approach combined with agile and lean engineering practices to deliver scalable, high-impact solutions. Backed by globally recognized certifications such as ISO 27001, SOC 1, SOC 2, SOC 3, UK Cyber Essentials Plus, and CMMI Level 3, we ensure the highest standards of security, compliance, process maturity, and quality across all our engagements.

Why Join Us:
    • Opportunity to work on global projects and Fortune 500 clients
    • Exposure to cutting-edge technologies
    • Strong learning, mentorship, and career growth programs
    • Collaborative and innovation-driven work culture

If you are passionate about working on innovative technologies and want to be part of a fast-growing organization, we encourage you to apply and be part of our journey.
Company Details:
Website: http://tblocks.com
LinkedIn: https://www.linkedin.com/company/techblocks/about/

Skills Required

  • 6-8 years of SRE or infrastructure engineering experience in cloud-native environments
  • GCP (GKE, Load Balancing, VPN, IAM)
  • Prometheus
  • Grafana
  • ELK
  • Datadog
  • Kubernetes
  • Docker
  • On-call incident management, RCA, SLIs/SLOs
  • Terraform
  • Helm
  • PagerDuty
  • OpsGenie
  • GCP Monitoring
  • Skywalking
  • Service Mesh
  • API Gateway
  • GCP Spanner
  • MongoDB (basic)
  • Cloud Armor
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Vaughan, Ontario
345 Employees
Year Founded: 2007

What We Do

@TechBlocks we power the software defined industries (SDI) of today and tomorrow. We are a software engineering and consulting firm. We build modern digital value chains and businesses reimagined to create frictionless experiences for innovative monetization methods and drive unforeseen efficiencies. We are known to build world class custom platforms and products that are cloud native for some of the worlds largest brands. We are the go to technology partners for born in digital businesses that grew with us from "Concept to Commercialization" and have revenues between $100M - $10B. We help modern businesses transition just from a technology outsourcing mentality to help create globally distributed digital COEs and mature them. Our converged COEs that we create in partnership with our clients help power software factories that are extremely dynamic. We have created modern digital COEs and factories that are created with a single minded goal to future proof our clients businesses. Everything we do is centred around two philosophies and practices - Design Thinking and Lean Engineering. Whether it is building digital commerce platforms, marketplace for worlds largest retailers or smart utilities applications and products or digital health products/platforms that power wearables, patches or devices across healthcare landscape; we do it all with speed and sophistication that is unmatched in the industry

Similar Jobs

MetLife Logo MetLife

Senior Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Hybrid
Hyderabad, Telangana, IND
43000 Employees

Vertafore Logo Vertafore

Senior Site Reliability Engineer

Information Technology • Insurance • Software
Hybrid
Hyderabad, Telangana, IND
2372 Employees

The TJX Companies, Inc. Logo The TJX Companies, Inc.

Site Reliability Engineer

eCommerce • Fashion • Retail
In-Office
500081, Hitech City, Telangana, IND
46062 Employees

Lloyds Technology Centre Logo Lloyds Technology Centre

Senior Site Reliability Engineer

Information Technology • Software • Consulting • Financial Services
In-Office
Hyderabad, Telangana, IND
65000 Employees

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account