L2 – Apigee + Microservice

Posted 8 Hours Ago
Be an Early Applicant
Mumbai, Maharashtra, IND
In-Office
Junior
Artificial Intelligence • Cloud • Information Technology • Automation
The Role
Provide L2 SRE support for Apigee and microservices: own escalations from L1, perform deep debugging, manage Kubernetes clusters, tune Nginx and API gateway policies, use observability tools to resolve incidents and automate runbook procedures to maintain uptime.
Summary Generated by Built In

Job Title: Site Reliability Engineer (SRE) – L2 Support (Apigee + Microservices)

Job Summary

We are seeking an analytical and technically skilled SRE – L2 Support Engineer with 2 to 3 years of experience specializing in API management and microservices infrastructure. In this role, you will act as the escalation point for the L1 monitoring team, taking ownership of deep technical triage, debugging complex application errors, and optimizing API gateways. You will focus on maintaining high uptime for critical banking application platforms, identifying systemic errors, and implementing permanent fixes or automated workarounds within a fast-paced cloud environment.

 

Key Responsibilities

L2 Triage & Advanced Incident Management

  • Escalation Ownership: Act as the direct L2 technical escalation point for complex infrastructure, application, and API connectivity issues routed from L1.
  • Deep-Dive Debugging: Analyze verbose log traces, evaluate application stack metrics, and dissect error signatures to quickly pinpoint core system failures.
  • Crisis Collaboration: Partner closely with L3 engineering, backend application development, and DevOps teams during major (P1/P2) production incidents to accelerate system recovery.

Apigee & Microservices Management

  • Apigee API Administration: Troubleshoot API proxy execution issues, configure security policies (OAuth, API keys, spike arrests), resolve SSL handshake errors, and debug payload routing failures.
  • Microservices Support: Monitor, debug, and trace distributed applications across complex microservice dependencies using distributed tracing logs.
  • Kubernetes Orchestration: Manage cluster health, troubleshoot pod eviction or crash loops, view container resource limits, and run diagnostic deployments utilizing advanced kubectl commands.

Observability, Reliability & Automated Recovery

  • Telemetry Dashboards: Navigate and customize production monitoring workflows within Datadog, Dynatrace, Prometheus, and Grafana to narrow down systemic bottlenecks.
  • Traffic Routing: Manage and debug complex reverse proxy, routing, and header rules configured within enterprise Nginx web servers.
  • Runbook Automation: Draft, refine, and execute technical runbooks and automated scripts to speed up manual triage and ensure consistent resolution steps.

 

Required Qualifications & Technical Skills

  • Experience: 2 to 3 years of dedicated hands-on experience in an L2 Application Support, Infrastructure Engineering, or SRE environment.
  • API Management: Strong hands-on engineering experience configuring and troubleshooting policies within Google Apigee.
  • Microservices & Containers: Deep familiarity with Microservice architectures running on Kubernetes (K8s) (e.g., executing container troubleshooting, reading pod events, managing namespaces).
  • Observability Toolsets: High competence investigating production incidents using tools like Datadog, Dynatrace, Prometheus, and Grafana.
  • Web Server Architecture: Intermediate understanding of Nginx configuration structures, upstream load balancing, and performance tuning.
  • Cloud & OS Infrastructure: Familiarity navigating Google Cloud Platform (GCP) systems along with strong Linux/Unix command-line tools for running performance diagnostics.
  • Operational Flexibility: Preparedness to participate in rotating shift schedules (including weekend support coverage) to secure zero-downtime banking services.

 



Skills Required

  • 2 to 3 years experience in L2 Application Support, Infrastructure Engineering, or SRE environment
  • Hands-on engineering experience configuring and troubleshooting Google Apigee policies
  • Experience with Microservice architectures and container troubleshooting on Kubernetes (K8s)
  • Proficiency with observability tools: Datadog, Dynatrace, Prometheus, and Grafana
  • Intermediate understanding of Nginx configuration, upstream load balancing, and performance tuning
  • Familiarity navigating Google Cloud Platform (GCP)
  • Strong Linux/Unix command-line skills for performance diagnostics
  • Experience configuring API security policies (OAuth, API keys, spike arrests) and troubleshooting SSL handshake errors
  • Willingness to participate in rotating shift schedules including weekend support
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
339 Employees
Year Founded: 2006

What We Do

SID Global Solutions (SIDGS) is an AI-first digital transformation and consulting company serving enterprises worldwide. It delivers intelligent full-stack solutions across artificial intelligence, cloud modernization, automation, API and application transformation, and enterprise intelligence. SIDGS helps organizations modernize legacy systems, optimize operations, integrate data and platforms, and accelerate measurable digital growth through scalable, customer-centric technology services for complex and demanding enterprise environments.

Similar Jobs

ZS Logo ZS

Manager - Governance, Risk & Compliance

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
Pune, Maharashtra, IND
15000 Employees

ZS Logo ZS

Senior Project Manager

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
Pune, Maharashtra, IND
15000 Employees

ZS Logo ZS

Senior Salesforce Cloud Administrator

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
Pune, Maharashtra, IND
15000 Employees

ZS Logo ZS

Technical Project Manager

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
Pune, Maharashtra, IND
15000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account