Associate Staff Engineer(SRE)

Posted 16 Days Ago
Be an Early Applicant
Shanghai, Shanghai Municipality, Shanghai, CHN
In-Office
Senior level
Artificial Intelligence • Information Technology • Machine Learning • Software • Virtual Reality • Analytics
The Role
Owns reliability, availability, scalability, security, and operational readiness for critical hospitality platforms. Responsibilities include cloud and Kubernetes operations, infrastructure as code, CI/CD, observability, incident response, root-cause analysis, automation, performance optimization, disaster recovery, vulnerability remediation, and 24x7 production support across global systems.
Summary Generated by Built In
Job Description

Must have Skills : DevOps - AWS (Strong)

Job Description :

Senior Site Reliability Engineer (SRE) Role

We are seeking an experienced Senior Site Reliability Engineer (SRE) to support highly available, business-critical platforms within a global hospitality environment.

The role is responsible for ensuring the reliability, availability, performance, scalability, security, and operational readiness of production systems supporting hotel operations, reservation services, digital channels, loyalty platforms, payment services, and other guest-facing applications.

The successful candidate will combine strong expertise in Cloud, DevOps, Kubernetes, Infrastructure as Code, Observability, Automation, and Production Support, and will work closely with Development, Infrastructure, Security, Architecture, and Service Management teams in a global 24x7 operating model.

Key Responsibilities Ensure the availability, reliability, performance, and scalability of critical production services.

Define and monitor SLIs, SLOs, SLAs, error budgets, availability, latency, and service health metrics.

Act as a senior technical escalation point for major production incidents and P1/P2 issues.

Lead troubleshooting, Root Cause Analysis, post-incident reviews, and corrective actions.

Reduce operational toil through automation, self-healing, and engineering improvements.

Operate and troubleshoot workloads across AWS, Azure, and/or GCP environments.

Support enterprise Kubernetes and container platforms, including EKS, AKS, GKE, or OpenShift.

Develop and maintain Infrastructure as Code using Terraform, CloudFormation, Bicep, or equivalent technologies.

Build and maintain CI/CD pipelines using tools such as Jenkins, GitHub Actions, GitLab CI, Azure DevOps, Argo CD, or Harness.

Implement and maintain observability solutions covering metrics, logs, tracing, alerts, dashboards, and APM. Support tools such as Prometheus, Grafana, Dynatrace, Datadog, Splunk, ELK/OpenSearch, or New Relic.

Perform performance analysis, capacity planning, load testing, and scalability optimization.

Support Disaster Recovery, Business Continuity, failover testing, and RTO/RPO validation.

Participate in production release, change, patching, vulnerability remediation, and operational readiness activities.

Maintain technical runbooks, operational procedures, monitoring standards, and knowledge documentation.

Participate in a global 24x7 on-call / production support model where required.

Hospitality Technology Scope The role may support platforms including: Property Management Systems (PMS) Central Reservation Systems (CRS) Hotel booking and reservation platforms Loyalty and membership systems Guest-facing websites and mobile applications Payment platforms Property connectivity and integration services API and middleware platforms Revenue management and hotel operational systems

The engineer will help ensure the reliability of critical guest journeys such as hotel search, booking, reservation modification, payment, check-in/check-out, loyalty transactions, and property system integration.

Required Qualifications Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline.

7+ years of experience in Cloud, DevOps, Infrastructure, Production Engineering, or IT Operations.

3+ years of hands-on experience in an SRE, DevOps, Cloud Operations, or Production Engineering role.

Strong hands-on experience with at least one major cloud platform: AWS, Azure, or GCP. Strong Kubernetes, Docker, and container troubleshooting skills.

Hands-on experience with Terraform or other Infrastructure as Code technologies.

Strong experience with CI/CD, deployment automation, and release management. Strong knowledge of Linux and production troubleshooting.

Experience with enterprise monitoring, logging, tracing, and observability platforms.

Experience managing major incidents, RCA, problem management, and production stability.

Good knowledge of networking concepts including DNS, HTTP/HTTPS, load balancing, firewall, routing, VPN, and CDN.

Experience supporting microservices, distributed systems, APIs, databases, and messaging platforms. Scripting or programming experience with Python, Bash, PowerShell, Go, Java, or similar languages.

Good understanding of ITIL-based Incident, Problem, Change, and Knowledge Management processes. Strong written and verbal English communication skills.

Preferred Qualifications Experience in hospitality, travel, airline, e-commerce, financial services, or other 24x7 high-availability industries.

Experience supporting high-volume transactional or reservation platforms.

Experience with PCI DSS, GDPR, ISO 27001, DevSecOps, and vulnerability management.

Experience with Kafka, Redis, API gateways, service mesh, or event-driven architectures.

Experience with cloud cost optimization / FinOps.

Experience with resilience testing, chaos engineering, or automated recovery.

Relevant certifications such as AWS, Azure, GCP, CKA, Terraform Associate, or ITIL.

Skills Required

  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline
  • 7+ years of experience in Cloud, DevOps, Infrastructure, Production Engineering, or IT Operations
  • 3+ years of hands-on experience in an SRE, DevOps, Cloud Operations, or Production Engineering role
  • Strong hands-on experience with AWS, Azure, or GCP
  • Strong Kubernetes, Docker, and container troubleshooting skills
  • Hands-on experience with Terraform or other Infrastructure as Code technologies
  • Strong experience with CI/CD, deployment automation, and release management
  • Strong knowledge of Linux and production troubleshooting
  • Experience with enterprise monitoring, logging, tracing, and observability platforms
  • Experience managing major incidents, root-cause analysis, problem management, and production stability
  • Knowledge of DNS, HTTP/HTTPS, load balancing, firewalls, routing, VPN, and CDN
  • Experience supporting microservices, distributed systems, APIs, databases, and messaging platforms
  • Scripting or programming experience with Python, Bash, PowerShell, Go, Java, or similar languages
  • Knowledge of ITIL-based Incident, Problem, Change, and Knowledge Management processes
  • Strong written and verbal English communication skills
  • Experience in hospitality, travel, airline, e-commerce, financial services, or other 24x7 high-availability industries
  • Experience supporting high-volume transactional or reservation platforms
  • Experience with PCI DSS, GDPR, ISO 27001, DevSecOps, and vulnerability management
  • Experience with Kafka, Redis, API gateways, service mesh, or event-driven architectures
  • Experience with cloud cost optimization or FinOps
  • Experience with resilience testing, chaos engineering, or automated recovery
  • AWS, Azure, GCP, CKA, Terraform Associate, or ITIL certification

Nagarro Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Nagarro and has not been reviewed or approved by Nagarro.

  • Pay Growth & Progression — Compensation is at times described as competitive, with salary hikes and perks occurring on certain occasions. Better growth opportunities and compensation are also positioned as an advantage versus other service-based companies.
  • Flexible Benefits — Work arrangements are framed around a “work-from-anywhere” mindset with flexitime and family-friendly working models. This flexibility appears to add meaningful value to the overall rewards package for many roles.
  • Healthcare Strength — Medical, dental, and vision coverage are described as available for employees and dependents, alongside life insurance. Mental-health support is also included via an Employee Assistance Program (EAP).

Nagarro Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Munich
19,994 Employees
Year Founded: 1996

What We Do

Nagarro helps future-proof your business through a forward-thinking, fluidic, and CARING mindset. We excel at digital engineering and help our clients become human-centric, digital-first organizations, augmenting their ability to be responsive, efficient, intimate, creative, and sustainable. Today, we are 19,000 experts across 36 countries, forming a Nation of Nagarrians, ready to help our customers succeed.

Similar Jobs

Takeda Logo Takeda

Innovative Access Expert, Market Access, Shanghai / Beijing

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Hybrid
2 Locations
50000 Employees

Takeda Logo Takeda

Head, China External Innovation

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Hybrid
Shanghai, Shanghai Municipality, Shanghai, CHN
50000 Employees

Takeda Logo Takeda

Head of Regulatory Affairs, R&D, Shanghai/Beijing

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Hybrid
2 Locations
50000 Employees

Takeda Logo Takeda

Marketing Manager

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Hybrid
2 Locations
50000 Employees

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account