Senior Engineer – Observability Engineering

Posted Yesterday
Be an Early Applicant
2 Locations
In-Office or Remote
Senior level
eCommerce • Fashion • Retail
The Role
Designs and leads enterprise observability solutions across metrics, logs, traces, synthetic monitoring, RUM, and business telemetry. Implements instrumentation across cloud-native and legacy systems, applies AI-powered AIOps for anomaly detection and automated remediation, and integrates monitoring into CI/CD and GitOps. The role establishes governance, evaluates tools, partners with engineering teams, and mentors teams on observability, reliability, and operational excellence.
Summary Generated by Built In

Job Location: Latin America

 

Calling all originals: At Levi Strauss & Co., you can be yourself — and be part of something bigger. We’re a company of people who like to forge our own path and leave the world better than we found it. Who believe that what makes us different makes us stronger. So add your voice. Make an impact. Find your fit — and your future. 

You will be part of the Observability Engineering Team within the Levi’s Shared Platform & Services Technology Organization—a group committed to delivering end-to-end visibility, actionable insights, and operational excellence across our digital ecosystem. The team’s mission is to empower engineering squads with the data, dashboards, and analytics they need to detect, diagnose, and resolve issues faster—ultimately improving reliability, performance, and customer experience. 

 

This role offers a unique opportunity for individuals who are passionate about Observability, Site Reliability, and Data-Driven Engineering to shape how Levi’s measures and manages its digital systems. You will play a key role in adoption of AI within our Observability capabilities and advancing Levi’s Observability Strategy—driving proactive monitoring, intelligent alerting, and continuous improvement across our platforms and services. 

 

This role demands strong expertise in applying AI-powered observability and AIOps solutions to enable proactive monitoring, intelligent alerting, automated incident analysis, and self-healing capabilities, driving faster issue resolution and improved engineering efficiency. 

 

As a member of the Observability Engineering Team, you will join a growing community focused on building a culture of visibility, accountability, and performance excellence, enabling Levi’s engineering squads to move with confidence and speed in our ongoing digital transformation journey. 

 

About the Job 

  • Design, build, and maintain end-to-end observability solutions covering metrics, logs, traces, synthetic monitoring, real user monitoring (RUM), and business telemetry using New Relic and modern cloud-native monitoring services. 

  • Develop and implement instrumentation of Observability standards for microservices, APIs, front-end, and legacy systems, ensuring consistent visibility across hybrid-cloud environments. 

  • Leverage AI driven observability capabilities within New Relic and modern AIOps platforms to enable intelligent anomaly detection, predictive alerting, automated root cause analysis, noise reduction, and self-healing workflows, improving incident response efficiency and accelerating developer productivity. 

  • Integrate observability into CI/CD and GitOps workflows, enabling automated monitoring, alerting, and feedback loops throughout the software delivery pipeline. 

  • Lead cross-functional collaboration with platform, Developer Velocity and application teams to establish scalable monitoring architectures and operational excellence practices. 

  • Evaluate, select, and implement next-generation observability tools and frameworks aligned with enterprise architecture, security, and scalability goals. 

  • Mentor and guide engineering teams on observability best practices, fostering a data-driven, performance-oriented culture across the technology organization. 

  • Establish and advance observability governance and maturity models, ensuring compliance with SLAs. 

  • Partner with technology and business leadership to translate observability insights into actionable improvements driving product quality, developer velocity, and operational efficiency. 

About You 

  • 7+ years of total IT industry experience with a strong focus on Observability and monitoring solutions.   

  • 5+ years of solid hands-on experience in administrating, managing and integrating the New Relic Solutions in GCP , Azure & AWS Cloud. 

  • Experience in managing Observability Solutions like Datadog , Dynatrace , Grafana , Prometheus etc. 

  • Deep understanding of the “Four pillars” of Observability—Metrics, Events, Logs, and Traces—and how they interconnect to drive reliability and performance insights. 

  • Strong background in distributed systems, microservices, and containerized applications (Kubernetes, EKS, Docker, service mesh architectures). 

  • Knowledge of automation, CI/CD, and DevSecOps practices in integrating observability into modern delivery pipelines. 

  • Excellent communication, collaboration & leadership skills with the ability to translate business observability needs into technical solutions 

  • Strong troubleshooting, analytical and problem-solving abilities. 

  • Understanding of SDLC, Agile Practices & ITIL Frameworks. 

  • Technical Certifications highly preferred: New Relic Certified Reliability Engineer, New Relic verified Foundation, New Relic Certified Performance Engineer. 

  • Required: Bachelor’s degree in computer science or equivalent; master’s degree preferred.

LOCATIONMexico, D.F., MexicoFULL TIME/PART TIMEFull timeCurrent LS&Co Employees, apply via your Workday account.

Skills Required

  • 7+ years of total IT industry experience with a strong focus on observability and monitoring solutions
  • 5+ years of hands-on experience administering, managing, and integrating New Relic solutions across GCP, Azure, and AWS
  • Experience managing observability solutions such as Datadog, Dynatrace, Grafana, and Prometheus
  • Deep understanding of observability metrics, events, logs, and traces
  • Experience with distributed systems, microservices, containerized applications, Kubernetes, EKS, Docker, and service mesh architectures
  • Knowledge of automation, CI/CD, and DevSecOps practices
  • Excellent communication, collaboration, and leadership skills
  • Strong troubleshooting, analytical, and problem-solving abilities
  • Understanding of SDLC, Agile practices, and ITIL frameworks
  • Bachelor's degree in computer science or equivalent
  • Master's degree
  • New Relic Certified Reliability Engineer, New Relic verified Foundation, or New Relic Certified Performance Engineer certification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA

What We Do

We’re a company of people who like to forge our own path. We invented the blue jean in 1873, and we reinvented khaki pants in 1986. We pioneered labor and environmental guidelines in manufacturing. And we work to build sustainability into everything we do. We just might be the original startup.

Similar Jobs

Cloudflare Logo Cloudflare

Senior Customer Engineer, LATAM - MCR Bogotá, Colombia.

Cloud • Information Technology • Security • Software • Cybersecurity
Remote or Hybrid
Colombia
4400 Employees

Shield AI Logo Shield AI

District Manager, LATAM

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
In-Office or Remote
2 Locations

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Sr. Sales Associate III

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
15-20 Hourly

Domino Data Lab Logo Domino Data Lab

Support Engineer

Artificial Intelligence • Machine Learning
Remote or Hybrid
10 Locations
200 Employees

Similar Companies Hiring

PRIMA Thumbnail
Travel • Software • Marketing Tech • Hospitality • eCommerce
US
15 Employees
Scotch Thumbnail
Artificial Intelligence • eCommerce • Fintech • Payments • Retail • Software • Analytics
US
35 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account