SRE Engineer

Reposted 2 Hours Ago
Be an Early Applicant
560064, Yelahanka, Karnataka, IND
In-Office
Mid level
Logistics • Transportation
The Role
Join the O&E SRE team to improve reliability, performance, and automation across container and cloud platforms. Participate in on-call rotations and incident response, develop Python/Ansible/shell automation, manage workloads on AWS/Azure/GCP, configure observability (Prometheus, Grafana, ELK), define SLIs/SLOs/error budgets, and explore AIOps-driven automation to reduce toil and speed remediation.
Summary Generated by Built In

Site Reliability Engineer (SRE) 

Ocean & Enablement Platform (O&E) – SRE Team 

 

Role Overview 

Join an exciting and cross-edge team shaping the future of container technology at Maersk. The O&E SRE Team ensures reliability, performance, and automation excellence across the Ocean &Enablement Platform. 

As a Site Reliability Engineer, you will collaborate closely with O&E platform teams to enhance system resilience, optimize cost and performance, and enable zero-touch operations through automation. 

You will gain hands-on experience in cloud technologies, incident management, and Python-based automation — while exploring AIOps and AI-driven automation to contribute to the reliability goals that power the Ocean & Enablement Platform ecosystem. 

Key Responsibilities 

  • Support and improve reliability, availability, and performance across O&E applications and shared services. 

  • Participate in on-call rotations, handle incidents, perform RCA, and implement corrective actions. 

  • Develop and maintain automation tools and scripts using Python, Ansible, or shell scripting to reduce manual toil. 

  • Deploy, monitor, and manage workloads on AWS / Azure / GCP, ensuring cost-efficient and reliable operations. 

  • Configure and enhance observability (Prometheus, Grafana, ELK) for proactive detection and fast recovery. 

  • Support SRE principles by defining SLIs, SLOs, and error budgets for O&E services. 

  • Explore AIOps and AI-driven automation to reduce alert noise, accelerate triage, and enable intelligent remediation. 

  • Collaborate with product, infra, and observability teams to improve reliability and incident response. 

  • Maintain runbooks, SOPs, and automation playbooks for operational readiness. 

Required Skills & Experience 

  • Bachelor’s degree in Computer Science, Engineering, or related field. 

  • 3–5 years of experience in SRE, DevOps, or Cloud Infrastructure roles. 

  • Strong programming skills in Python for automation and scripting. 

  • Familiarity with AWS, Azure, or GCP cloud environments. 

  • Working knowledge of Docker, Kubernetes, and Infrastructure as Code tools (Terraform / Ansible). 

  • Experience with observability tools (Prometheus, Grafana, ELK, Datadog, etc.). 

  • Exposure to incident response, troubleshooting, and root cause analysis. 

  • Ownership-driven mindset with passion for improving reliability through automation. 

Nice to Have 

  • Exposure to AIOps or AI/ML-driven automation for reliability and operations. 

  • Interest in building agentic AI workflows and AI-assisted automation for SRE use cases. 

  • Familiarity with LLM platforms (e.g., Azure AI Foundry / Azure OpenAI) and prompt-based tooling. 

  • Familiarity with cloud cost optimization and reliability frameworks. 

  • Knowledge of CI/CD pipelines and AIOps or AI/ML automation. 

  • Experience working with or administering databases (Oracle DBA, PostgreSQL, or other relational databases). 

  • Experience developing or maintaining service health dashboards. 

What You’ll Gain 

  • Hands-on experience in multi-cloud environments supporting global Ocean & Enablement platforms. 

  • Opportunity to contribute to automation, AIOps, and observability frameworks improving reliability at scale. 

  • Mentorship from senior SREs and architects on scalability and operational excellence. 

  • A career path toward advanced SRE roles within Maersk’s global technology ecosystem. 

Maersk is committed to a diverse and inclusive workplace, and we embrace different styles of thinking. Maersk is an equal opportunities employer and welcomes applicants without regard to race, colour, gender, sex, age, religion, creed, national origin, ancestry, citizenship, marital status, sexual orientation, physical or mental disability, medical condition, pregnancy or parental leave, veteran status, gender identity, genetic information, or any other characteristic protected by applicable law. We will consider qualified applicants with criminal histories in a manner consistent with all legal requirements.

 

We are happy to support your need for any adjustments during the application and hiring process. If you need special assistance or an accommodation to use our website, apply for a position, or to perform a job, please contact us by emailing  [email protected]

Skills Required

  • Bachelor's degree in Computer Science, Engineering, or related field
  • 3-5 years experience in SRE, DevOps, or Cloud Infrastructure roles
  • Strong programming skills in Python for automation and scripting
  • Shell scripting / Bash experience
  • Experience with public cloud environments (AWS, Azure, or GCP)
  • Working knowledge of Docker and Kubernetes
  • Experience with Infrastructure as Code tools (Terraform, Ansible)
  • Experience with observability tools (Prometheus, Grafana, ELK, Datadog)
  • Exposure to incident response, troubleshooting, and root cause analysis
  • Ownership-driven mindset focused on reliability through automation
  • Exposure to AIOps or AI/ML-driven automation for reliability and operations
  • Familiarity with LLM platforms (Azure AI Foundry / Azure OpenAI) and prompt-based tooling
  • Familiarity with cloud cost optimization and reliability frameworks
  • Experience with CI/CD pipelines
  • Experience administering databases (Oracle DBA, PostgreSQL, or other relational databases)
  • Experience developing or maintaining service health dashboards
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Copenhagen
58,338 Employees

What We Do

A.P. Moller - Maersk is an integrated transport and logistics company; going all the way, together, for our customers and society. ALL THE WAY is our commitment to connect the world so that everyone has both the possibility and the ability to trade, grow and thrive. The company employs roughly 110.000 employees across operations in 130 countries.

Similar Jobs

GitLab Logo GitLab

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
2500 Employees

JPMorganChase Logo JPMorganChase

Data Engineer

Financial Services
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
5550 Employees

AT&T Logo AT&T

Site Reliability Engineer

Internet of Things • Mobile • Retail
In-Office or Remote
3 Locations
150000 Employees

Similar Companies Hiring

Blissway Thumbnail
Computer Vision • Fintech • Hardware • Internet of Things • Machine Learning • Software • Transportation
Denver, CO
24 Employees
Toro TMS Thumbnail
Cloud • Enterprise Web • Sales • Software • Transportation
Chicago, IL
80 Employees
Axle Health Thumbnail
Artificial Intelligence • Healthtech • Information Technology • Logistics
Santa Monica, CA
25 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account