AIOps Leader

Posted 16 Days Ago
Be an Early Applicant
Bangalore, Bengaluru Urban, Karnataka, IND
In-Office
Senior level
Aerospace
The Role
Leads the architecture and implementation of AIOps automation across cloud AI platforms. Builds self-healing infrastructure, RAG knowledge tools, conversational support bots, observability pipelines, predictive anomaly detection, and operational dashboards. Eliminates support toil, improves MTTR and ticket deflection, standardizes logs and telemetry, and mentors L1/L2 engineers. Requires strong AWS or GCP, Kubernetes, infrastructure-as-code, scripting, observability, and generative AI production experience.
Summary Generated by Built In

Job Description:

Job Description - JR

Job Description Date: August 2026 

Role : AIOps Leader

Number of positions : 1 

Description: 

We are seeking a seasoned, hands-on AIOps Leader (7+ Years of Experience) to serve as the principal technical architect and lead builder for our centralized AI Product Support & Operations function. Operating within a high-growth AI PSL (Product/Service Line), you will design, architect, and execute the end-to-end automation strategy that transforms raw operational chaos into scalable, self-healing, and data-driven infrastructure.

In this role, you will bridge software engineering, site reliability, and AI/ML architectures. You will lead the creation of intelligent diagnostic pipelines, custom RAG-driven knowledge tools, self-healing systems, and automated triage engines. You will work closely with cross-functional leadership, L1/L2 support teams, and platform engineering to systematically eliminate operational toil, optimize MTTR, and build proactive anomaly detection mechanisms across our AI ecosystem.

Qualification & Experience: 

Education: Bachelor’s or Master’s degree in Computer Science, Software Engineering, Information Technology, or a related quantitative field.

Overall Experience: 7+ years of hands-on experience across Software Engineering, Site Reliability Engineering (SRE), DevOps, or Systems Operations—with a focused concentration on cloud infrastructure automation and AI/ML operational tooling.

Platform Specialization: At least 3+ years architecting and running operations directly within major cloud ecosystems (AWS or GCP), including native AI/ML compute platforms.


Key Responsibilities 

1. Architecture & Advanced AI/ML Automation

  • Self-Healing Infrastructure & Workflows: Architect, build, and deploy auto-remediation routines, script-based diagnostic runners, and event-driven automation triggers that autonomously resolve platform issues.

  • LLM & RAG Systems Engineering: Design, implement, and maintain advanced Retrieval-Augmented Generation (RAG) knowledge tools, vector databases, and LLM utilities that index telemetry, historic logs, and RCAs for instant incident context.

  • Bot & Agent Development: Lead the development and production rollout of conversational AI agents, custom webhooks, and self-service bots integrated into ticketing engines to automate Tier-1 and Tier-2 resolutions.

2. Observability, Telemetry & Predictive Analytics

  • Observability Architecture: Build enterprise-grade telemetry ingestion workflows, automated log scraping, and context-enrichment pipelines that dynamically append system metrics directly to incident tickets upon creation.

  • Predictive Anomaly Detection: Configure and tune real-time predictive alerting, log-pattern analysis, and AI-driven monitoring models across AWS, Azure, or GCP microservices.

  • Operational BI & Analytics: Architect and own centralized executive and operational dashboards (e.g., ServiceNow, Datadog) tracking MTTR velocity, system uptime, defect density, ticket deflection rates, and SLA/CSAT compliance.

3. Operational Engineering & L1/L2 Empowerment

  • Toil Elimination: Continuously audit support operational bottlenecks across product teams, transforming high-volume manual intervention points into production-grade, single-click, or fully autonomous workflows.

  • Log & Metadata Standardization: Standardize system log outputs, stack-trace formatting, and tagging taxonomy across all AI products to ensure platform telemetry remains machine-readable for AI engines.

  • Technical Mentorship: Guide L1/L2 support engineers on best practices for automation, code-based triage, and log parsing.

Technical Essentials
  • AWS & GCP Native AI Platforms: Deep hands-on experience orchestrating production AI/ML workflows on AWS (Bedrock, SageMaker AI, OpenSearch, AWS Lambda) or GCP (Vertex AI, Vertex AI Agent Builder, Cloud Run, BigQuery).

  • Cloud Infrastructure & Infrastructure-as-Code (IaC): Advanced experience writing and managing cloud provisioning scripts using Terraform, AWS CloudFormation, or Google Cloud Deployment Manager to deploy auto-scaling, resilient operations environments.

  • Containerization & Orchestration: Production experience managing microservices via Kubernetes (EKS/GKE) and Docker to support agentic AI workers, vector indexing engines, and automated micro-tasks.

  • Observability & Cloud Telemetry: Proven capability to configure full-stack observability across cloud environments using AWS CloudWatch or GCP Cloud Logging/Monitoring to trigger automated alerts and log enrichment.

  • Advanced Automation & Scripting: Strong engineering capability in Python, Go, or Shell to build autonomous cloud functions (AWS Lambda/GCP Cloud Run), self-healing infrastructure scripts, and custom ITSM connectors (Jira/ServiceNow APIs).

  • Enterprise Generative AI Stack: Production execution experience deploying RAG (Retrieval-Augmented Generation) architectures using cloud vector engines (Amazon Bedrock Knowledge Bases, GCP Vertex AI Search, Pinecone, or Qdrant) for automated incident context retrieval.

Soft Skills & Behavioral Attributes
  • Technical Leadership & Influence: Proven ability to architect complex automation systems independently, establish technical standards, and convince cross-functional teams to adopt new operational models.

  • Root-Cause Obsession: Relentless drive to solve underlying architectural flaws rather than applying temporary operational workarounds.

  • Cross-Functional Bridge: Ability to speak fluidly with product developers, AI researchers, enterprise clients, and technical support teams.

  • Thrives in Ambiguity: Highly adaptable problem solver capable of establishing order, structure, and code standards in unstructured environments.

Nice-to-Have / Added Advantages
  • Linguistic Skill: Professional working proficiency or full fluency in French (spoken and written).

  • Certifications: ITIL v5 Foundation or Managing Professional, Certified Agile Service Manager (CASM), PMP, or introductory Cloud/AI certifications (AWS Certified Cloud Practitioner, Azure AI Fundamentals).

  • AI Bot Implementation Experience: Direct hands-on experience implementing conversational AI, RAG-based internal search, or auto-remediation bots within support portals.

Success Metrics
  • MTTR (Mean Time to Resolution) - Reduction in overall incident MTTR via automated routing, diagnostic bots, and clear escalation paths .

  • Ticket Deflection & Reduction - Percentage of Tier-1 issues resolved via self-service bots, automated workflows, and root-cause resolution.

  • SLA & First Contact Resolution - High compliance across overall uptime, response, and resolution SLAs for all AI products. 

  • Operational Centralization - Onboarding of existing and newly building AI products into the unified service delivery framework.

  • Stakeholder & CSAT Score - Satisfaction scores from internal product units, business leadership, and end-users.

This job requires an awareness of any potential compliance risks and a commitment to act with integrity, as the foundation for the Company’s success, reputation and sustainable growth.

Company:

Airbus India Private Limited

Employment Type:

Permanent

-------

Experience Level:

Professional

Job Family:

Digital

By submitting your CV or application you are consenting to Airbus using and storing information about you for monitoring purposes relating to your application or future employment. This information will only be used by Airbus.
Airbus is committed to achieving workforce diversity and creating an inclusive working environment. We welcome all applications irrespective of social and cultural background, age, gender, disability, sexual orientation or religious belief.

Airbus is, and always has been, committed to equal opportunities for all. As such, we will never ask for any type of monetary exchange in the frame of a recruitment process. Any impersonation of Airbus to do so should be reported to [email protected].

At Airbus, we support you to work, connect and collaborate more easily and flexibly. Wherever possible, we foster flexible working arrangements to stimulate innovative thinking.

Skills Required

  • Bachelor's or Master's degree in Computer Science, Software Engineering, Information Technology, or a related quantitative field
  • 7+ years of hands-on experience in software engineering, SRE, DevOps, or systems operations
  • 3+ years architecting and operating within AWS or GCP cloud ecosystems
  • Experience with native AI/ML compute platforms in AWS or GCP
  • Experience with AWS Bedrock, SageMaker AI, OpenSearch, and AWS Lambda or GCP Vertex AI, Vertex AI Agent Builder, Cloud Run, and BigQuery
  • Advanced infrastructure-as-code experience using Terraform, AWS CloudFormation, or Google Cloud Deployment Manager
  • Production experience managing Kubernetes, EKS or GKE, and Docker microservices
  • Experience configuring observability and telemetry using AWS CloudWatch or GCP Cloud Logging and Monitoring
  • Strong Python, Go, or Shell scripting capability
  • Experience building ITSM connectors using Jira or ServiceNow APIs
  • Production experience deploying RAG architectures using Bedrock Knowledge Bases, Vertex AI Search, Pinecone, or Qdrant
  • Ability to architect automation systems, establish technical standards, and influence cross-functional teams
  • Professional working proficiency or fluency in French
  • ITIL v5, CASM, PMP, AWS Cloud Practitioner, or Azure AI Fundamentals certification
  • Hands-on experience implementing conversational AI, RAG internal search, or auto-remediation bots in support portals

Airbus Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Airbus and has not been reviewed or approved by Airbus.

  • Healthcare Strength Healthcare coverage is positioned as comprehensive in several locations, including medical, dental, and vision options available from day one in the U.S. Access to life insurance, disability coverage, and employee assistance/wellbeing support adds breadth to the health offering.
  • Retirement Support Retirement support is framed as a meaningful part of the package through plans such as a 401(k) with company matching in the U.S. These programs strengthen long-term financial security beyond base wages.
  • Leave & Time Off Breadth Time-off provisions are described as generous in some settings, including vacation availability from day one and extended holiday coverage. Flexible working arrangements and hybrid options further increase the perceived value of time-related benefits.

Airbus Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Toulouse
52,655 Employees
Year Founded: 2014

What We Do

Airbus is a global leader in aeronautics, space and related services. In 2020, it generated revenues of €49.9 billion and employed a workforce of around 130,000. Airbus offers the most comprehensive range of passenger airliners. Airbus is also a European leader providing tanker, combat, transport and mission aircraft, as well as one of the world’s leading space companies. In helicopters, Airbus provides the most efficient civil and military rotorcraft solutions worldwide. Airbus is an international pioneer in the aerospace industry and a leader in designing, manufacturing and delivering aerospace products, services and solutions to customers on a global scale. We believe that it’s not just what we make, but how we make it that counts; promoting responsible, sustainable and inclusive business practices and acting with integrity. Our people work with passion and determination to make the world a more connected, safer and smarter place, on the ground, in the sky and in space.

Similar Jobs

Ericsson Logo Ericsson

Window Exchange Expert

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
88000 Employees

Expedia Group Logo Expedia Group

Senior Manager, Software Development Engineering

AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Hybrid
Bangalore, Bengaluru Urban, Karnataka, IND
16000 Employees

Graphcore Logo Graphcore

Design Engineer

Artificial Intelligence • Semiconductor
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
903 Employees

Wells Fargo Logo Wells Fargo

Operations Associate

Fintech • Financial Services
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
205000 Employees

Similar Companies Hiring

Red 6 Thumbnail
Aerospace • Hardware • Software • Virtual Reality • Defense
Orlando, Florida
186 Employees
Outpost Space Thumbnail
Aerospace • Defense
US
24 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account