Monitoring and Observability Engineer

Posted 12 Hours Ago
Be an Early Applicant
Bangalore, Bengaluru Urban, Karnataka, IND
In-Office
Mid level
Artificial Intelligence • Hardware • Automation • Manufacturing
The Role
Design, implement, and maintain monitoring and observability infrastructure using Prometheus, Grafana, Nagios, SolarWinds and Kafka. Build KPI standards, AI-based anomaly detection, SLAs/SLOs, automation scripts (bash/Python/Go), runbooks, and collaborate with DevOps and application teams to optimize performance and reliability.
Summary Generated by Built In

About Analog Devices

Analog Devices, Inc. (NASDAQ: ADI) is a global semiconductor leader that bridges the physical and digital worlds to enable breakthroughs at the Intelligent Edge. ADI combines analog, digital, AI, and software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in automation and robotics, mobility, healthcare, energy and data centers. With revenue of more than $11 billion in FY25, ADI ensures today's innovators stay Ahead of What's Possible. Learn more at www.analog.com and on LinkedIn and X.

          

Position Overview

We are seeking an experienced Monitoring and Observability Engineer to be part of our monitoring and observability initiatives. This role requires to implement monitoring standards and KPIs with hands-on technical expertise in building and maintaining robust observability infrastructure. You will work on key technologies and implement best practices to enhance the user experience across our monitoring ecosystem.

Key Responsibilities
  • Implement and manage KPI/Metrics standards and best practices for the organization, ensuring consistency and measurability
  • Deploy monitoring solutions using industry-leading tools including Prometheus, Grafana, Nagios, and SolarWinds
  • Develop and implement Kafka-based event streaming and monitoring pipelines for real-time data collection and analysis
  • Integrate and manage AI monitoring capabilities to enable intelligent alerting, anomaly detection, and predictive insights
  • Create and maintain comprehensive documentation and runbooks for monitoring infrastructure, troubleshooting procedures, and operational playbooks
  • Develop custom scripts and automation in bash, Python, or Go to enhance monitoring capabilities and reduce operational overhead
  • Establish and refine SLAs, SLOs, and alerting thresholds based on business requirements and operational data
  • Conduct performance analysis and optimization of monitoring infrastructure to reduce latency and improve system reliability
  • Collaborate with DevOps, Platform, and Application teams to ensure comprehensive observability across the entire stack
  • Stay current with emerging monitoring and observability trends, tools, and technologies
  • Drive continuous improvement initiatives by analyzing metrics, gathering feedback, and implementing enhancements to the user experience
Required Qualifications
  • Minimum 3+ years of professional experience in monitoring, observability, systems engineering, or DevOps roles
  • Proven experience in monitoring or infrastructure domains
  • Deep hands-on expertise with opensource monitoring tools such as Prometheus, Grafana, and Nagios
  • Strong experience with enterprise monitoring platforms like SolarWinds
  • Proficiency in bash scripting and Linux system administration (RHEL/CentOS/Ubuntu preferred)
  • Experience with event-driven architectures and Apache Kafka for real-time data processing
  • Knowledge of monitoring and observability across Linux systems, storage infrastructure, applications, and network platforms
  • Understanding of metrics collection, time-series data management, and data-driven decision-making
  • Experience implementing AI/ML-based monitoring solutions or anomaly detection systems
  • Strong problem-solving skills and ability to work effectively in fast-paced, complex environments
  • Excellent communication skills with ability to present technical concepts to both technical and non-technical stakeholders
  • Experience with Infrastructure-as-Code (IaC) tools (Terraform, Ansible, etc.)
Preferred Qualifications
  • Experience with containerized environments and Kubernetes monitoring
  • Proficiency in Python or Go for custom tooling development
  • Experience with cloud platforms (AWS, Azure, GCP) and their native monitoring services
  • Experience with observability backends such as OpenTelemetry.
  • Knowledge of application performance monitoring (APM) tools
  • Experience with incident response and post-mortem processes
  • Certification in relevant areas (e.g., Kubernetes, cloud platforms, monitoring tools)
  • Contribution to open-source monitoring projects
Required Technical Skills
  • Monitoring Platforms: Prometheus, Grafana, Nagios, Azure Monitor, AWS CloudWatch, SolarWinds
  • Data Streaming: Apache Kafka, event processing pipelines
  • AI Monitoring: Anomaly detection, intelligent alerting, ML-based insights
  • Scripting Languages: Bash, Python, Go
  • Operating Systems: Linux (RHEL, CentOS, Ubuntu), system administration
  • Infrastructure: Storage systems, networking, application monitoring
  • Data Visualization: Grafana dashboards, custom visualizations
  • Cloud & Containerization: Docker, Kubernetes (optional but valuable)
What We're Looking For

We seek a hands-on Engineer who is passionate about observability and excellence. Ideal candidates will:
• Balance strategic thinking with hands-on technical expertise
• Team player
• Possess strong written and verbal communication skills
• Approach challenges with creativity and persistence
• Stay current with evolving monitoring and observability technologies

For positions requiring access to technical data, Analog Devices, Inc. may have to obtain export  licensing approval from the U.S. Department of Commerce - Bureau of Industry and Security and/or the U.S. Department of State - Directorate of Defense Trade Controls.  As such, applicants for this position – except US Citizens, US Permanent Residents, and protected individuals as defined by 8 U.S.C. 1324b(a)(3) – may have to go through an export licensing review process.

Analog Devices is an equal opportunity employer. We foster a culture where everyone has an opportunity to succeed regardless of their race, color, religion, age, ancestry, national origin, social or ethnic origin, sex, sexual orientation, gender, gender identity, gender expression, marital status, pregnancy, parental status, disability, medical condition, genetic information, military or veteran status, union membership, and political affiliation, or any other legally protected group.

Job Req Type: Experienced

          

Required Travel: Yes, 10% of the time

          

Shift Type: 1st Shift/Days

Skills Required

  • Minimum 3+ years experience in monitoring, observability, systems engineering, or DevOps
  • Hands-on expertise with Prometheus
  • Hands-on expertise with Grafana
  • Hands-on expertise with Nagios
  • Experience with SolarWinds enterprise monitoring
  • Experience with Apache Kafka and event-driven architectures
  • Proficiency in bash scripting and Linux system administration (RHEL/CentOS/Ubuntu)
  • Proficiency in Python or Go for custom tooling
  • Knowledge of metrics collection, time-series data management, and observability across systems, storage, applications, and networks
  • Experience implementing AI/ML-based monitoring or anomaly detection
  • Experience with Infrastructure-as-Code tools (Terraform, Ansible, etc.)
  • Strong problem-solving and communication skills; ability to present technical concepts to varied stakeholders
  • Experience with containerized environments and Kubernetes monitoring
  • Experience with cloud platforms (AWS, Azure, GCP) and native monitoring services
  • Experience with observability backends such as OpenTelemetry and APM tools
  • Experience with incident response and post-mortem processes
  • Relevant certifications (Kubernetes, cloud platforms, monitoring tools)
  • Contributions to open-source monitoring projects

Analog Devices Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Analog Devices and has not been reviewed or approved by Analog Devices.

  • Retirement Support The 401(k) program is described as a standout feature, with company contribution up to 8% of base salary and immediate vesting. This structure strengthens long-term value even when cash compensation perceptions vary.
  • Healthcare Strength Health coverage is positioned as comprehensive, including medical, dental, and vision options along with disability and life insurance. Day-one eligibility and multiple plan choices add to perceived robustness.
  • Leave & Time Off Breadth Paid time off appears broad, with vacation ranging from roughly 17–25 days and increasing up to five weeks with tenure, alongside sick time and paid holidays. Parental leave and related time-off provisions further expand coverage.

Analog Devices Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Wilmington, MA
20,292 Employees
Year Founded: 1965

What We Do

Analog Devices, Inc. (NASDAQ: ADI) operates at the center of the modern digital economy, converting real-world phenomena into actionable insight with its comprehensive suite of analog and mixed signal, power management, radio frequency (RF), and digital and sensor technologies. ADI serves 125,000 customers worldwide with more than 75,000 products in the industrial, communications, automotive, and consumer markets. ADI is headquartered in Wilmington, MA.

Similar Jobs

Wells Fargo Logo Wells Fargo

Site Reliability Engineer

Fintech • Financial Services
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
205000 Employees

Analog Devices Logo Analog Devices

Sr. Monitoring and Observability Engineer

Artificial Intelligence • Hardware • Automation • Manufacturing
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
20292 Employees

Airwallex Logo Airwallex

Sales Development Representative

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
2300 Employees
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees

Similar Companies Hiring

Legora Thumbnail
Artificial Intelligence • Legal Tech • Software
New York, New York
700 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account