Senior Lead Observability - Dynatrace

Posted 3 Hours Ago
Be an Early Applicant
2 Locations
Hybrid
Senior level
Artificial Intelligence • Cloud • Fintech • Information Technology • Analytics • Financial Services • Cybersecurity
Dynamic careers. Brighter futures. Achieve greater.
The Role
Hands-on individual contributor who designs, implements, and automates enterprise observability across Dynatrace, SCOM, and ServiceNow ITOM. Improve alert quality, event correlation, onboarding, dashboards, and integrations using scripting and CI/CD. Partner with application, infrastructure, and operations teams to reduce noise, accelerate incident response, and increase monitoring coverage and automation.
Summary Generated by Built In

About Northern Trust


As a global leader in innovative wealth management, asset servicing, asset management and banking services, Northern Trust (Nasdaq: NTRS) is proud to guide the world’s most successful individuals, families, corporations and institutions.


Since 1889, we have aligned our efforts with our three guiding Principles That Endure: Service, Expertise, and Integrity. Together, they reflect the three cornerstones of business conduct which we strive to instill in our employees, whom we call partners, and to provide to our clients and the communities we serve worldwide.


With more than 135 years of financial experience and over 24,000 partners, we serve the world’s most sophisticated clients using leading technology and exceptional service.


Senior Lead, Observability – DynatraceRole Overview

We are seeking a highly skilled Senior Observability Engineer to support and enhance enterprise observability capabilities across infrastructure, applications, cloud, and business services. This is an individual contributor role focused on hands-on engineering, platform operations, monitoring standardization, automation, and continuous improvement across observability platforms including Dynatrace, Microsoft SCOM, ServiceNow ITOM & Event Management, and related monitoring integrations.

The successful candidate will play a key role in improving monitoring effectiveness, reducing alert noise, enhancing event correlation, enabling faster incident response, and supporting operational resilience through automation and observability best practices. The role requires strong technical expertise, analytical thinking, and the ability to work closely with platform, application, infrastructure, and operations teams to deliver reliable and scalable monitoring solutions.

This role builds on the observability focus outlined in the reference document, including enterprise visibility across applications, infrastructure, cloud, and business services through Dynatrace, ServiceNow ITOM Event Management, AI-driven observability, and automation-led efficiency improvements.

Key Responsibilities
  • Implement, maintain, and continuously improve enterprise observability capabilities across Dynatrace, SCOM, ServiceNow ITOM & Event Management, and supporting monitoring tools.
  • Configure and support monitoring for infrastructure, applications, services, databases, middleware, cloud, and hybrid environments to ensure end-to-end visibility and operational stability.
  • Develop, tune, and optimize monitoring alerts, dashboards, thresholds, synthetic checks, anomaly detection, and service-level views to improve signal quality and reduce false positives.
  • Drive alert hygiene activities, including duplicate alert reduction, threshold tuning, suppression logic, event enrichment, and monitoring standardization across technology teams.
  • Support integrations between observability platforms and ServiceNow Event Management, ensuring events are enriched, correlated, deduplicated, and routed effectively for incident response.
  • Manage and enhance monitoring integrations from tools such as Dynatrace, SCOM, infrastructure monitoring sources, and application monitoring platforms into ServiceNow Event Management.
  • Use Dynatrace capabilities such as service flow, distributed tracing, Davis AI, anomaly detection, problem correlation, dashboards, management zones, tags, and alerting profiles to improve root cause analysis and operational insights.
  • Support Microsoft SCOM monitoring activities including management pack configuration, alert rule tuning, agent health checks, monitoring coverage, and operational troubleshooting.
  • Build automation scripts and reusable solutions using Python, PowerShell, REST APIs, YAML/JSON, CI/CD pipelines, and other automation frameworks to improve onboarding, monitoring configuration, health checks, reporting, and operational efficiency.
  • Contribute to observability-as-code practices by supporting standardized, repeatable, and automated monitoring onboarding patterns.
  • Partner with application, infrastructure, cloud, and support teams to onboard new applications and services into enterprise monitoring platforms.
  • Troubleshoot monitoring gaps, integration failures, agent issues, event flow problems, and alerting defects across observability systems.
  • Support operational reporting and KPI tracking related to alert volume, noise reduction, MTTR improvement, event quality, monitoring coverage, and automation adoption.
  • Maintain technical documentation, operational runbooks, configuration standards, troubleshooting guides, and onboarding procedures.
  • Participate in incident reviews and problem management discussions to identify opportunities for monitoring improvement and proactive detection.
  • Apply ITIL practices and event lifecycle management principles to improve incident quality, operational response, and service reliability.
Required Skills and Experience
  • 12+ years of overall IT experience, with strong hands-on experience in observability, monitoring, event management, infrastructure operations, application support, or platform engineering.
  • Strong practical experience with Dynatrace including OneAgent, dashboards, alerts, management zones, synthetic monitoring, service flow, problem detection, tagging, and Davis AI capabilities.
  • Hands-on experience with Microsoft SCOM, including alert configuration, management packs, agent monitoring, rule tuning, infrastructure monitoring, and operational troubleshooting.
  • Strong working knowledge of ServiceNow Event Management / ITOM, including event ingestion, event rules, alert correlation, deduplication, enrichment, service mapping awareness, and incident integration.
  • Experience integrating monitoring tools with ServiceNow or similar ITSM platforms.
  • Good understanding of observability concepts including metrics, logs, traces, events, topology, service health, synthetic monitoring, and full-stack monitoring.
  • Strong scripting and automation experience using Python, PowerShell, REST APIs, shell scripting, or similar technologies.
  • Experience with CI/CD tools, source control, configuration files, automation workflows, and infrastructure-as-code or observability-as-code practices.
  • Familiarity with cloud and hybrid environments such as Azure, AWS, VMware, Windows, Linux, databases, middleware, and enterprise infrastructure platforms.
  • Understanding of ITIL processes, especially Incident Management, Problem Management, Change Management, Event Management, and operational support models.
  • Strong analytical and troubleshooting skills with the ability to identify monitoring gaps, alert quality issues, and event flow problems.
  • Ability to work independently as an individual contributor while collaborating effectively with engineering, operations, application, and support teams.
Preferred Skills
  • Experience with OpenTelemetry, log analytics, AIOps, machine learning-based monitoring, predictive monitoring, or GenAI-assisted incident diagnosis.
  • Exposure to enterprise observability platforms beyond Dynatrace and SCOM, such as Splunk, Elastic, Azure Monitor, Grafana, Prometheus, or other monitoring ecosystems.
  • Experience building dashboards, operational reports, executive metrics, and service health views.
  • Knowledge of CMDB, service mapping, application dependency mapping, and event-to-CI relationships in ServiceNow.
  • Experience supporting large-scale enterprise monitoring environments in financial services or regulated industries.
  • Familiarity with Ansible, Terraform, GitHub, Azure DevOps, Jenkins, or other automation and DevOps tooling.
  • Ability to contribute to monitoring governance, standards, best practices, and platform maturity improvements.
Certifications

Preferred certifications include:

  • Dynatrace Associate, Professional, or equivalent certification.
  • Microsoft SCOM or Microsoft infrastructure/platform certifications.
  • ServiceNow ITOM / Event Management certification.
  • Microsoft Azure or AWS certification.
  • ITIL Foundation or higher certification.
Success Measures

The individual in this role will be measured on the ability to deliver measurable improvements in observability maturity, monitoring quality, automation, and operational effectiveness. Key success measures include:

  • Improved incident quality and faster triage through actionable alerts and contextual event information.
  • Increased monitoring coverage across applications, infrastructure, and services.
  • Faster and more consistent onboarding of applications into observability platforms.
  • Increased use of automation for repeatable monitoring and operational tasks.
  • Improved dashboard quality, service health visibility, and operational reporting.
  • Stronger alignment with enterprise monitoring standards and ITIL event management practices.
  • Reduction in duplicate, noisy, unactionable, or low-value alerts.
  • Improved signal-to-noise ratio across monitoring platforms.
  • Better event correlation and enrichment through ServiceNow Event Management.

Working with Us


As a Northern Trust partner, you will be part of a flexible and collaborative work culture, which has a strong history of financial strength and stability. Movement within the organization is encouraged, senior leaders are accessible, and you can take pride in working for a company committed to an inclusive workplace and assisting the communities we serve.


Philanthropy is deeply rooted in Northern Trust’s history and is an essential element of our culture. Employees around the world give their time and talent to work for the greater good of their communities.


Reasonable Accommodation


Northern Trust is committed to working with and providing adjustments to individuals with health conditions and disabilities. If you need a reasonable accommodation for any part of the employment process, please email our HR Service Center at [email protected], or alternatively you can discuss your individual requirements with the recruiter you are working with.


About Our Pune Office


The Northern Trust Pune office, established in 2016, is now home to over 3,000 employees. The office handles various functions, including Operations for Asset Servicing and Wealth Management, as well as delivering critical technology solutions that support business operations across the globe.


Our Pune team takes our commitment to service to heart. In 2024, they volunteered more than 10,000+ hours into the communities where they live and work. Learn more.

Skills Required

  • 12+ years of overall IT experience in observability, monitoring, event management, infrastructure operations, application support, or platform engineering
  • Strong practical experience with Dynatrace (OneAgent, dashboards, synthetic monitoring, service flow, tagging, Davis AI, alerting)
  • Hands-on experience with Microsoft SCOM (management packs, alert configuration, agent monitoring, rule tuning, troubleshooting)
  • Working knowledge of ServiceNow Event Management / ITOM (event ingestion, correlation, deduplication, enrichment, incident integration, service mapping awareness)
  • Experience integrating monitoring tools with ServiceNow or similar ITSM platforms
  • Strong scripting and automation using Python, PowerShell, REST APIs, shell scripting, YAML/JSON
  • Experience with CI/CD tools, source control, automation workflows, and observability-as-code or infrastructure-as-code practices
  • Familiarity with cloud and hybrid environments (Azure, AWS, VMware) and infrastructure platforms (Windows, Linux, databases, middleware)
  • Understanding of ITIL processes (Incident, Problem, Change, Event Management) and event lifecycle management principles
  • Strong analytical and troubleshooting skills to identify monitoring gaps, alert quality issues, and integration failures
  • Ability to work independently as an individual contributor and collaborate with engineering, operations, application, and support teams
  • Experience with OpenTelemetry, log analytics, AIOps, machine learning-based monitoring, or GenAI-assisted incident diagnosis
  • Exposure to Splunk, Elastic, Azure Monitor, Grafana, Prometheus or other monitoring ecosystems
  • Familiarity with Ansible, Terraform, GitHub, Azure DevOps, Jenkins or other automation/DevOps tooling
  • Experience supporting large-scale enterprise monitoring environments in financial services or regulated industries
  • Preferred certifications: Dynatrace Associate/Professional, Microsoft SCOM/infra certs, ServiceNow ITOM/Event Management, Azure/AWS cert, ITIL Foundation

Northern Trust Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Northern Trust and has not been reviewed or approved by Northern Trust.

  • Retirement Support A 401(k) with company match alongside an employer‑funded defined‑benefit pension is highlighted in employer‑verified materials and filings, setting the retirement package apart from many private employers.
  • Leave & Time Off Breadth PTO is portrayed as solid overall, and two paid volunteer days per year are clearly documented across corporate materials.
  • Parental & Family Support Company sources and postings reference paid parental and caregiver leave and recent enhancements to parental benefits, indicating meaningful family support within the package.

Northern Trust Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Chicago, IL
24,000 Employees
Year Founded: 1889

What We Do

As a global leader in innovative wealth management, asset servicing and investment solutions, Northern Trust (Nasdaq: NTRS) is proud to guide the world’s most successful individuals, families and institutions by remaining true to our enduring principles of service, expertise and integrity. A globally recognized Fortune 500 Company in continuous operation since 1889, we’ve built a legacy of empowering clients to reach their goals with confidence. Since our roots as a trust bank, we’ve grown to a global presence with more than 24,000 employees in more than 20 countries and across six core business units: Wealth Management Asset Management Asset Servicing Technology Corporate Functions Enterprise Operations Join a Team That’s Achieving Greater At Northern Trust, we refer to our employees as partners – with good reason. We understand that relationships are the key to our success. Here you’ll join a diverse and inclusive team of innovators with the drive to challenge the way things have always been done. Instead of choosing between a dynamic career and work-life balance, enjoy working with a team that supports your goals in the office and at home. We’ll help you get where you want to go without sacrificing what matters most to you. Delivering value and adhering to our enduring principles What are enduring principles? Since our founding, they have guided our strategy and success. Thanks to the dedication of our partners, Northern Trust continues to thrive by adhering to three enduring principles: service, expertise and integrity . What does this mean? Service Northern Trust has a relentless drive to provide exceptional service to our clients, our partners and our communities. We set new standards and go above and beyond in our commitment to delivering greater results. Expertise Expertise is at the core of who we are. We focus sharply on what we do well. From expanding our capabilities, to hiring talented professionals to developing innovative solutions, our expertise is why we continue to be a trusted advisor for generations of families and institutions. Integrity Operating with uncompromising ethics is central to Northern Trust’s heritage. As a result, our clients, partners and communities know they can rely on us. For more than 130 years, our integrity has been our guide – and that will never change.

Why Work With Us

At Northern Trust, we go further because we go together. We embrace flexibility, encourage balance, and prioritize inclusion at all levels, working together to keep you connected. We are committed to our employees—all 24,000 of them. Whether this is a first step or a bold new leap in your career, we’re here to help you move forward.

Gallery

Gallery

Similar Jobs

Northern Trust Logo Northern Trust

Senior Lead - Observability – Dynatrace

Artificial Intelligence • Cloud • Fintech • Information Technology • Analytics • Financial Services • Cybersecurity
Hybrid
2 Locations
24000 Employees

Northern Trust Logo Northern Trust

Senior Lead – Observability (ServiceNow ITOM & Dynatrace)

Artificial Intelligence • Cloud • Fintech • Information Technology • Analytics • Financial Services • Cybersecurity
Hybrid
Pune, Maharashtra, IND
24000 Employees

Morningstar Logo Morningstar

Apprentice

Artificial Intelligence • Big Data • Enterprise Web • Fintech • Software • Financial Services
Hybrid
Mumbai, Maharashtra, IND
11500 Employees
22K-22K Annually
Hybrid
Mumbai, Maharashtra, IND
289097 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account