The Role
Build AI agents that analyze ServiceNow Discovery logs, classify failures, identify root causes, recommend or generate probe and pattern corrections, and monitor MID Server health, throughput, latency, and coverage. Convert troubleshooting knowledge into agentic workflows with testing guardrails and human approval. The role also involves ServiceNow Discovery, CMDB/CSDM data quality, observability integrations, infrastructure troubleshooting, and production-grade automation.
Summary Generated by Built In
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.
WHY JOIN US
If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!
ABOUT THE ROLE
We are looking for a ServiceNow Discovery Engineer to join an enterprise-scale Discovery Engineering squad that executes discovery across millions of infrastructure targets on daily and weekly schedules. You will design and build AI agents that analyze Discovery logs, classify failures, identify root causes, and generate corrections to Discovery probes and patterns with appropriate testing guardrails, while also developing observability agents that continuously assess MID Server health, throughput, and coverage. The role requires 4+ years of hands-on ServiceNow Discovery experience including MID Servers, Patterns/Probes, and CMDB/CSDM data quality.
WHAT YOU WILL DO
- Design and build AI agents that continuously analyze Discovery logs, execution results, errors, and failed discoveries at scale.
- Develop agents capable of classifying and correlating failures, performing troubleshooting, identifying probable root causes, and recommending corrective actions.
- Build capabilities for agents to analyze and recommend or generate corrections to ServiceNow Discovery probes and patterns, with appropriate testing, guardrails, and human approval before production changes.
- Convert operational knowledge, runbooks, and recurring troubleshooting procedures into reusable agentic workflows, progressively reducing manual investigation and improving Discovery reliability.
- Design and build AI agents that continuously assess the health of the ServiceNow Discovery ecosystem, including MID Servers, discovery schedules/playbooks, target execution, throughput, latency, failures, retries, and discovery coverage.
- Develop agents that regularly collect, aggregate, and interpret metrics and telemetry from ServiceNow Discovery, MID Servers, logs, monitoring platforms, and other relevant data sources.
MUST HAVES
- Hands-on experience with ServiceNow Discovery of at least 4 years, including MID Servers, Discovery Patterns/Probes, credentials, schedules, Discovery Status, ECC Queue, and troubleshooting Discovery failures.
- Experience developing or modifying ServiceNow Discovery Patterns and Probes and understanding how infrastructure attributes and relationships are discovered.
- Experience with ServiceNow CMDB/CSDM, Configuration Items (CIs), identification/reconciliation, relationships, and CMDB data quality.
- Experience with observability technologies such as Prometheus, OpenTelemetry, and Grafana.
- Upper-intermediate English level.
NICE TO HAVES
- Advanced Python development experience and REST API design capabilities with experience building APIs, data-processing pipelines, integrations, automation, and production-grade engineering solutions.
- Strong experience with modern software development and SDLC practices and toolchains, including Git/GitHub, Jenkins or equivalent CI/CD platforms, artifact repositories, automated testing, code quality/security scanning, release management, and deployment automation.
- Practical experience integrating Gen AI APIs into software applications (e.g., OpenAI, Anthropic) and a working understanding of developer-level concepts like Retrieval-Augmented Generation (RAG) and Vector databases. (Note: We need software builders, not Machine Learning researchers).
- Practical knowledge of Infrastructure Engineering with Linux, Windows Server, compute, virtualization/cloud infrastructure, authentication, processes/services, and infrastructure troubleshooting.
- Working knowledge of TCP/IP, DNS, routing, firewalls, ports, SSH, WMI/WinRM, SNMP, HTTP/S, and common network and infrastructure troubleshooting techniques.
- Experience with high-volume logs, metrics, errors, and operational datasets to identify patterns, anomalies, trends, and root causes and turn them into actionable engineering improvements.
PERKS AND BENEFITS
- Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
- Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
- Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
- Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
- Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
- Well-being & support: access local well-being programs and people-focused support tailored to your location
Meet Our Recruitment Process
Application → Coding Challenge → Video Interview → Technical Interview or Hiring Manager Interview
Each step helps us understand your skills and overall fit.
If it’s a match, you’ll receive an offer.
Skills Required
- At least 4 years of hands-on ServiceNow Discovery experience
- Experience with MID Servers, Discovery Patterns and Probes, credentials, schedules, Discovery Status, ECC Queue, and troubleshooting Discovery failures
- Experience developing or modifying ServiceNow Discovery Patterns and Probes
- Understanding of infrastructure attributes and relationships discovered through ServiceNow Discovery
- Experience with ServiceNow CMDB/CSDM, Configuration Items, identification and reconciliation, relationships, and CMDB data quality
- Experience with Prometheus, OpenTelemetry, and Grafana
- Upper-intermediate English proficiency
- Advanced Python development experience
- REST API design and experience building APIs, data-processing pipelines, integrations, automation, and production-grade engineering solutions
- Experience with modern software development and SDLC practices, Git/GitHub, CI/CD, automated testing, code quality and security scanning, release management, and deployment automation
- Practical experience integrating generative AI APIs such as OpenAI or Anthropic
- Working knowledge of Retrieval-Augmented Generation and vector databases
- Infrastructure Engineering experience with Linux, Windows Server, compute, virtualization or cloud infrastructure, authentication, processes, services, and troubleshooting
- Working knowledge of TCP/IP, DNS, routing, firewalls, ports, SSH, WMI/WinRM, SNMP, HTTP/S, and network troubleshooting
- Experience analyzing high-volume logs, metrics, errors, and operational datasets to identify patterns, anomalies, trends, and root causes
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
AgileEngine is a privately held company established in 2010 that builds dedicated teams of designers and developers. We turn good ideas into awesome software that people actually want to use. Some of the biggest names and the hottest startups around the world chose us to build their tech.


.png)






