Site Reliability Engineer (NOC Automation) - Splunk

Posted 12 Days Ago
Be an Early Applicant
Kraków, Małopolskie, POL
In-Office
Mid level
Cloud • Information Technology • Internet of Things • Professional Services • Software
The Role
Support large-scale Splunk Cloud deployments through monitoring, incident response, troubleshooting, on-call operations, and production support. Develop reliable automation by converting manual runbooks into repeatable workflows. Investigate issues across cloud infrastructure, Linux, networking, distributed services, and application code. Participate in post-incident reviews, code reviews, documentation, and cross-functional efforts to improve service reliability and reduce operational toil.
Summary Generated by Built In
Meet the team

Join us as we pursue our disruptive new vision to make machine data accessible, usable and valuable to everyone. We are a company filled with people who are passionate about our product and seek to deliver the best experience for our customers. At Cisco, we are building a more resilient digital world through our security and observability portfolio. Splunk, a Cisco company, helps organizations turn data into action and strengthen the resilience of their digital systems. Learn more about Cisco careers and how you can become a part of our journey!


Splunk Cloud is looking for an operations and automation engineer to support Network Operations Center (NOC) responsibilities and manage large-scale Splunk Cloud deployments. This role balances hands-on operational support with the development of reliable, maintainable automation.


Your impact

You will monitor and troubleshoot distributed cloud systems, respond to production incidents, and turn manual operational runbooks into safe, repeatable workflows. You will work with operations and engineering teams to improve how we detect, diagnose, and resolve service issues while reducing manual toil.


What you'll do
  • Monitor and support large-scale Splunk Cloud deployments, ensuring service reliability, availability, and performance.
  • Respond to monitoring alerts and production incidents according to defined playbooks and procedures.
  • Participate in 12/7 operations and on-call support
  • Investigate complex issues across cloud infrastructure, Linux systems, networking, distributed services, and application code.
  • Participate in post-incident reviews and turn findings into operational and automation improvements.
  • Identify repetitive, manual, or error-prone operational tasks that are suitable for automation.
  • Convert operational runbooks into reliable automated workflows.
  • Use Git-based development, peer review, documentation, and cross-functional collaboration to maintain and improve operational automation.

Minimum qualifications
  • B.S. in a related field or equivalent work experience (5+ years).
  • 3+ years of experience in systems or network administration in a cloud environment, including incident response and major incident management
  • Proven experience developing maintainable python/go automation, operational tooling, or production support utilities.
  • Experience using Unix or Linux systems and shell scripting.
  • Experience using Git and participating in code review workflows.
  • Working knowledge of software engineering practices, including modular design, testing, debugging, exception handling, logging, and documentation.
  • Ability to translate a manual operational procedure into a safe, repeatable automated workflow.
  • Understanding of production automation risks, including permissions, secrets management, validation, rollback, and human approval points.
  • Strong troubleshooting, prioritization, collaboration, and communication skills, including the ability to remain effective during major service outages.
Why Cisco? 

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.

Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. 

We are Cisco, and our power starts with you. 

Skills Required

  • Bachelor of Science degree in a related field or equivalent work experience of 5 or more years
  • 3 or more years of systems or network administration experience in a cloud environment
  • Experience with incident response and major incident management
  • Experience developing maintainable Python or Go automation, operational tooling, or production support utilities
  • Experience using Unix or Linux systems and shell scripting
  • Experience using Git and participating in code review workflows
  • Working knowledge of software engineering practices, including modular design, testing, debugging, exception handling, logging, and documentation
  • Ability to translate manual operational procedures into safe, repeatable automated workflows
  • Understanding of production automation risks, including permissions, secrets management, validation, rollback, and human approval points
  • Strong troubleshooting, prioritization, collaboration, and communication skills, including effectiveness during major service outages

Cisco Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Cisco and has not been reviewed or approved by Cisco.

  • Healthcare Strength — Health coverage is described as broad, with medical options including PPOs, high-deductible plans, and regional HMOs. Feedback suggests robust wellness and mental-health resources, with some locations offering on-site support and second medical opinions.
  • Leave & Time Off Breadth — Time off is highlighted through paid holidays, paid time off, and additional recharge days such as “Days for Me,” with generous volunteer time also mentioned. Feedback suggests these programs help support rest, volunteering, and critical life events.
  • Parental & Family Support — Parental and family support appears extensive, including paid child-bonding leave, family medical leave, and caregiving resources. Observations also point to family-planning assistance and adoption or surrogacy support in some regions.

Cisco Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Jose, CA
77,500 Employees
Year Founded: 1984

What We Do

Cisco (NASDAQ: CSCO) enables people to make powerful connections--whether in business, education, philanthropy, or creativity. Cisco hardware, software, and service offerings are used to create the Internet solutions that make networks possible--providing easy access to information anywhere, at any time. Cisco was founded in 1984 by a small group of computer scientists from Stanford University. Since the company's inception, Cisco engineers have been leaders in the development of Internet Protocol (IP)-based networking technologies. Today, with more than 71,000 employees worldwide, this tradition of innovation continues with industry-leading products and solutions in the company's core development areas of routing and switching, as well as in advanced technologies such as home networking, IP telephony, optical networking, security, storage area networking, and wireless technology. In addition to its products, Cisco provides a broad range of service offerings, including technical support and advanced services. Cisco sells its products and services, both directly through its own sales force as well as through its channel partners, to large enterprises, commercial businesses, service providers, and consumers.

Similar Jobs

GitLab Logo GitLab

Systems Engineer

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
In-Office or Remote
2 Locations
2500 Employees

Circle (circle.so) Logo Circle (circle.so)

Lead Engineer, AI Platform

Artificial Intelligence • Consumer Web • Digital Media • Information Technology • Social Impact • Software
In-Office or Remote
43 Locations
250 Employees

Sprout Social Logo Sprout Social

Senior Workplace Experience Specialist (Part-Time, B2B Contract)

Marketing Tech • Social Media • Software • Analytics • Business Intelligence
Easy Apply
Remote or Hybrid
Kraków, Małopolskie, POL
1400 Employees
70K-130K Annually

Capco Logo Capco

Business Analyst

Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Remote or Hybrid
Poland
6000 Employees

Similar Companies Hiring

Ford Energy Thumbnail
Automotive • Software • Energy • Utilities • Manufacturing • Renewable Energy
US
55 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account