Site Reliability Engineer

Posted 2 Days Ago
Atlanta, GA, USA
In-Office
Junior
Fintech • Information Technology • Payments • Software
The Role
Supports and scales cloud infrastructure and applications through automation, monitoring, incident response, troubleshooting, infrastructure as code, and CI/CD improvements. Partners with software and platform engineering teams to improve reliability, scalability, security, and operational readiness. Participates in on-call rotations, root cause analysis, documentation, and preventive solution implementation for production systems.
Summary Generated by Built In

About NCR VOYIX

NCR Voyix Corporation (NYSE: VYX) is a global platform-powered leader in unified commerce for shopping and dining. Combining a flexible, intelligent platform with end-to-end payments capabilities and services developed through its deep industry experience, NCR Voyix empowers retailers and restaurants to accelerate new possibilities for their operations, experiences and business outcomes. NCR Voyix is headquartered in Atlanta, Georgia, and serves customers in more than 35 countries worldwide.Site Reliability Engineer

Onsite: Atlanta, GA


Job Summary

At NCR Voyix, we're looking for a Site Reliability Engineer II to help build, support, and scale the cloud platforms that power our products and services. This role is ideal for an engineer who enjoys solving technical challenges, automating manual processes, and improving the reliability and performance of applications in a cloud-native environment.

As part of our Engineering organization, you'll work closely with software engineers, cloud architects, and operations teams to ensure our platforms remain secure, available, and scalable. You'll have the opportunity to gain hands-on experience with cloud technologies, automation, observability tools, and modern infrastructure practices while contributing to mission-critical systems used by customers around the world.

What You'll Do
  • Support and enhance cloud-based infrastructure and applications across modern cloud platforms.
  • Implement automation solutions that improve operational efficiency, reduce manual effort, and increase system reliability.
  • Monitor application and infrastructure health using observability and monitoring tools to identify and resolve performance issues.
  • Participate in incident response and on-call support rotations for production systems.
  • Troubleshoot infrastructure, application, and deployment issues in partnership with engineering teams.
  • Assist with root cause analysis (RCA) activities and support implementation of preventive solutions.
  • Contribute to Infrastructure as Code (IaC) initiatives using modern automation and cloud provisioning tools.
  • Partner with software development teams to improve application reliability, scalability, and operational readiness.
  • Support CI/CD pipelines and deployment automation processes.
  • Create and maintain technical documentation, operational runbooks, and troubleshooting guides.
  • Stay current with emerging cloud and platform technologies and recommend improvements where applicable.
Required Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.
  • 1-3 years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, Systems Engineering, Infrastructure Engineering, or a related technical role.
  • Experience with at least one public cloud platform such as AWS, Azure, or Google Cloud Platform (GCP).
  • Experience with scripting or programming languages such as Python, Bash, Go, or PowerShell.
  • Familiarity with Linux operating systems and basic system administration concepts.
  • Understanding of cloud infrastructure, networking, and application architecture fundamentals.
  • Familiarity with container technologies such as Docker and Kubernetes.
  • Knowledge of CI/CD concepts and experience with tools such as GitHub Actions, GitLab CI, Jenkins, or similar platforms.
  • Exposure to monitoring and observability tools such as Datadog, Grafana, Prometheus, Splunk, or ELK.
  • Strong troubleshooting, analytical, and problem-solving skills.
  • Excellent communication and collaboration skills with the ability to work effectively across technical teams.
Preferred Qualifications
  • Exposure to Infrastructure as Code tools such as Terraform, CloudFormation, or Pulumi.
  • Understanding of SRE concepts including service availability, reliability, incident management, and operational excellence.
  • Experience supporting production applications in a cloud-native environment.
  • Familiarity with Agile software development methodologies.
  • Relevant cloud certifications (AWS, Azure, or GCP) are a plus.
What Success Looks Like
  • Contributes to reliable, scalable, and secure cloud services.
  • Helps reduce operational toil through automation and process improvements.
  • Responds effectively to incidents and contributes to long-term reliability improvements.
  • Builds strong partnerships with development and platform engineering teams.
  • Continuously expands technical expertise while supporting business-critical systems.

This job description outlines the primary responsibilities of the position and is not intended to be an exhaustive list of duties. Additional responsibilities may be assigned based on business needs and individual skills and experience.

Offers of employment are conditional upon passage of screening criteria applicable to the job

EEO Statement

Integrated into our shared values is NCR Voyix’s commitment to equal employment opportunity.  All qualified applicants will receive consideration for employment without regard to sex, age, race, color, creed, religion, national origin, disability, sexual orientation, gender identity, veteran status, military service, genetic information, or any other characteristic or conduct protected by law.  NCR Voyix is committed to being a globally inclusive company where all people are treated fairly, recognized for their individuality, promoted based on performance and encouraged to strive to reach their full potential.  We believe in understanding and respecting differences among all people.  Every individual at NCR Voyix has an ongoing responsibility to respect and support a globally diverse environment.

Statement to Third Party Agencies
To ALL recruitment agencies: NCR Voyix only accepts resumes from agencies on the preferred supplier list. Please do not forward resumes to our applicant tracking system, NCR Voyix employees, or any NCR Voyix facility. NCR Voyix is not responsible for any fees or charges associated with unsolicited resumes

“When applying for a job, please make sure to only open emails that you will receive during your application process that come from a @ncrvoyix.com email domain.”

Skills Required

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience
  • 1-3 years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, Systems Engineering, Infrastructure Engineering, or a related technical role
  • Experience with at least one public cloud platform such as AWS, Azure, or Google Cloud Platform
  • Experience with scripting or programming languages such as Python, Bash, Go, or PowerShell
  • Familiarity with Linux operating systems and basic system administration
  • Understanding of cloud infrastructure, networking, and application architecture fundamentals
  • Familiarity with Docker and Kubernetes
  • Knowledge of CI/CD concepts and experience with GitHub Actions, GitLab CI, Jenkins, or similar platforms
  • Exposure to monitoring and observability tools such as Datadog, Grafana, Prometheus, Splunk, or ELK
  • Strong troubleshooting, analytical, and problem-solving skills
  • Excellent communication and collaboration skills across technical teams
  • Exposure to Infrastructure as Code tools such as Terraform, CloudFormation, or Pulumi
  • Understanding of SRE concepts including availability, reliability, incident management, and operational excellence
  • Experience supporting production applications in a cloud-native environment
  • Familiarity with Agile software development methodologies
  • Relevant AWS, Azure, or GCP cloud certification

NCR Corporation Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about NCR Corporation and has not been reviewed or approved by NCR Corporation.

  • Healthcare Strength — Healthcare coverage is described as comprehensive, including medical, dental, and vision, alongside HSA/HRA funding and an employee assistance program. This breadth increases the perceived baseline value of the total rewards package even when pay satisfaction varies.
  • Retirement Support — Retirement support includes a 401(k) match structure and an employee stock purchase program with a stated discount. These elements provide longer-term wealth-building mechanisms beyond base salary.
  • Leave & Time Off Breadth — Time-off provisions include paid vacation, holidays (including floating days), sick time, and defined maternity and paternity leave. This breadth can improve the overall rewards experience for those who can fully utilize the leave policies.

NCR Corporation Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Atlanta, GA
36,000 Employees
Year Founded: 1884

What We Do

Shaping the future for 135 years, NCR is the world’s enterprise technology leader for restaurants, retailers and banks. The #1 global POS software provider for retail and hospitality, and the #1 provider of multi-vendor ATM software, we create software, hardware and services that run the enterprise from back office to the front end and everything in between for our clients.

Similar Jobs

MetLife Logo MetLife

Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Remote or Hybrid
United States
43000 Employees
111K-180K Annually

MetLife Logo MetLife

Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Remote or Hybrid
United States
43000 Employees
111K-180K Annually

PwC Logo PwC

Site Reliability Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
58 Locations
370000 Employees
151K-187K Annually

Jackpocket Logo Jackpocket

Site Reliability Engineer

Consumer Web • Gaming • Mobile • News + Entertainment • Software
Remote or Hybrid
United States
330 Employees
168K-210K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account