Senior Site Reliability Engineer (SRE)

Posted 6 Hours Ago
Be an Early Applicant
Atlanta, GA, USA
Hybrid
115K-140K Annually
Senior level
Energy
The Role
Build and support scalable, resilient cloud-native platforms and applications. Responsibilities include reliability engineering, observability, incident response, platform engineering, cloud infrastructure automation, Infrastructure as Code, CI/CD enablement, disaster recovery, chaos engineering, and AI-driven operations. The role leads reliability improvements, develops self-service tooling, partners cross-functionally on platform modernization, and mentors junior engineers.
Summary Generated by Built In

Job Title:

Senior Site Reliability Engineer (SRE)

Job Description

We're Concentrix. The intelligent transformation partner. Solution-focused. Tech-powered. Intelligence-fueled.
The global technology and services leader that powers the world’s best brands, today and into the future. We’re solution-focused, tech-powered, intelligence-fueled. With unique data and insights, deep industry expertise, and advanced technology solutions, we’re the intelligent transformation partner that powers a world that works, helping companies become refreshingly simple to work, interact, and transact with. We shape new game-changing careers in over 70 countries, attracting the best talent.
The Concentrix Technical Products and Services team is the driving force behind Concentrix’s transformation, data, and technology services. We integrate world-class digital engineering, creativity, and a deep understanding of human behavior to find and unlock value through tech-powered and intelligence-fueled experiences. We combine human-centered design, powerful data, and strong tech to accelerate transformation at scale. You will be surrounded by the best in the world providing market leading technology and insights to modernize and simplify the customer experience. Within our professional services team, you will deliver strategic consulting, design, advisory services, market research, and contact center analytics that deliver insights to improve outcomes and value for our clients. Hence achieving our vision.
Our gamechangers around the world have devoted their careers to ensuring every relationship is exceptional. And we’re proud to be recognized with awards such as "World's Best Workplaces," “Best Companies for Career Growth,” and “Best Company Culture,” year after year.
Join us and be part of this journey towards greater opportunities and brighter futures.

Job Summary:

We are seeking a Senior Site Reliability Engineer (SRE) to build, automate, and support highly scalable cloud-native platforms and digital applications. This role is responsible for improving system reliability, observability, performance, and operational excellence through Infrastructure as Code, Kubernetes, CI/CD automation, monitoring, and incident management. The ideal candidate will have strong experience with AWS/Azure, distributed systems, platform engineering, and enterprise observability tools, with a passion for reducing operational toil and enhancing platform resilience through automation and intelligent operations.

Responsibilities:

Reliability Engineering & Operational Excellence

  • Design, implement, and support highly available, scalable, and resilient cloud-native platforms and services.
  • Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to ensure platform reliability.
  • Identify, analyze, and eliminate reliability and performance bottlenecks across distributed systems.
  • Lead incident response, Root Cause Analysis (RCA), and implementation of preventive measures.
  • Participate in on-call support and production operations for mission-critical applications and services.

Observability & Monitoring

  • Build and manage enterprise observability solutions leveraging tools such as Splunk, Grafana, Prometheus, Datadog, New Relic, and Open Telemetry.
  • Develop dashboards, alerts, monitoring strategies, and reporting frameworks that provide actionable operational insights.
  • Improve Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR) through proactive monitoring and automation.
  • Establish best practices for logging, distributed tracing, application monitoring, and system health management.

Platform Engineering

  • Develop and enhance shared platform services and infrastructure utilized by multiple engineering teams.
  • Create self-service capabilities and automation tools that improve developer productivity and reduce operational overhead.
  • Design reusable platform components, frameworks, and operational tooling.
  • Collaborate with architecture, engineering, and product teams to define and execute platform modernization strategies.

Cloud Infrastructure & Automation

  • Design, deploy, and manage cloud infrastructure across AWS and Azure environments.
  • Implement Infrastructure as Code (IaC) using Terraform, CloudFormation, Helm, and Kubernetes manifests.
  • Automate infrastructure provisioning, configuration management, deployments, scaling, and recovery processes.
  • Improve infrastructure consistency, governance, security, and scalability through automation and standardization.

CI/CD & DevOps Enablement

  • Build and optimize CI/CD pipelines to enable secure, scalable, and reliable software delivery.
  • Implement deployment automation, release management processes, and validation controls.
  • Support GitOps methodologies and continuous delivery practices.
  • Partner with development teams to improve deployment frequency, quality, and operational stability.

Incident Management & Resilience Engineering

  • Lead operational readiness assessments, disaster recovery planning, and business continuity exercises.
  • Develop and maintain runbooks, playbooks, escalation procedures, and automated remediation workflows.
  • Drive resiliency testing, chaos engineering initiatives, and fault-tolerance improvements.
  • Ensure compliance with operational, security, and reliability standards.

AI-Driven Operations & Innovation

  • Utilize AI-powered observability and incident management platforms to enhance operational efficiency.
  • Leverage predictive analytics and automation to proactively identify risks and prevent service disruptions.
  • Drive adoption of intelligent operational capabilities that improve system reliability, engineering productivity, and customer experience.
  • Explore innovative approaches to reduce manual effort and operational toil through automation and AI-driven solutions.

Cross-Functional Collaboration

  • Partner closely with Software Engineering, Platform Engineering, Cloud Architecture, Security, and Product teams to deliver reliable and scalable solutions.
  • Provide technical leadership and guidance on reliability best practices, automation strategies, and operational excellence.
  • Mentor junior engineers and contribute to continuous improvement initiatives across the organization.

Qualifications

  • 5+ years of experience in Site Reliability Engineering (SRE), DevOps, Platform Engineering, Cloud Operations, or a related field.
  • Strong experience with Linux administration, Kubernetes, Docker, and cloud platforms such as AWS and/or Azure.
  • Hands-on experience with Infrastructure as Code (Terraform, CloudFormation, Helm) and CI/CD pipeline automation.
  • Proficiency in scripting and automation using Python and Bash.
  • Experience supporting distributed systems, APIs, microservices, and customer-facing applications in production environments.
  • Strong knowledge of monitoring and observability tools such as Splunk, Grafana, Prometheus, Open Telemetry, and New Relic.
  • Experience with incident management, root cause analysis (RCA), performance tuning, disaster recovery, and reliability engineering.
  • Understanding of Git, DevOps best practices, automation, and cloud-native technologies.
  • Strong problem-solving, troubleshooting, and collaboration skills.
  • Bachelor’s degree in computer science, Engineering, Information Technology, or a related field, or equivalent practical experience.

The base salary for this position is $115,000 - $140,000, plus incentives that align with individual and company performance. Actual salaries will vary based on work location, qualifications, skills, education, experience, and competencies. Benefits available to eligible employees in this role include medical, dental, and vision insurance, comprehensive employee assistance program, 401(k) retirement plan, paid time off and holidays and paid learning days.

The deadline to apply for this position is: 09/02/2026 This position is for an existing, immediate vacancy. We are currently seeking to fill this role with an individual who can start as soon as possible.

As part of the hiring process, candidates may be required to undergo background screening and identity verification, where permitted by applicable law and consistent with the requirements of the role.  Certain verification processes used by the Company or its service providers may involve technologies that rely on biometric identifiers or biometric information, where permitted by law.  If biometric identifiers or biometric information are collected, used, or stored, the Company will provide the legally required disclosures and obtain any required written consent prior to such collection, and will handle such information in accordance with applicable biometric privacy laws and Company policies.

# WFO

#LI-Hybrid

#Concentrix

    Location:

    USA Seattle 1401 NW 46th St

    Language Requirements:

    Time Type:

    Full time

    Physical and Mental Requirements:
     
    The employee is regularly required to operate a computer, keyboard, telephone/headset, and/or other office equipment as essential functions of this position. Work is generally sedentary in nature.

     

    Equal Employment Opportunity:

    Concentrix is an equal opportunity and affirmative action (EEO-AA) employer. We promote equal opportunity to all qualified individuals and do not discriminate in any phase of the employment process based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, pregnancy or related condition, disability, status as a protected veteran, or any other basis protected by law.


    For more information regarding your EEO rights as an applicant, please visit the following websites:

    ·    English  

    ·   Spanish

    Accommodation:

    Concentrix welcomes and encourages applications from candidates with disabilities and is committed to providing an inclusive recruitment process. If you require reasonable accommodation to participate in any stage of the application or interview process, please let us know. Requests may be made by contacting [email protected].  All information will be treated confidentially and used solely to facilitate your participation in the recruitment process.

     

    Artificial Intelligence:

    As part of our recruitment process, we may use artificial intelligence (AI) tools to assist in the screening and/or assessment of job applicants. These tools could be used to evaluate resumes, applications, and other materials submitted to help us identify the best candidates for the role.

     

    Work Authorization:


    In accordance with federal law, only applicants who are legally authorized to work in the United States will be considered for this position. Must reside in the United States or have a valid U.S. address for residence.  

     

    For further information on available work states and Equal Employment Opportunity as an applicant, please click HERE

    Skills Required

    • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Cloud Operations, or a related field
    • Strong experience with Linux administration, Kubernetes, Docker, and AWS and/or Azure
    • Hands-on experience with Infrastructure as Code using Terraform, CloudFormation, Helm, and CI/CD pipeline automation
    • Proficiency in scripting and automation using Python and Bash
    • Experience supporting distributed systems, APIs, microservices, and customer-facing applications in production
    • Strong knowledge of monitoring and observability tools including Splunk, Grafana, Prometheus, OpenTelemetry, and New Relic
    • Experience with incident management, root cause analysis, performance tuning, disaster recovery, and reliability engineering
    • Understanding of Git, DevOps best practices, automation, and cloud-native technologies
    • Strong problem-solving, troubleshooting, and collaboration skills
    • Bachelor's degree in computer science, Engineering, Information Technology, or a related field, or equivalent practical experience
    Am I A Good Fit?
    beta
    Get Personalized Job Insights.
    Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

    The Company
    HQ: Canonsburg, PA
    843 Employees
    Year Founded: 1864

    What We Do

    CNX Resources Corporation (NYSE: CNX) is the premier natural gas development, midstream, and technology company in Appalachia – one of the most energy abundant regions in the world. We believe in an Appalachia First approach to our work – prioritizing investments and utilizing home-grown resources that truly make a Tangible, Impactful, and Local difference in our regional communities first, and then far beyond. Over the past 150+ years, our 100% local workforce has produced low cost and low emission natural gas to help meet the world’s growing energy demand and catalyze environmental progress. Learn more about our company at CNX.com. Or, to view stories on our vision, innovations and actions, visit PositiveEnergyHub.com

    Similar Jobs

    CrowdStrike Logo CrowdStrike

    Senior Engineer

    Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
    Remote or Hybrid
    USA
    11000 Employees
    140K-215K Annually
    In-Office
    Alpharetta, GA, USA
    10001 Employees

    Fiserv Logo Fiserv

    Senior Site Reliability Engineer

    eCommerce • Fintech • Information Technology • Payments • Financial Services
    In-Office
    3 Locations
    41000 Employees
    128K-216K Annually
    Remote or Hybrid
    United States
    1557 Employees
    155K-172K Annually

    Similar Companies Hiring

    UL Solutions Thumbnail
    Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
    Chicago, IL
    15000 Employees
    Runwise Thumbnail
    Greentech • Hardware • Real Estate • Software • Energy • PropTech
    New York, NY
    199 Employees
    Energy CX Thumbnail
    Greentech • Professional Services • Business Intelligence • Consulting • Energy • Financial Services • Utilities
    Chicago, IL
    108 Employees

    Sign up now Access later

    Create Free Account

    Please log in or sign up to report this job.

    Create Free Account