Top Site Reliability Engineer Jobs

Reposted One Month AgoSaved
Hybrid
San Francisco, CA, USA
205K-225K Annually
Senior level
205K-225K Annually
Senior level
Artificial Intelligence • HR Tech • Professional Services
Design, build, and operate reliable, scalable cloud infrastructure. Maintain AWS/GCP and Linux systems, manage Kubernetes clusters, implement IaC (Ansible/Puppet/Terraform), automate CI/CD (Jenkins), monitor with Prometheus/ELK, triage alerts, participate in design/reviews, migrate apps to Kubernetes, and improve operational automation.
Top Skills: AnsibleAWSC++ElkGCPGoJenkinsKubernetesLinuxPrometheusPuppetRustTerraformTypescript
Reposted One Month AgoSaved
Hybrid
Atlanta, GA, USA
Senior level
Senior level
Healthtech • Information Technology • Software
Design, implement, and automate scalable, highly available AWS infrastructure. Improve reliability and performance, build CI/CD and IaC with Terraform, manage containers (ECS/EKS), monitor with observability tools, participate in on-call rotations, run incident response and root cause analysis, and advise engineering teams on SRE best practices.
Top Skills: Aws CodebuildAws Ec2Aws EcsAws EksAws LambdaClaudeCloudwatchDatadogDnsDockerElasticacheGitGithub CopilotIamJenkinsKubernetesPagerdutyRdsRedshiftS3SnsSqsTeamcityTerraform
Reposted One Month AgoSaved
In-Office or Remote
3 Locations
100K-125K Annually
Senior level
100K-125K Annually
Senior level
Healthtech • Pet • Biotech
Senior SRE responsible for designing and modernizing CI/CD and deployment systems, automating AWS Serverless infrastructure, improving observability and incident response, enforcing release and security practices, and guiding engineering teams to scale resilient global services.
Top Skills: AuroradbAws CloudformationAws LambdaAzure Entra IdCloudfrontDynamoDBEventbridgeGitGitGithub ActionsMavenOauth2Openid ConnectS3SnsSqsTerraform
Reposted One Month AgoSaved
In-Office
Oakland Estates, San Antonio, TX, USA
Senior level
Senior level
Digital Media • Events • Music
Lead and manage a team of SRE/DevOps engineers to ensure reliability, availability, and performance of cloud-based systems. Oversee incident response, operational troubleshooting, process improvements, and cross-team collaboration while mentoring and delegating tasks to meet business objectives.
Top Skills: Cloud Services
Reposted One Month AgoSaved
In-Office
Sunnyvale, CA, USA
170K-196K Annually
Senior level
170K-196K Annually
Senior level
Software • Cybersecurity
Senior SRE responsible for ensuring reliability, scalability, and performance of cloud systems (AWS/Azure). Duties include monitoring, on-call support, incident response and RCA, releases and maintenance, security controls, and automation to improve infrastructure efficiency.
Top Skills: AWSAzureAzure DevopsDockerGitlab Ci/CdGoJenkinsKubernetesPowershellPython
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
Reposted One Month AgoSaved
In-Office
Sunnyvale, CA, USA
170K-196K Annually
Senior level
170K-196K Annually
Senior level
Software • Cybersecurity
Drive reliability, scalability, and performance of cloud-based systems on AWS/Azure. Monitor systems, handle on-call production support, lead incident response and root cause analysis, perform releases and hotfixes, implement cloud security controls, and automate infrastructure improvements.
Top Skills: AWSAzureAzure DevopsCloud-NativeDockerGitlab Ci/CdGoJenkinsKubernetesMicroservicesPowershellPython
Reposted One Month AgoSaved
In-Office
Reston, VA, USA
133K-238K Annually
Senior level
133K-238K Annually
Senior level
Cloud • Fintech • HR Tech
The Senior Site Reliability Engineer will maintain and optimize the Kubernetes-based platform, ensuring high availability, automating infrastructure, and adhering to security compliance, while troubleshooting and collaborating with development teams.
Top Skills: Argo CdAWSC#Ci/CdGoKubernetesPythonRubyRustTerraform
Reposted 21 Days AgoSaved
Remote or Hybrid
4 Locations
165K-330K Annually
Mid level
165K-330K Annually
Mid level
Software
As a Site Reliability Engineer, you'll build and maintain infrastructure for ML models, automate processes, and collaborate cross-functionally.
Top Skills: Circle CiCloudFormationElk StackGithub ActionsGitlab CiGrafanaJenkinsKubernetesOpentelemetryPrometheusPulumiTerraform
2 Months AgoSaved
In-Office
Santa Monica, CA, USA
31-56 Hourly
Junior
31-56 Hourly
Junior
Gaming • Hardware
Entry-level Site Reliability Engineer responsible for monitoring service health, incident response, troubleshooting Kubernetes, networking, DNS, and application issues, building observability (dashboards, alerts, runbooks), automating repetitive tasks, and supporting release reliability and post-incident remediation.
Top Skills: BashCloudContainersDashboardsDnsGitHTTPKubernetesLinuxLoggingMetricsMonitoringPython
10 Months AgoSaved
In-Office
Alexandria, VA, USA
126K-228K Annually
Senior level
126K-228K Annually
Senior level
Information Technology • Software
As a Cloud Site Reliability Engineer, you'll design, implement, and maintain cloud systems across AWS and Azure, ensuring reliability and performance through automation and collaboration with various teams.
Top Skills: Arm TemplatesAWSAws CodepipelineAzureAzure DevopsAzure MonitorBashBicepCloudwatchGithub ActionsGrafanaKubernetesPowershellPrometheusPythonTerraform
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account