Site Reliability Engineer

Posted 7 Hours Ago
Be an Early Applicant
Atlanta, GA, USA
In-Office
Mid level
Digital Media • Real Estate • Software
The Role
Own AWS production infrastructure reliability, scalability, performance, and cost optimization. Build automated CI/CD pipelines, infrastructure as code, observability, and reliability practices. Lead incident response, on-call operations, postmortems, disaster recovery, security hardening, and compliance efforts while partnering with development teams to improve system design and releases.
Summary Generated by Built In

About SeekNow 

Seek Now is transforming property inspections through technology, data, and human expertise. We deliver faster, smarter, more reliable insights to insurance carriers and single-family rental markets, and we’re just getting started. If you want to be part of a product-driven, tech-forward team building real-world impact at scale, you’re in the right place. 

The Role 

As a Site Reliability Engineer, you'll be responsible for the availability, scalability, and operational health of our AWS-hosted infrastructure. You'll lead incident response, build the automation that lets our engineering teams ship safely and often, and drive a culture of measurable reliability across the organization. 

What You'll Do 

  • Own the reliability and performance of production services running in AWS, including capacity planning, cost optimization, and architecture reviews 
  • Design, build, and maintain fully automated CI/CD pipelines that take code from commit to production with minimal manual intervention 
  • Lead incident management: serve in the on-call rotation, coordinate response during outages, run blameless postmortems, and drive remediation to completion 
  • Define and track SLOs, SLIs, and error budgets in partnership with product and engineering teams 
  • Build and improve observability through monitoring, logging, alerting, and distributed tracing 
  • Manage infrastructure as code and eliminate toil through automation 
  • Partner with development teams to embed reliability best practices into system design and release processes 
  • Contribute to disaster recovery planning, security hardening, and compliance efforts 

What We're Looking For 

  • 4+ years in SRE, DevOps, or cloud operations roles supporting production systems 
  • Deep hands-on experience operating workloads in AWS (e.g., EC2, ECS/EKS, Lambda, RDS, S3, IAM, VPC, CloudWatch) 
  • Proven experience with incident management: on-call ownership, incident command, root cause analysis, and postmortem processes 
  • Demonstrated track record of building fully automated CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, CodePipeline, or similar) 
  • Strong infrastructure-as-code skills with Terraform, CloudFormation, or CDK 
  • Proficiency in at least one scripting or programming language (Python, Go, Bash) 
  • Experience with containers and orchestration (Docker, Kubernetes) 
  • Familiarity with observability tooling such as Datadog, Prometheus/Grafana, or the ELK stack 
  • Clear communicator who stays calm under pressure and can explain complex issues to technical and non-technical audiences 

Nice to Have 

  • AWS certifications (Solutions Architect, DevOps Engineer) 
  • Experience with GitOps workflows (ArgoCD, Flux) 
  • Background in security operations, compliance frameworks (SOC 2, ISO 27001) 

Why You'll Love It Here 

  • Tech-First Culture: We believe in building smart, scalable systems—and we invest in them. 
  • Real-World Impact: Your work will touch thousands of users every day, improving workflows and outcomes. 
  • Autonomy + Collaboration: Own your space while being part of a highly connected, supportive team. 
  • Growth-Minded Environment: We prioritize learning, innovation, and pushing the limits of what’s possible. 

What We Offer 

  • Competitive salary 
  • Comprehensive health, dental, and vision coverage 
  • 401(k) with company match 
  • Flexible PTO and hybrid work arrangement in Atlanta 

Location 

This role is based in Atlanta, GA, with a hybrid schedule.

Skills Required

  • 4+ years of experience in SRE, DevOps, or cloud operations roles supporting production systems
  • Hands-on experience operating production workloads in AWS, including services such as EC2, ECS/EKS, Lambda, RDS, S3, IAM, VPC, and CloudWatch
  • Experience with incident management, on-call ownership, incident command, root cause analysis, and postmortems
  • Experience building fully automated CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, CodePipeline, or similar
  • Infrastructure-as-code skills with Terraform, CloudFormation, or CDK
  • Proficiency in at least one scripting or programming language: Python, Go, or Bash
  • Experience with containers and orchestration, including Docker and Kubernetes
  • Familiarity with observability tools such as Datadog, Prometheus/Grafana, or the ELK stack
  • Clear communication skills and ability to explain complex issues to technical and non-technical audiences
  • AWS certification such as Solutions Architect or DevOps Engineer
  • Experience with GitOps workflows using ArgoCD or Flux
  • Background in security operations and compliance frameworks such as SOC 2 or ISO 27001
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Louisville, KY
322 Employees
Year Founded: 2012

What We Do

Seek Now is an inspection platform and services provider to the Property & Casualty (P&C) Insurance industry.

Similar Jobs

PagerDuty Logo PagerDuty

Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Easy Apply
Hybrid
2 Locations
1200 Employees
98K-149K Annually

PagerDuty Logo PagerDuty

Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Easy Apply
Hybrid
Atlanta, GA, USA
1200 Employees
113K-172K Annually

PwC Logo PwC

Site Reliability Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
9 Locations
370000 Employees
124K-280K Annually

PwC Logo PwC

Site Reliability Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
9 Locations
370000 Employees
99K-232K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account