Site Reliability Engineer

Posted 6 Hours Ago
Be an Early Applicant
2 Locations
Remote
160K-210K Annually
Senior level
Professional Services • Consulting
The Role
Own and improve CI/CD, developer tooling, agent harnesses, Kubernetes infrastructure, observability, cost management, and incident-response practices. Build platform capabilities that reduce engineering toil and improve delivery speed for conversational AI products. Participate in an on-call rotation for core infrastructure and help shape reliable, scalable systems across a fully remote engineering organization.
Summary Generated by Built In

About the company

Our client is a Series B conversational AI company building voice agents for customer service. Its platform has been running in production since 2020 and handles millions of calls each month for enterprise customers. A small, senior engineering team works remotely across North America to build and improve the systems behind these interactions.

The role / why it matters

As Site Reliability Engineer, you will build and operate the platform that helps engineering teams deliver reliable AI products at scale. You will own critical infrastructure domains across CI/CD, developer experience, observability, and agent harness engineering, with room to shape how the platform evolves.

What you'll do

- Own and improve CI/CD pipelines, including caching, architecture, developer self-service, and deployment workflows.

- Build developer tools and platform capabilities that reduce toil and help engineering teams ship faster.

- Extend the agent harness, including continuous integration, sandboxes, guardrails, and validation for autonomous agents.

- Operate Kubernetes-based cloud infrastructure and improve cost management, reliability, and observability.

- Develop monitoring, alerting, and incident-response practices across logs, metrics, and traces.

- Participate in an on-call rotation for core infrastructure; non-business-hours pages are rare.

What we're looking for

- Six or more years in software development enablement roles such as SRE, platform engineering, or DevEx.

- Experience owning CI/CD platforms end to end, including caching, system design, and developer self-service.

- Hands-on experience with TypeScript or Node.js, Python, Terraform, and Kubernetes; familiarity with Helm.

- Practical experience with observability, including logs, metrics, tracing, monitoring, alerting, and incident management.

- Understanding of LLMs and experience using AI development tools with sound judgment about where they help.

- Experience on fully remote teams and the ability to pass a software-engineer-oriented technical screen.

Bonus points

- Experience building platforms for autonomous AI agents, including harness engineering.

Compensation and benefits

The base salary range is $160,000 to $210,000 USD, plus competitive equity. Canadian compensation bands are lower. Benefits include medical, dental, and vision coverage, flexible vacation, a monthly wellness stipend, and a technology and learning stipend.

Location / work model

This is a fully remote role for candidates based in Canada or the United States who can work U.S. time zones. New visa sponsorship is not available; visa transfers may be considered for exceptional senior candidates.

Skills Required

  • Six or more years in software development enablement roles such as SRE, platform engineering, or developer experience
  • End-to-end ownership experience of CI/CD platforms, including caching, system design, and developer self-service
  • Hands-on experience with TypeScript or Node.js
  • Hands-on experience with Python
  • Hands-on experience with Terraform
  • Hands-on experience with Kubernetes
  • Familiarity with Helm
  • Practical experience with observability, including logs, metrics, tracing, monitoring, alerting, and incident management
  • Understanding of LLMs and experience using AI development tools with sound judgment
  • Experience working on fully remote teams
  • Ability to pass a software-engineer-oriented technical screen
  • Experience building platforms for autonomous AI agents, including harness engineering
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
28 Employees
Year Founded: 2021

What We Do

Raydar is a talent acquisition and business consulting firm that connects world-class and emerging talent with growing organizations. It supports companies through team development, strategic hiring, and customized growth solutions, helping clients recruit roles such as engineers, product managers, executives, legal counsel, and quantitative traders. Raydar focuses on understanding each organization’s needs, culture, and long-term goals to build high-impact teams.

Similar Jobs

GitLab Logo GitLab

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
3 Locations
2500 Employees
223K-380K Annually
In-Office or Remote
2 Locations
105 Employees
80K-110K Annually

MLabs Logo MLabs

Site Reliability Engineer

Artificial Intelligence • Blockchain • Information Technology • Consulting
Remote
5 Locations
90K-120K Annually

Inviso Corporation Logo Inviso Corporation

Site Reliability Engineer

Cloud • Information Technology • Business Intelligence • Consulting
Remote
CA
265 Employees

Similar Companies Hiring

Fora Thumbnail
Agency • On-Demand • Professional Services • Sales • Software • Travel • Hospitality
New York, NY
250 Employees
Energy CX Thumbnail
Greentech • Professional Services • Business Intelligence • Consulting • Energy • Financial Services • Utilities
Chicago, IL
150 Employees
Northslope Thumbnail
Artificial Intelligence • Information Technology • Software • Analytics • Consulting • Generative AI
London, GB
100 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account