Integration Reliability Engineer

Posted Yesterday
Be an Early Applicant
2 Locations
In-Office
150K-170K Annually
Mid level
Artificial Intelligence • Logistics • Robotics • Software
The Role
Own reliability of distributed systems across cloud (Kubernetes), edge, and on-site deployments. Build observability, monitoring, alerting, and incident response processes. Improve deployment workflows, diagnose infra/networking/distributed issues, and partner with engineering to prevent recurrence and scale repeatable deployments in imperfect real-world environments.
Summary Generated by Built In

We’re looking for a Integration Reliability Engineer to own the reliability of our system across cloud, edge, and real-world environments. Our platform runs across distributed infrastructure—connecting cloud services, on-site compute, and live video/data pipelines inside warehouses. This role is responsible for making systems observable, diagnosable, and repeatable as we scale across deployments. You’ll work closely with engineering and deployment teams to ensure the system performs reliably in production—not just in ideal conditions.

What You’ll Own
  • Own reliability of systems across cloud (Kubernetes), edge compute, and on-site deployments

  • Build and maintain monitoring, alerting, and observability systems

  • Define and improve incident response, severity levels, and on-call processes

  • Improve deployment and bring-up workflows across facilities

  • Diagnose issues across infrastructure, networking, and distributed systems

  • Partner with engineering to identify root causes and prevent recurring issues

  • Improve system visibility, debugging, and operational tooling

  • Help make deployments repeatable and scalable across sites

Required Qualifications
  • 3+ years of experience in SRE, infrastructure, or distributed systems

  • Strong Linux and networking fundamentals

  • Experience operating systems in production environments

  • Experience working with networking in constrained or distributed environments (e.g., VPNs, secure tunnels, on-site networking)

  • Experience with:

    • Kubernetes and containerized systems

    • Cloud platforms (GCP, AWS, or Azure)

    • Observability tools (Prometheus, Grafana, OpenTelemetry, etc.)

  • Ability to debug issues across multiple layers of the stack (infra → services → network)

  • Comfortable working in real-world, imperfect environments (not just clean cloud systems)

  • Strong ownership and ability to drive issues to resolution

Preferred Qualifications
  • Experience with multi-site or edge deployments

  • Experience with event-driven systems (Kafka or similar)

  • Familiarity with video or streaming systems (RTSP, WebRTC)

  • Experience working with hardware-integrated systems

  • Exposure to security/compliance frameworks (SOC2, ISO27001, etc.)

  • US citizen/ permanent resident

  • Located in SFBAY or NY area

Why This Role Matters

We’re scaling from a small number of deployments to many, and this role is critical to making the following happen:

  • Systems that work outside ideal environments

  • Fast, reliable diagnosis and recovery when things break

  • Repeatable deployments across real-world facilities

Equal Opportunity Statement

We’re an equal opportunity employer that values diversity and inclusion. We welcome teammates of all backgrounds and don’t discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.

Benefits

At Claryo, we offer a competitive benefits package that supports your health and well-being, including — top-tier medical, dental, and vision coverage, 401k with employer matching, equity, parental leave, and unlimited vacation.

Skills Required

  • 3+ years of experience in SRE, infrastructure, or distributed systems
  • Strong Linux fundamentals
  • Strong networking fundamentals
  • Experience operating systems in production environments
  • Experience with networking in constrained or distributed environments (VPNs, secure tunnels, on-site networking)
  • Experience with Kubernetes and containerized systems
  • Experience with cloud platforms (GCP, AWS, or Azure)
  • Experience with observability tools (Prometheus, Grafana, OpenTelemetry, etc.)
  • Ability to debug issues across multiple layers of the stack (infrastructure -> services -> network)
  • Comfortable working in real-world, imperfect environments
  • Strong ownership and ability to drive issues to resolution
  • Experience with multi-site or edge deployments
  • Experience with event-driven systems (Kafka or similar)
  • Familiarity with video or streaming systems (RTSP, WebRTC)
  • Experience working with hardware-integrated systems
  • Exposure to security/compliance frameworks (SOC2, ISO27001, etc.)
  • US citizen or permanent resident
  • Located in SFBAY or NY area
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
14 Employees
Year Founded: 2022

What We Do

Claryo is a leading provider of Spatial Generative AI that helps industrial facilities and warehouses achieve ideal operational efficiency. Through its AI-powered Virtual Facility, the company creates photorealistic, spatially accurate digital representations of facilities, enabling warehouses to upgrade their entire ecosystem, improve automation, and optimize workflows, inventory, and space usage.

Similar Jobs

ZS Logo ZS

Product Marketing Lead

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
10 Locations
15000 Employees
120K-130K Annually

ZS Logo ZS

Product Marketing Manager

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
10 Locations
15000 Employees
170K-187K Annually

Pontera Logo Pontera

Operations Specialist

Fintech • Software • Financial Services
Hybrid
New York, NY, USA
250 Employees
100K-120K Annually

RigUp Logo RigUp

Lead, Accounts Receivable Operations

Information Technology • Professional Services • Software • Energy
Remote or Hybrid
USA
260 Employees

Similar Companies Hiring

Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
300 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account