Senior Site Reliability Engineer, Platform Infrastructure

Posted 4 Hours Ago
Be an Early Applicant
South Jordan, UT, USA
Hybrid
Senior level
Other
The Role
Own the architecture, reliability, scalability, security, and cost management of Cricut’s AWS platform infrastructure. Lead infrastructure-as-code practices, monitoring, incident response, SLO/SLI development, and blameless postmortems. Partner with engineering teams on platform strategy, evaluate AI/ML infrastructure, mentor engineers, and participate in 24/7 SRE on-call coverage.
Summary Generated by Built In
Company Description

Cricut® empowers people to make and personalize almost anything—from custom cards and apparel to everyday items and home décor. Our smart cutting machines, design apps, and materials make creativity easy and accessible for everyone. We believe everyone is born creative, and our mission is to put the power of handmade into the hands of all. With a passionate community of Makers around the world, Cricut helps turn inspiration into real, tangible creations—one project at a time.

Let’s make.

Job Description

We're looking for a Senior Site Reliability Engineer, Platform Infrastructure to take hands-on technical ownership of the architecture, reliability, and scalability of our entire AWS infrastructure. Reporting to the Engineering Manager, Platform Infrastructure & SRE, you'll set technical direction, review designs, and raise the bar for reliability engineering across a growing and globally distributed engineering organization.

This is a senior individual contributor role. You'll work side by side with our onsite SRE team, Software Engineers, and other Lead Engineers to support seamless 24/7 reliability. It's ideal for an AI-forward engineer with a strong software engineering background who uses AI-assisted development tools to move faster, has a passion for infrastructure-as-code, and a proven track record of mentoring engineers to build highly reliable, scalable, and performant systems.

Key Responsibilities

  • Own the architecture, reliability, and scalability of critical AWS infrastructure, working hands-on across the full stack.
  • Partner with Software Engineers and other Lead Engineers to shape the roadmap and technical strategy for Cricut's platform infrastructure.
  • Take ownership of our AWS environment, driving best practices in security, cost management, and scalability.
  • Champion and expand our "infrastructure-as-code" philosophy across the organization.
  • Use AI-assisted development tools (e.g., Claude Code, GitHub Copilot) to accelerate delivery, applying prompt engineering and context management practices to get reliable results, and evaluate AI/ML infrastructure (e.g., model serving, vector databases, LLM tooling) as it becomes part of the platform.
  • Oversee production monitoring, incident response, and blameless post-mortem processes to continuously improve system reliability.
  • Develop and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for critical production systems.
  • Act as a key consultant for our feature-focused pillar and pod teams, ensuring they have the infrastructure resources and support required to deliver their projects successfully.
  • Mentor software engineers who have an affinity for infrastructure, helping them grow their skills in reliability engineering.
  • Collaborate closely with the onsite SRE team, sharing the on-call rotation to ensure seamless 24/7 reliability coverage.

Qualifications

  • 4+ years of experience in a software engineering or site reliability engineering role.
  • An AI-forward mindset: fluency with AI-assisted development tools (e.g., Claude Code, GitHub Copilot, Cursor) to accelerate delivery, practical prompt engineering skills, and experience managing context (e.g., structuring prompts, memory, and retrieved data) to keep LLM-based workflows accurate and reliable, plus hands-on exposure to AI/ML infrastructure (e.g., model serving, vector databases, LLM operations).
  • A strong background in software engineering, ideally with experience in backend microservices (.NET is a strong plus).
  • Deep, hands-on expertise with AWS and its core services (e.g., EC2, S3, RDS, Kinesis, VPC, IAM).
  • Proven experience building and managing infrastructure with Infrastructure-as-Code (IaC) tools like Terraform or CloudFormation.
  • A solid understanding of SRE principles and a proven track record of improving site reliability.
  • Experience with modern observability, monitoring, and logging platforms, ideally Datadog and OpenTelemetry.
  • A Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent industry experience.

Soft Skills

  • Communicates complex technical concepts with clarity—written and verbal—to diverse audiences across engineering, product, and leadership.
  • Mentors with intent, with a demonstrated ability to develop engineers and help them grow in their careers.
  • Influences without authority, building genuine alignment on reliability and infrastructure standards across teams that don't report to you.
  • Stays calm and decisive under production pressure, leading incident response and blameless postmortems that turn outages into durable fixes.
  • Stays genuinely curious, tracking advances in cloud infrastructure, reliability engineering, and AI tooling, and pulls the best of what's new into the team's everyday practice.

Additional Information

We’ve Got You Covered

At Cricut, we take care of our people. Enjoy competitive Medical, Dental, and Vision coverage, a 401(k) match, generous PTO, tuition reimbursement, and a yearly lifestyle stipend to support your wellness and passions. You’ll also receive exclusive employee discounts—and best of all, you’ll be surrounded by some of the most talented, creative, and curious minds out there.

A Quick Note Before You Apply…

Cricut is in an exciting chapter of transformation. We’re evolving fast—refining our strategy, growing our teams, and raising the bar across everything we do. This is an incredible opportunity for the right kind of person—but it’s not for everyone.

We’re looking for A-players—people who thrive in dynamic environments, turn challenges into momentum, and consistently deliver their best work. If that sounds like you, read on.

Here’s what makes someone a great fit for this role (and for this moment at Cricut):

  • You move with urgency. You don’t wait for perfect clarity to act—you start, learn, and adjust.
  • You set high standards. You take ownership, deliver quality, and hold yourself accountable.
  • You stay focused when things move fast. You prioritize what matters most and tune out the noise.
  • You collaborate like a pro. You elevate others, communicate clearly, and bring a low-ego, high-output energy.
  • You embrace AI as part of your toolkit. From idea exploration to data analysis and creative problem-solving, you leverage AI to accelerate innovation and amplify impact—because technology and creativity go hand-in-hand here.

One More Thing (It’s a Big One)

This role is in-office at least 4–5 days per week. We believe real collaboration, innovation, and culture are built face-to-face. If you’re energized by working alongside smart, kind, creative people—and love those hallway conversations that spark the next great idea—you’ll feel right at home.

If you’re looking for a fully remote role, this may not be the right fit. But if you’re excited by challenge, purpose, and building something better—let’s make something amazing together.

Relocation Statement:

• This position is eligible for relocation assistance.

What to Do Next: Please attach your resume, cover letter and/or include links to your portfolio or other social presence. If you want to show your super powers in other ways – include that information too. You can be sure that Cricut® is an employer who values individuality, equality and diversity, so tell us what you’re all about. If you are a Maker or a DIY enthusiast, whether you think you are a good one or not, we would love to hear about it when you send us your information.

Cricut® is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. This position is contingent on successfully completing a Criminal Background Check upon hire. Cricut participates in E-Verify.

Skills Required

  • 4+ years of software engineering or site reliability engineering experience
  • Experience with AI-assisted development tools such as Claude Code, GitHub Copilot, or Cursor
  • Practical prompt engineering and LLM context-management experience
  • Hands-on exposure to AI/ML infrastructure, including model serving, vector databases, or LLM operations
  • Strong software engineering background, preferably with backend microservices
  • Deep hands-on expertise with AWS and services including EC2, S3, RDS, Kinesis, VPC, and IAM
  • Experience managing infrastructure with Terraform or CloudFormation
  • Understanding of SRE principles and experience improving site reliability
  • Experience with observability, monitoring, and logging platforms
  • Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent industry experience
  • Experience with .NET backend microservices
  • Experience with Datadog and OpenTelemetry
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: South Jordan, UT
707 Employees
Year Founded: 1962

What We Do

At Cricut, we believe that we’re all born makers. When we built our first cutting machine, we saw the potential for a simple yet powerful tool to completely transform the way people craft, design, and DIY. Since then, we continue to innovate with new machines, platforms, materials, and tools, but that’s just what we do. Who we are is a bustling worldwide community, a means for connection, and an outlet for unbridled creativity. Join us as we place the power of handmade into the hands of ALL.

Similar Jobs

PwC Logo PwC

Salt Lake City - Digital Assurance & Transparency (DAT) - Intern - Summer 2027

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
Salt Lake City, UT, USA
370000 Employees
29K-48K Hourly

PwC Logo PwC

AWS Cloud Infrastructure and Operations Delivery Manager

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
67 Locations
370000 Employees
99K-232K Annually

PwC Logo PwC

AWS Cloud Infrastructure and Operations Delivery Senior Manager

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
67 Locations
370000 Employees
124K-280K Annually

PwC Logo PwC

Procurement-Senior Associate

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Remote or Hybrid
67 Locations
370000 Employees
151K-187K Annually

Similar Companies Hiring

T-Mobile Thumbnail
Other • Utilities
Bellevue, WA
89016 Employees
Rosendin Thumbnail
Other • Manufacturing
San Jose, CA
6219 Employees
OmniCable Thumbnail
Other
Houston, Texas
815 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account