Site Reliability Engineer - Public Sector

Posted One Month Ago
Hiring Remotely in USA
Remote
140K-170K Annually
Mid level
Artificial Intelligence • Software • Generative AI • Automation
The Role
Operate and harden Blitzy's self-hosted, Kubernetes-based AI platform inside customer-controlled secure cloud environments. Own deployments, upgrades, capacity planning, observability, incident response, and customer-facing technical coordination while championing security and feeding operational learnings back into the product roadmap.
Summary Generated by Built In

Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into enterprise grade code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.

Our Culture

Who we are:

Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.

How we work:

  • We move Blitzy Fast: Time is both our company’s and our clients’ most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.

  • Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.

  • Passion for Invention: We’re pushing the frontier of what’s possible, requiring constant innovation and iteration.

  • We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.

  • We believe in being ‘everyday athletes’: taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for optimal mental performance. It makes for a happier and more productive team.

About the Role

As a Site Reliability Engineer on Blitzy's Public Sector team, you will be the backbone of our platform's reliability and operational excellence for a dedicated enterprise customer opportunity in a highly regulated industry. You'll deploy and operate Blitzy's self-hosted platform within the customer's secure cloud environment, serving as Blitzy's embedded engineer on the account. You'll work at the intersection of software engineering, infrastructure, and customer success, ensuring our AI-powered development platform remains highly available and performant in one of the most demanding security environments in enterprise software. This is a high-impact, hands-on role for an engineer who thrives in a fast-moving environment and takes deep ownership of the systems they operate.

Responsibilities
  • Deploy, operate, and maintain Blitzy's self-hosted platform within a customer-controlled, secure cloud environment.

  • Own the Kubernetes-based deployment: releases, upgrades, capacity planning, and performance benchmarking for compute-intensive AI workloads.

  • Design and maintain observability logging, metrics, tracing, and alerting — that operates fully within the customer's security boundary.

  • Serve as Blitzy's on-account technical presence: partner with customer infrastructure, security, and governance teams on provisioning, reviews, documentation, and operational escalations.

  • Handle sensitive customer data in accordance with customer security requirements; champion security best practices across the deployment.

  • Feed lessons learned back into Blitzy's product and infrastructure roadmap to strengthen our self-hosted offering for future public sector customers.

 
Qualifications
  • 3+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles.

  • Ability to complete a customer background/badging process.

  • Strong proficiency in Kubernetes and container orchestration, including deploying software into customer-controlled, isolated, or otherwise restricted environments (defense, government, financial services, or similar regulated industries).

  • Hands-on experience deploying software into customer-controlled or restricted environments.

  • Experience operating in isolated, restricted, or otherwise highly regulated network environments (defense, government, financial services, or similar).

  • Hands-on experience with infrastructure-as-code tools (Terraform, Pulumi, or equivalent) and at least one major cloud platform.

  • Deep expertise in observability tooling, incident management, and on-call practices.

  • Strong scripting and automation skills (Python, Go, Bash, or similar).

  • Ability to effectively communicate to stakeholders — you will work directly with customer engineering, security, and governance stakeholders and represent Blitzy on the account.

  • Familiarity with handling sensitive data and security frameworks common to regulated industries.

  • Experience supporting AI/ML workloads or their supporting infrastructure.

  • Prior forward-deployed, residency, or embedded-engineer experience at an enterprise customer site.

Salary Range: $140,000 - $170,000 + Bonus + Equity

Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger. Base salary ranges are determined by country, role, level, experience, and skills. The range displayed on each job posting reflects Blitzy’s good faith determination of the minimum and maximum targets for new hire salaries across all US locations. Individual pay is determined by related factors, including job skills, experience, and relevant education or training, which may impact a final offer. Your Talent Partner can share more about the specific salary range during the hiring process.

Skills Required

  • U.S. citizenship and ability to complete customer background/badging process.
  • 3+ years experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
  • Strong proficiency in Kubernetes and container orchestration; hands-on deployment into customer-controlled or restricted environments.
  • Experience operating in isolated, restricted, or highly regulated network environments (defense, government, financial services, or similar).
  • Hands-on experience with infrastructure-as-code tools (Terraform, Pulumi, or equivalent) and at least one major cloud platform.
  • Deep expertise in observability tooling, incident management, and on-call practices.
  • Strong scripting and automation skills (Python, Go, Bash, or similar).
  • Excellent communication skills for direct customer engineering, security, and governance collaboration.
  • Experience deploying or operating software in government-accredited or similarly certified cloud environments.
  • Familiarity with handling sensitive data and security frameworks common to regulated industries.
  • Experience supporting AI/ML workloads or their supporting infrastructure.
  • Prior forward-deployed, residency, or embedded-engineer experience at an enterprise customer site.
  • Prior experience in a high-growth startup environment where you wore multiple hats.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Year Founded: 2023

What We Do

Blitzy is an autonomous software development platform that enables development teams to transform six-month software projects into six-day turnarounds. By leveraging an agentic platform with thousands of specialized AI agents and 'System 2 Thinking,' Blitzy automates over 80% of the software development lifecycle for enterprise codebases, delivering high-quality, production-ready code with precision and speed.

Similar Jobs

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Remote or Hybrid
10 Locations
5550 Employees
127K-249K Annually
In-Office or Remote
Chicago, IL, USA
1805 Employees
68K-79K Annually
In-Office or Remote
Chicago, IL, USA
1805 Employees
165K-184K Annually

Liberty Mutual Insurance Logo Liberty Mutual Insurance

Inside Sales Representative

Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Remote or Hybrid
9 Locations
40000 Employees
45K-85K Annually

Similar Companies Hiring

LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
200 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account