Sr. Cloud Infrastructure Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Arlington, VA, USA
In-Office or Remote
135K-220K Annually
Senior level
Aerospace • Software • Agriculture
The Role
Design, implement, and operate scalable AWS and Kubernetes infrastructure using Terraform and automation. Improve CI/CD, observability, security posture, and developer experience; troubleshoot production incidents, maintain runbooks, and support customer-managed deployments.
Summary Generated by Built In

We are seeking a Senior Cloud Infrastructure Engineer to join our Engineering team. This role is central to building, operating, and scaling the cloud infrastructure that powers our platform, and to making our software easy to deploy both in our cloud-hosted environments and in customer-managed environments.

This engineer will work closely with platform and product engineering to improve the reliability, automation, and operational maturity of our AWS and Kubernetes environments. Some of our environments serve regulated, security-sensitive customers; while deep security specialization is not required, this engineer should be comfortable operating within an established, security-conscious baseline.

The ideal candidate combines strong hands-on infrastructure expertise with a practical, execution-focused mindset and a bias toward automation.

Core Responsibilities

  • Design, implement, and maintain scalable, reliable infrastructure in AWS

  • Operate and improve Kubernetes-based environments, including production workloads

  • Build and maintain infrastructure as code using Terraform

  • Improve CI/CD pipelines, deployment workflows, and release automation in partnership with engineering teams

  • Build and maintain the packaging and reference architectures customers use to install our software in their own environments

  • Strengthen observability across the platform, including monitoring, logging, alerting, and actionable dashboards

  • Improve developer experience through tooling, environment automation, and self-service infrastructure

  • Operate within and preserve the established security and compliance posture of our environments

  • Monitor, troubleshoot, and resolve complex infrastructure issues with clear and timely communication

  • Participate in incident response and post-incident analysis

  • Develop and maintain documentation, runbooks, and technical standards

  • Identify opportunities to improve cost efficiency, performance, and resilience across environments

Required Qualifications

  • Minimum of 5 years of experience in DevOps, Infrastructure Engineering, Platform Engineering, or Site Reliability Engineering

  • Strong hands-on experience with AWS in production environments

  • Proven experience operating Kubernetes in production

  • Strong experience with Terraform and infrastructure-as-code practices

  • Proven experience building or improving CI/CD pipelines and deployment automation

  • Solid understanding of cloud networking, IAM, secrets management, and operational controls

  • Experience with monitoring, logging, and observability tooling

  • Scripting proficiency in Python, Go, Bash, or similar

  • Excellent troubleshooting and problem-solving skills in complex production environments

  • Strong communication skills with the ability to explain technical concepts to both technical and non-technical stakeholders

  • Must live/work in the U.S.

  •  Preferred Qualifications
  • Experience operating stateful workloads on Kubernetes, such as databases or message queues, including persistent storage and backup/recovery

  • Experience with GitOps-based deployment workflows

  • Experience packaging software for customer-managed or self-hosted deployment (e.g., Helm charts)

  • Familiarity with compliance or security frameworks such as FedRAMP, NIST, SOC 2, or similar

  • Experience with PostgreSQL, cloud storage platforms, and production networking patterns

  • Experience with configuration management tools such as Ansible

  • Experience with additional cloud platforms such as Azure or GCP

  • Experience with service mesh or advanced Kubernetes networking

  • Experience supporting customer-facing or mission-critical production infrastructure

  • Top Secret Security Clearance

Skills Required

  • Minimum of 5 years experience in DevOps, Infrastructure, Platform, or Site Reliability Engineering
  • Hands-on experience with AWS in production environments
  • Proven experience operating Kubernetes in production
  • Strong experience with Terraform and infrastructure-as-code practices
  • Proven experience building or improving CI/CD pipelines and deployment automation
  • Solid understanding of cloud networking, IAM, secrets management, and operational controls
  • Experience with monitoring, logging, and observability tooling
  • Scripting proficiency in Python, Go, Bash, or similar
  • Excellent troubleshooting and problem-solving skills in complex production environments
  • Strong communication skills to explain technical concepts to technical and non-technical stakeholders
  • Must live/work in the U.S.
  • Experience operating stateful workloads on Kubernetes (preferred)
  • Experience with GitOps-based deployment workflows (preferred)
  • Experience packaging software for customer-managed deployment (e.g., Helm charts) (preferred)
  • Familiarity with compliance/security frameworks such as FedRAMP, NIST, SOC 2 (preferred)
  • Experience with PostgreSQL, cloud storage platforms, and production networking (preferred)
  • Experience with configuration management tools such as Ansible (preferred)
  • Experience with additional cloud platforms such as Azure or GCP (preferred)
  • Experience with service mesh or advanced Kubernetes networking (preferred)
  • Experience supporting customer-facing or mission-critical production infrastructure (preferred)
  • Top Secret Security Clearance (preferred)
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Arlington, VA
48 Employees
Year Founded: 2022

What We Do

Digital twins are revolutionizing industries from aerospace to agriculture. Istari Digital makes them simple and more secure, unlocking models and simulations for better products - better everything. A faster, cheaper, greener digital future awaits. With Istari Digital, simple and secure collaboration finally comes to software-enabled physical systems, empowering your team with internet-like possibilities.

Similar Jobs

Dragos Logo Dragos

Infrastructure Engineer

Security • Cybersecurity
Remote
United States
295 Employees
165K-165K Annually
Easy Apply
Remote
US
163 Employees
150K-175K Annually
In-Office or Remote
10 Locations
843 Employees
Remote
USA
30 Employees
164K-220K Annually

Similar Companies Hiring

Outpost Space Thumbnail
Aerospace • Defense
US
24 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account