We are seeking a Senior Cloud Infrastructure Engineer to join our Engineering team. This role is central to building, operating, and scaling the cloud infrastructure that powers our platform, and to making our software easy to deploy both in our cloud-hosted environments and in customer-managed environments.
This engineer will work closely with platform and product engineering to improve the reliability, automation, and operational maturity of our AWS and Kubernetes environments. Some of our environments serve regulated, security-sensitive customers; while deep security specialization is not required, this engineer should be comfortable operating within an established, security-conscious baseline.
The ideal candidate combines strong hands-on infrastructure expertise with a practical, execution-focused mindset and a bias toward automation.
Core Responsibilities
Design, implement, and maintain scalable, reliable infrastructure in AWS
Operate and improve Kubernetes-based environments, including production workloads
Build and maintain infrastructure as code using Terraform
Improve CI/CD pipelines, deployment workflows, and release automation in partnership with engineering teams
Build and maintain the packaging and reference architectures customers use to install our software in their own environments
Strengthen observability across the platform, including monitoring, logging, alerting, and actionable dashboards
Improve developer experience through tooling, environment automation, and self-service infrastructure
Operate within and preserve the established security and compliance posture of our environments
Monitor, troubleshoot, and resolve complex infrastructure issues with clear and timely communication
Participate in incident response and post-incident analysis
Develop and maintain documentation, runbooks, and technical standards
Identify opportunities to improve cost efficiency, performance, and resilience across environments
Required Qualifications
Minimum of 5 years of experience in DevOps, Infrastructure Engineering, Platform Engineering, or Site Reliability Engineering
Strong hands-on experience with AWS in production environments
Proven experience operating Kubernetes in production
Strong experience with Terraform and infrastructure-as-code practices
Proven experience building or improving CI/CD pipelines and deployment automation
Solid understanding of cloud networking, IAM, secrets management, and operational controls
Experience with monitoring, logging, and observability tooling
Scripting proficiency in Python, Go, Bash, or similar
Excellent troubleshooting and problem-solving skills in complex production environments
Strong communication skills with the ability to explain technical concepts to both technical and non-technical stakeholders
Must live/work in the U.S.
Preferred QualificationsExperience operating stateful workloads on Kubernetes, such as databases or message queues, including persistent storage and backup/recovery
Experience with GitOps-based deployment workflows
Experience packaging software for customer-managed or self-hosted deployment (e.g., Helm charts)
Familiarity with compliance or security frameworks such as FedRAMP, NIST, SOC 2, or similar
Experience with PostgreSQL, cloud storage platforms, and production networking patterns
Experience with configuration management tools such as Ansible
Experience with additional cloud platforms such as Azure or GCP
Experience with service mesh or advanced Kubernetes networking
Experience supporting customer-facing or mission-critical production infrastructure
Top Secret Security Clearance
Skills Required
- Minimum of 5 years experience in DevOps, Infrastructure, Platform, or Site Reliability Engineering
- Hands-on experience with AWS in production environments
- Proven experience operating Kubernetes in production
- Strong experience with Terraform and infrastructure-as-code practices
- Proven experience building or improving CI/CD pipelines and deployment automation
- Solid understanding of cloud networking, IAM, secrets management, and operational controls
- Experience with monitoring, logging, and observability tooling
- Scripting proficiency in Python, Go, Bash, or similar
- Excellent troubleshooting and problem-solving skills in complex production environments
- Strong communication skills to explain technical concepts to technical and non-technical stakeholders
- Must live/work in the U.S.
- Experience operating stateful workloads on Kubernetes (preferred)
- Experience with GitOps-based deployment workflows (preferred)
- Experience packaging software for customer-managed deployment (e.g., Helm charts) (preferred)
- Familiarity with compliance/security frameworks such as FedRAMP, NIST, SOC 2 (preferred)
- Experience with PostgreSQL, cloud storage platforms, and production networking (preferred)
- Experience with configuration management tools such as Ansible (preferred)
- Experience with additional cloud platforms such as Azure or GCP (preferred)
- Experience with service mesh or advanced Kubernetes networking (preferred)
- Experience supporting customer-facing or mission-critical production infrastructure (preferred)
- Top Secret Security Clearance (preferred)
What We Do
Digital twins are revolutionizing industries from aerospace to agriculture. Istari Digital makes them simple and more secure, unlocking models and simulations for better products - better everything. A faster, cheaper, greener digital future awaits. With Istari Digital, simple and secure collaboration finally comes to software-enabled physical systems, empowering your team with internet-like possibilities.









