Senior Software Systems Engineer (Storage) - remote in the US

Posted 7 Hours Ago
Be an Early Applicant
Hiring Remotely in USA
Remote
Senior level
Software
The Role
Deploy and operate high-performance NFS storage for Kubernetes, bare-metal, GPU, and AI workloads. Integrate storage through CSI, tune Linux, NFS, networking, and kernel parameters, and manage capacity, quotas, snapshots, and lifecycle. Automate provisioning with Terraform/OpenTofu and GitOps, build observability, and troubleshoot performance and reliability issues across hybrid, edge, and air-gapped environments. The role also supports k0s, Cluster API, K0rdent, Harbor, and secure disconnected deployments.
Summary Generated by Built In
Company Description

About Mirantis

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Job Description

Overview

Deploy, integrate, and operate high-performance storage for GPU-accelerated compute and AI platforms. You will own the storage layer where Kubernetes meets bare metal — standing up NFS-based high-performance storage, wiring it into clusters via CSI, and tuning it to keep data flowing to GPU workloads at scale. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack.

About the Role

We are looking for a senior DevOps engineer who treats storage as infrastructure to be automated, observed, and tuned — not hand-managed. The right candidate is fluent in Kubernetes storage, deeply versed in Linux storage and networking fundamentals down to the kernel and NFS-client layer, and knows how to make high-performance NAS actually perform under demanding workloads. You should reach for infrastructure-as-code and GitOps by default, be self-directed in diagnosing performance and reliability issues end to end, set operational standards for others to follow, and communicate clearly across teams. Bare-metal hardware experience is a strong plus, but deep Linux storage knowledge is essential.

Responsibilities
1. Storage Integration & Operation

  • Integrate NFS-based high-performance storage (e.g., VAST, Dell PowerScale) into Kubernetes clusters via CSI, storage classes, and persistent volumes.
  • Tune the NFS data path — mount options, nconnect/RDMA, Linux client, and network settings — for high-throughput, low-latency GPU/AI workloads.
  • Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle.

2. Linux Platform & System Integration

  • Configure and optimize Linux systems for storage workloads, including driver setup, file system layout, network tuning, and kernel parameter optimization.
  • Deliver storage integration for k0s-based Kubernetes via Cluster API (CAPI) and K0rdent management/child cluster topologies.
  • Operate storage in fully disconnected (air-gapped) environments, including local artifact/mirror connectivity (Harbor) and PKI/TLS considerations.

3. Automation & Observability

  • Automate storage provisioning and configuration with infrastructure-as-code (Terraform/OpenTofu) and GitOps pipelines (ArgoCD or Flux).
  • Build monitoring, alerting, and observability for storage performance, capacity, and health.
  • Diagnose and resolve performance, reliability, and scaling issues across the storage stack.

Qualifications

Required qualifications:

  • 7+ years of experience in SRE or infrastructure operations

  • 5+ years of building/operating distributed production Storage systems at scale

  • Hands-on with High Performance Storage solutions (VAST, Weka, DDN, PowerScale)

  • Linux and Kubernetes storage fundamentals (NFS, CSI)

 

Preferred:

  • Bare-metal experience: hands-on experience with bare-metal host provisioning, raw disk/hardware layout, and physical server storage configurations.

  • Hands-on experience with VAST and/or Dell PowerScale.

  • Experience with GPUDirect Storage and RDMA/RoCE data paths

  • Experience with the Mirantis K0rdent stack (K0rdent Enterprise, K0rdent AI, k0s, MKE) and Cluster API.

  • Familiarity with other storage backends (Ceph, object/S3) and CSI driver operations.

  • Proven experience in sovereign or high-security air-gapped environments.

 

Additional Information

What does Mirantis offer you?

- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge, open-source innovation;
- Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings, happy hours, hackathons, and tech talks;
- Receive a competitive compensation package with a strong benefits plan.

It is understood that Mirantis, Inc. may use automated decision-making technology (ADMT) for specific employment-related decisions. Opting out of ADMT use is requested for decisions about evaluation and review connected with the specific employment decision for the position applied for. You also have the right to appeal any decisions made by ADMT by sending your request to [email protected]

By submitting your resume, you consent to the processing and storage of your personal data in accordance with applicable data protection laws, for the purposes of considering your application for current and future job opportunities.

We are a Leader for Container Management in G2 (#2 after AWS)!

Skills Required

  • 7+ years of experience in SRE or infrastructure operations
  • 5+ years of building and operating distributed production storage systems at scale
  • Hands-on experience with high-performance storage solutions such as VAST, Weka, DDN, or Dell PowerScale
  • Linux and Kubernetes storage fundamentals, including NFS and CSI
  • Hands-on bare-metal host provisioning, raw disk and hardware layout, and physical server storage configurations
  • Hands-on experience with VAST and/or Dell PowerScale
  • Experience with GPUDirect Storage and RDMA/RoCE data paths
  • Experience with the Mirantis K0rdent stack, k0s, MKE, and Cluster API
  • Familiarity with Ceph, object/S3 storage, and CSI driver operations
  • Experience in sovereign or high-security air-gapped environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Campbell, CA
729 Employees
Year Founded: 1999

What We Do

We are dedicated to helping organizations increase developer productivity and ship code faster on public and private clouds. We provide a ZeroOps experience to remove the stress of managing cloud native infrastructure by combining software and automation tools with our cloud native expertise to deliver the industry's leading secure cloud platforms. Our capabilities allow us to provide a secure and reliable cloud native platform that includes validated FIPS-140-2 Encryption and DISA STIG ready capabilities. Who do we serve? We serve a wide range of industries, building on our extensive customer experience to provide distinct value in specific verticals including Financial Services, Government & Education, Healthcare, Manufacturing, and Telecommunications. Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Inmarsat, PayPal, Reliance Jio, Societe Generale, Splunk, and S&P Global. Learn more at www.mirantis.com.

Similar Jobs

Golden Pet Brands Logo Golden Pet Brands

Assistant Manager, Customer Care

Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
Easy Apply
Remote
USA
178 Employees
67K-83K Annually

Cash App Logo Cash App

Designer

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
3500 Employees
252K-377K Annually

SailPoint Logo SailPoint

Consultant

Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
2461 Employees
117K-197K Annually
Remote or Hybrid
Houston, TX, USA
589 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account