Senior Storage Platform Engineer

Posted 16 Days Ago
Be an Early Applicant
Santa Clara, CA, USA
Hybrid
168K-334K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
The Role
Design, deploy, operate, and automate large-scale multi-vendor storage platforms supporting EDA, software engineering, AI/ML, and HPC workloads. Own storage infrastructure lifecycle, configuration standards, IaC frameworks, CI/CD pipelines, CMDB and observability integrations, self-service provisioning, capacity planning, lifecycle management, performance tuning, and root cause analysis. Partner with engineering and research teams while championing GitOps and scalable platform practices.
Summary Generated by Built In

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables unique creativity and discovery, and powers what were once science fiction inventions, from artificial intelligence to autonomous cars. NVIDIA is looking for phenomenal people like you to help us accelerate the next wave of artificial intelligence.

Join our team at NVIDIA as a Senior Storage Platform Engineer responsible for designing, deploying, and operating the storage platforms that power NVIDIA's EDA FARM, software engineering, AI/ML teams, and engineering workflows at scale. You will own the full lifecycle of our multi-vendor storage infrastructure while simultaneously building the automation, pipelines, and integrations that turn storage into a scalable, self-service platform.You will drive Infrastructure as Code adoption across various Storage platforms, and ensure our storage estate is tightly integrated with CMDB, observability, configuration management, and self-service tooling.

What You'll Be Doing:

  • Lead end-to-end deployment of storage systems across NetApp, Pure Storage, Cloudian, and DDN, owning timelines, configuration quality, and delivery against project milestones.

  • Define and enforce configuration standards, baselines, and operational runbooks across all platforms; ensure every deployment is consistent, documented, and auditable.

  • Design and implement Infrastructure as Code frameworks (Ansible, Terraform, or equivalent) to automate provisioning, configuration management, and lifecycle operations across all storage platforms.

  • Build and maintain CI/CD pipelines for storage configuration deployments — ensuring every change is reviewed, tested, and rolled out in a repeatable, validated manner.

  • Develop and own integrations between storage platforms and the broader infrastructure ecosystem: CMDB for asset discovery and inventory sync, observability stacks for metrics and alerting, configuration management tools for drift detection, and self-service portals that let engineering teams provision storage on demand.

  • Manage day-2 operations including capacity planning, firmware and software lifecycle management, performance tuning, and root cause analysis.

  • Partner with stakeholders like chip design teams, software engineering, research and application teams to deliver integrated, fit-for-purpose storage solutions for engineering workloads

  • Champion GitOps practices for storage — every configuration tracked, every change reviewed, no manual snowflakes.

​

What we need see:

  • 8+ years of hands-on experience with enterprise storage systems in large-scale production environments.

  • Deep working knowledge of at least three platforms from our stack: NetApp ONTAP (NFS/NAS), Pure Storage FlashArray or FlashBlade, Cloudian HyperStore (S3 object), DDN (Lustre/EXAScaler or high-performance NAS).

  • Strong command of storage protocols — NFS, SMB, iSCSI, NVMe-oF, S3, and Lustre.

  • Proven experience building infrastructure automation with Ansible, Terraform, or equivalent IaC tools.

  • Proficiency in Python, Go, or similar, with a track record of building reusable tooling rather than one-off scripts.

  • Experience integrating storage systems with CMDB platforms (ServiceNow, Nautobot, Cerebro or equivalent) and observability stacks (Prometheus/Grafana, Splunk, Datadog, or equivalent).

  • Solid experience with CI/CD platforms (GitHub Actions, Jenkins, GitLab CI, or equivalent) and Git-based workflows.

  • Ability to work across teams and translate infrastructure needs into platform capabilities.

  • MS Degree in Computer Science or equivalent experience

Way to stand out from the crowd:

  • You've worked in HPC, AI/ML, or large-scale research computing environments where storage is mission-critical and throughput matters.

  • You've built self-service storage provisioning workflows — not just automated deployments, but systems that enable other teams to move independently.

  • You are experienced with containerization technologies, such as Docker, Mesosphere DCOS, Kubernetes (k8s).

  • You've contributed to or maintained internal developer tools or platform APIs — not just consumed them.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

#LI-Hybrid

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 270,250 USD for Level 4, and 208,000 USD - 333,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 14, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Skills Required

  • 8+ years of hands-on experience with enterprise storage systems in large-scale production environments
  • Deep working knowledge of at least three storage platforms, including NetApp ONTAP, Pure Storage, Cloudian HyperStore, or DDN
  • Strong command of NFS, SMB, iSCSI, NVMe-oF, S3, and Lustre storage protocols
  • Experience building infrastructure automation with Ansible, Terraform, or equivalent IaC tools
  • Proficiency in Python, Go, or similar programming languages, with experience building reusable tooling
  • Experience integrating storage systems with CMDB platforms and observability stacks
  • Experience with CI/CD platforms such as GitHub Actions, Jenkins, or GitLab CI, and Git-based workflows
  • Ability to work across teams and translate infrastructure needs into platform capabilities
  • MS degree in Computer Science or equivalent experience
  • Experience in HPC, AI/ML, or large-scale research computing environments
  • Experience building self-service storage provisioning workflows
  • Experience with Docker, Mesosphere DCOS, or Kubernetes
  • Experience contributing to or maintaining internal developer tools or platform APIs

NVIDIA Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about NVIDIA and has not been reviewed or approved by NVIDIA.

  • Equity Value & Accessibility — Equity awards and a discounted ESPP are highlighted as core parts of total compensation, enabling employees to share in the company’s success. Stock-based compensation and the two-year lookback ESPP are consistently described as especially valuable.
  • Healthcare Strength — Health coverage is portrayed as robust, with comprehensive medical, dental, and vision options alongside mental health support and on-site care resources. Employer HSA contributions and wellness perks reinforce the depth of the offering.
  • Retirement Support — Retirement programs are depicted as strong, featuring a meaningful 401(k) match with Roth options and support for Mega Backdoor Roth contributions. These elements position long-term savings as a notable advantage of the total rewards package.

NVIDIA Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Santa Clara, CA
21,960 Employees
Year Founded: 1993

What We Do

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, NVIDIA is increasingly known as “the AI computing company.”

Similar Jobs

Railway Logo Railway

Senior Platform Engineer

Cloud • Information Technology • Software • Infrastructure as a Service (IaaS)
In-Office or Remote
7 Locations
50 Employees

NVIDIA Logo NVIDIA

Software Engineer

Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
In-Office or Remote
Santa Clara, CA, USA
21960 Employees
184K-357K Annually
In-Office or Remote
7 Locations
6273 Employees

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account