Senior Solution Architect, HPC and AI - NVIS

Sorry, this job was removed at 10:09 a.m. (CST) on Friday, Aug 29, 2025
Be an Early Applicant
Santa Clara, CA
In-Office
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
The Role

Do you want to be part of the team that brings Artificial Intelligence (AI) emerging technology to the field? We are looking for a hardworking Solution Architect (SA) to join the NVIDIA AI Enterprise (NVAIE) SA Segment Team. The mission of the NVAIE Segment team is to guide and enable the successful adoption at scale of NVIDIA AI Enterprise Software in production.

In our Solutions Architecture team, we work with NVIDIA's pioneering hardware and software, driving the latest breakthroughs in artificial intelligence. We need people who enable customer adoption of NVIDIA technology and develop lasting relationships with our technology partners, making NVIDIA a key design choice for end-user solutions. On this team, you will support full stack deployment including architectural designs, workload orchestration and application optimization. At NVIDIA, you will be immersed in a diverse, encouraging environment where everyone is inspired to do their life's work. Come join the team and see how you can make a lasting impact on the world!

What You’ll Be Doing:

  • Primary responsibilities will include building and enabling robust AI/HPC infrastructure for customers

  • Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, training stability, real-time monitoring, logging, and alerting

  • Engage in and improve services from inception and design through deployment, operation, and optimization

  • Co-design telemetry of AI workloads to help engineering build solutions for more robust workloads at scale

  • Communicate across internal teams to support the continuous improvement of NVIDIA's offerings and software designs

What We Need to See:

  • Strong foundational expertise, from a BS, MS, or Ph.D. degree in Engineering, Mathematics, Physics, Computer Science, Data Science, or similar (or equivalent experience).

  • 8+ years of experience and knowledge of neural networks including good understanding of transformer architectures. Experience designing large scale AI workloads with SLURM and/or Kubernetes

  • Proficiency with Python / C++ / Rust or other popular software languages

  • Excellent verbal, written communication, and technical presentation skills in English

  • You are motivated to work with multiple levels and teams across organizations

  • Strong analytical and problem-solving skills

  • Strong time-management and organization skills for coordinating multiple initiatives, priorities and implementations of new technology and products into very sophisticated projects

  • You are a curious self-starter with a desire for continuous learning and sharing knowledge across the team

Ways to Stand Out from The Crowd:

  • Experience orchestrating distributed Deep Learning training with SLURM

  • Proficiency in DevOps, including hands-on experience with Ansible, Terraform or similar tools. Equivalent experience will be accepted as well.

  • 8+ years designing solutions with one or more Tier-1 Clouds (AWS, Azure, GCP or OCI) and cloud-native architectures and software

  • Technical leadership with a strong understanding of NVIDIA technologies, and success in working with customers

  • Expertise with parallel file systems (e.g. Lustre, GPFS, BeeGFS, WekaIO) and high-speed interconnects (InfiniBand, Omni Path, and Gig-E)

  • Experience with integration and deployment of software products in production enterprise environments, and microservices software architecture

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 148,000 USD - 235,750 USD for Level 4, and 176,000 USD - 276,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 29, 2025.NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar Jobs

Cash App Logo Cash App

Visual Designer

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
3500 Employees
252K-377K Annually

Cash App Logo Cash App

Creative Strategy Lead

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
3500 Employees
240K-359K Annually

Cash App Logo Cash App

Machine Learning Engineer

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
3500 Employees
277K-415K Annually

Cash App Logo Cash App

Machine Learning Engineer

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
3500 Employees
277K-415K Annually
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Santa Clara, CA
21,960 Employees
Year Founded: 1993

What We Do

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, NVIDIA is increasingly known as “the AI computing company.”

Similar Companies Hiring

Granted Thumbnail
Insurance • Healthtech • Financial Services • Artificial Intelligence
New York, New York
23 Employees
Milestone Systems Thumbnail
Software • Security • Other • Big Data Analytics • Artificial Intelligence • Analytics
Lake Oswego, OR
1500 Employees
Idler Thumbnail
Artificial Intelligence
San Francisco, California
6 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account