Senior Specialist Field Engineer - Compute Infrastructure

Posted 2 Hours Ago
Be an Early Applicant
6 Locations
In-Office
188K-275K Annually
Senior level
Cloud • Information Technology • Machine Learning
We empower creators and innovators with access to GPU resources they need to work more efficiently.
The Role
Lead end-to-end technical delivery of large-scale bare-metal GPU clusters for strategic customers: facility/rack design, GPU cluster bring-up, InfiniBand/RoCE fabric validation, HPC benchmarking and remediation, operational models for BMaaS, and cross-team product feedback. Act as primary technical customer contact, run proofs-of-concept, collaborate with engineering teams, and support security-sensitive, production-ready supercomputers.
Summary Generated by Built In
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.

What You'll Do:

The Field Engineering organization at CoreWeave is dedicated to ensuring every customer running AI workloads at scale has a seamless, reliable, and high-performance experience. This team supports the infrastructure that powers the AI revolution—working across data centers, hardware systems, and customer workloads to maintain the integrity of our cloud platform. Field Engineering aligns closely with internal and customer engineering teams, offering valuable insights from the field and the chance to shape the CoreWeave product roadmap and development.

About the role:

As a Specialist Field Engineer - Compute Infrastructure at CoreWeave, you'll own the technical path for some of our largest customers as they go from facility and rack design to a validated, production-ready supercomputer. Working alongside the teams that build and operate each layer, you are the deep technical expert who turns raw data center hardware—racks, GPUs, high-speed fabric, firmware—into reliable compute that customers can train and inference on at scale, spanning infrastructure engineering, provisioning, validation, operations, and support.

You'll engage hands-on across the entire customer lifecycle: leading new GPU cluster bring-up and acceptance, driving InfiniBand/RoCE fabric validation and HPC performance benchmarking, defining how we operate customer bare-metal fleets at rack-level-and-up (IT service, break-fix, network, and firmware), and standing up locked-down, security-sensitive environments for our most strategic AI customers. You'll partner closely with Data Center Operations, Fleet Operations, Networking, and Product Engineering, and your work in the field will directly shape how CoreWeave delivers compute infrastructure. If you're driven by innovation, thrilled by the possibilities of what specialized compute can enable, and eager to be part of a team that's shaping the future, then CoreWeave is the place for you. Join us and let's embark on this adventure together!

In this role, you will:

  • Serve as the primary technical point of contact for customers, establishing strong technical relationships and ensuring their success with CoreWeave's cloud infrastructure offerings, focusing on bare-metal compute infrastructure and end-to-end cluster delivery within high-performance compute (HPC) environments.
  • Own the technical path from facility and rack design to a validated, production-ready supercomputer—spanning logical design, infrastructure engineering, provisioning, validation, operations, and support. 
  • Lead bring-up and acceptance of new large-scale GPU clusters, driving InfiniBand/RoCE fabric validation, HPC performance benchmarking (e.g., NCCL, ib_write_bw), and remediation of fabric, optics, firmware, and node-level issues to meet customer performance targets.
  • Define and operationalize models for managing customer bare-metal fleets at rack-level-and-up—IT service, break-fix, network and firmware management—including Bare Metal as a Service (BMaaS) and customer self-service patterns.
  • Partner with Data Center Operations, Fleet Operations, and Networking teams to align facility, hardware, and fabric readiness with customer go-live timelines and operational SLAs.
  • Review and advise on customer-facing technical contract terms, including service scope, operational responsibilities, SLAs, isolation requirements, and support boundaries.
  • Lead proof of concept initiatives to showcase the value and viability of CoreWeave's solutions within specific environments.
  • Drive technical leadership and direction during customer meetings, presentations, and workshops, addressing any technical queries or concerns that arise.
  • Act as a virtual member of CoreWeave's Compute Infrastructure, Fleet Operations, and Networking engineering teams, identifying opportunities for product enhancement and collaborating with engineers to implement your suggestions.
  • Offer valuable insights on product features, functionality, and performance, contributing regularly to discussions about product strategy and architecture. 
  • Stay informed of the latest developments and trends in Kubernetes, cloud computing and infrastructure, sharing your thought leadership with customers and internal stakeholders.
  • Lead the prototyping and initiation of research and development efforts for emerging products and solutions, delivering prototypes and key insights for internal consumption.
  • Represent CoreWeave at conferences and industry events, with occasional travel as required.

Who You Are:

  • B.S. in Computer Science or a related technical discipline, or equivalent experience
  • 7+ years of proven experience as a Solutions Architect, Field Engineer, Infrastructure/Systems Engineer, or Technical Account Manager in Cloud Infrastructure, focusing on building or operating distributed systems or HPC/cloud services, with an expertise focused on bare-metal compute infrastructure and large-scale GPU cluster delivery
  • Fluency in cloud computing concepts, architecture, and technologies with hands-on experience in designing and implementing cloud solutions
  • Proven track record with building customer relationships, communicating clearly and the ability to break down complex technical concepts to both technical and non-technical audiences
  • Deep expertise with modern rack-scale GPU server hardware (e.g., NVIDIA HGX / GB200-class systems), high-speed interconnects (InfiniBand, NVLink), and the firmware/BMC/BIOS layer
  • Expert-level Linux system administration and command-line troubleshooting, paired with strong networking fundamentals (routing, fabric topologies, TCP/IP)
  • Hands-on experience bringing up, validating, and operating large GPU clusters—including bare metal node pxe boot, hardware health, fabric validation, and HPC acceptance/performance testing—and integrating bare metal with orchestration layers such as Kubernetes and Slurm

Preferred:

  • Experience operating security-sensitive, air-gapped, or otherwise locked-down customer environments
  • Experience with scripting and automation related to bare-metal provisioning, infrastructure validation, and lifecycle management (Python, Bash, Ansible, or similar)
  • Experience designing AI supercomputers from MEP designs
  • Experience delivering bare-metal infrastructure at scale for large strategic customers or AI research labs
  • Experience with building solutions across multi-cloud or hybrid environment

Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk. 

  • You love to help solve challenging technical problems
  • You’re curious about the latest and greatest technologies in the AI space
  • You’re an expert in managing conflict and achieving mutually beneficial technical outcomes

Why CoreWeave?

At CoreWeave, we work hard, have fun, and move fast!  We’re in an exciting stage of hyper-growth that you will not want to miss out on. We’re not afraid of a little chaos, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values: 

  • Be Curious at Your Core
  • Act Like an Owner
  • Empower Employees
  • Deliver Best-in-Class Client Experiences
  • Achieve More Together

We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us! 

The base salary range for this role is $188,000 to $275,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).


What We Offer

The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.

In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings; for roles in other locations, benefits vary and are shared during the hiring process. These include:

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave
  • Company-paid Life Insurance 
  • Voluntary supplemental life insurance 
  • Short and long-term disability insurance 
  • Flexible Spending Account
  • Health Savings Account
  • Tuition Reimbursement 
  • Ability to Participate in Employee Stock Purchase Program (ESPP)
  • Mental Wellness Benefits through Spring Health 
  • Family-Forming support provided by Carrot
  • Paid Parental Leave 
  • Flexible, full-service childcare support with Kinside
  • 401(k) with a generous employer match
  • Flexible PTO
  • Catered lunch each day in our office and data center locations
  • A casual work environment
  • A work culture focused on innovative disruption

California Applicants

California Consumer Privacy Act 

Equal Opportunity & Accommodations

CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.

As part of this commitment and consistent with the Americans with Disabilities Act (ADA), CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: [email protected].

Export Control Compliance

This position requires access to export controlled information.  To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency.  CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.

Skills Required

  • B.S. in Computer Science or related technical discipline, or equivalent experience
  • 7+ years experience as a Solutions Architect, Field Engineer, Infrastructure/Systems Engineer, or Technical Account Manager in cloud infrastructure or HPC, focused on bare-metal compute and large-scale GPU cluster delivery
  • Hands-on experience designing and implementing cloud computing solutions and architecture
  • Deep expertise with modern rack-scale GPU server hardware (e.g., NVIDIA HGX / GB200-class systems) and NVLink
  • Experience with high-speed interconnects and fabric validation (InfiniBand, RoCE) and HPC performance benchmarking (NCCL, ib_write_bw)
  • Expert-level Linux system administration and command-line troubleshooting; strong networking fundamentals (routing, fabric topologies, TCP/IP)
  • Hands-on experience bringing up, validating, and operating large GPU clusters including PXE boot, hardware health, fabric validation, HPC acceptance/performance testing
  • Experience integrating bare-metal infrastructure with orchestration layers such as Kubernetes and Slurm
  • Deep familiarity with firmware/BMC/BIOS layer and node-level remediation/troubleshooting
  • Must meet export control access requirements (U.S. person or eligible to access export-controlled information or obtain required authorization)
  • Experience operating security-sensitive, air-gapped, or locked-down customer environments
  • Scripting and automation for bare-metal provisioning and lifecycle management (Python, Bash, Ansible, or similar)
  • Experience designing AI supercomputers from MEP designs
  • Experience delivering bare-metal infrastructure at scale for large strategic customers or AI research labs
  • Experience building solutions across multi-cloud or hybrid environments

What the Team is Saying

Alex
Andy
Sasha
Louis
Taylor
Anthony
Ivy
Darrell
Yitzy
Nicolas
Vaibhav
Robert

CoreWeave Compensation & Benefits Highlights

  • Healthcare Strength Health coverage is described as comprehensive, including medical, dental, vision, and mental-health resources. Feedback suggests employer-paid employee premiums and inclusive provisions (such as gender-affirming care) make coverage especially attractive.
  • Retirement Support Retirement support includes a 401(k) with company matching. Feedback suggests this is a dependable component of the package alongside other financial benefits.
  • Flexible Benefits Flexible PTO and hybrid/remote options are offered. Feedback suggests flexibility is supported by employer-verified language even as on-site hubs remain important.

CoreWeave Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Livingston, NJ
1,450 Employees
Year Founded: 2017

What We Do

CoreWeave, the AI Hyperscaler™, delivers a cloud platform of cutting-edge software powering the next wave of AI. The company's technology provides enterprises and leading AI labs with cloud solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers across the US and Europe. CoreWeave was ranked as one of the TIME100 most influential companies and featured on Forbes Cloud 100 ranking in 2024. Learn more at www.coreweave.com.

Why Work With Us

At CoreWeave we work hard, have fun and move fast! Today we are a small, growing team of intelligent, genuine people, that value different perspectives and approaches to solving complex problems. We foster an environment that champions collaboration and prioritizes innovative solutions. Here, you are surrounded by the best.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

CoreWeave Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Typical time on-site: Flexible
HQLivingston, NJ
Bellevue, WA
London, UK
New York, NY
Philadelphia, PA
Sunnyvale, CA
Learn more

Similar Jobs

CoreWeave Logo CoreWeave

Senior Manager, International Reporting

Cloud • Information Technology • Machine Learning
In-Office
Dallas, TX, USA
1450 Employees
135K-198K Annually

CoreWeave Logo CoreWeave

Accounting Manager

Cloud • Information Technology • Machine Learning
In-Office
Dallas, TX, USA
1450 Employees
115K-168K Annually

CoreWeave Logo CoreWeave

Web Engineer

Cloud • Information Technology • Machine Learning
In-Office
6 Locations
1450 Employees
115K-168K Annually

CoreWeave Logo CoreWeave

User Experience Designer

Cloud • Information Technology • Machine Learning
In-Office
6 Locations
1450 Employees
135K-198K Annually

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account