Senior Software Engineer, HPC Network

Posted 8 Days Ago
4 Locations
Mid level
Cloud • Information Technology • Machine Learning
We empower creators and innovators with access to GPU resources they need to work more efficiently.
The Role
Design, develop, and implement tooling to integrate InfiniBand fabrics with the network stack. Support large infrastructure projects and maintain high-performance network fabrics for AI, ML, and VFX workloads. Work on automation and monitoring to create highly automated networks.
Summary Generated by Built In

CoreWeave is the AI Hyperscaler™, delivering a cloud platform of cutting edge services powering the next wave of AI. The company’s technology provides enterprises and leading AI labs with the most performant, efficient and resilient solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers covering every region of the US and across Europe. CoreWeave was ranked as one of the TIME100 most influential companies of 2024.

As the leader in the industry, we thrive in an environment where adaptability and resilience are key. Our culture offers career-defining opportunities for those who excel amid change and challenge. If you’re someone who thrives in a dynamic environment, enjoys solving complex problems, and is eager to make a significant impact, CoreWeave is the place for you. Join us, and be part of a team solving some of the most exciting challenges in the industry. 

CoreWeave powers the creation and delivery of the intelligence that drives innovation. To learn more about our values, please visit our careers website.

About the Role:

Our HPC Network teams have a maniacal focus on delivering world-class network infrastructure by way of top notch automation on top of modern architectural and design concepts. Our goal is to build the most resilient, high performance network fabrics possible to accelerate our unique, and bleeding edge AI, ML, and VFX workloads for our customers, but also have fun doing it! CoreWeave maintains and runs numerous cloudscale datacenter fabrics, which are central to the direct success of our customers' workloads.  

As our next amazing Engineer, you will be responsible for helping design, develop, and implement tooling to integrate our InfiniBand Fabrics with the rest of our stack. Your day to day will consist of writing code in close cooperation with our HPC Network Engineering team to build, operate, and monitor CoreWeave’s Infiniband fabrics. Your goal is to make our network so highly automated and intelligent, that you forget it’s even there.

Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are some qualities we’ve found compatible with our team. If a portion of this resonates with you, we’d love to talk. 

4+ years experience in the following areas:

  • A basic understanding of computer networks and networking.
  • Experience supporting large infrastructure projects.
  • Familiarity with cloud native tooling and protocols such as:
    • Helm, ArgoCD, Prometheus, Grafana, Alert Manager, REST/gRPC APIs.
  • Good understanding and working knowledge of Linux.
  • Python & Shell scripting are required, Go is a big plus.
  • Using Kubernetes to automate all things is second nature to you:
    • You create controllers and operators at any time
    • Can recite "Kubernetes API Conventions" by heart
  • A great attitude, and a willingness to help those more junior, and learn from those more senior.

Nice to Have:

Previous experience with the following would be greatly appreciated:

  • CI/CD
  • Software Development
  • Monitoring
  • HPC

Our compensation reflects the cost of labor across several US geographic markets. The base pay for this position ranges from $160,000-$210,000. Pay is based on a number of factors including market location and may vary depending on job-related knowledge, skills, and experience.

What We Offer

The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.

In addition to a competitive salary, we offer a variety of benefits to support your needs, including:

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave
  • Company-paid Life Insurance 
  • Voluntary supplemental life insurance 
  • Short and long-term disability insurance 
  • Flexible Spending Account
  • Health Savings Account
  • Tuition Reimbursement 
  • Mental Wellness Benefits through Spring Health 
  • Family-Forming support provided by Carrot
  • Paid Parental Leave 
  • Flexible, full-service childcare support with Kinside
  • 401(k) with a generous employer match
  • Flexible PTO
  • Catered lunch each day in our office and data center locations
  • A casual work environment
  • A work culture focused on innovative disruption

Our Workplace

At CoreWeave, we are committed to operating as a hybrid workplace, offering employees flexibility in how they structure their time between in-office and remote work. We recognize the significance of fostering connections, collaboration, and creativity within our office culture and its positive impact on our business. Our philosophy operating as a hybrid workplace underscores our dedication to enabling employees to tailor work-life balance to their individual preferences.

For those who do not live within 30 miles of one of our offices, we are open to considering remote work for candidates whose skills and experience strongly align with the role. While we prioritize a hybrid work environment for most roles, we understand the importance of flexibility and are open to remote work for specific positions and specialized skill sets. Onboarding is essential to your success. New employees not based out of an office will be invited to attend onboarding training at one of our hubs within their first month of employment. We continue to foster a collaborative environment by bringing teams together quarterly.

 

California Consumer Privacy Act - California applicants only

CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.

As part of this commitment and consistent with the Americans with Disabilities Act (ADA), CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: [email protected].

Top Skills

Go
Python
Shell Scripting

What the Team is Saying

Alex
Andy
Sasha
Louis
Taylor
Anthony
Ivy
Darrell
Yitzy
Nicolas
Vaibhav
Robert
The Company
HQ: Roseland, NJ
806 Employees
Hybrid Workplace
Year Founded: 2017

What We Do

CoreWeave, the AI Hyperscaler™, delivers a cloud platform of cutting-edge software powering the next wave of AI. The company's technology provides enterprises and leading AI labs with cloud solutions for accelerated computing. Since 2017, CoreWeave has operated a growing footprint of data centers across the US and Europe. CoreWeave was ranked as one of the TIME100 most influential companies and featured on Forbes Cloud 100 ranking in 2024. Learn more at www.coreweave.com.

Why Work With Us

At CoreWeave we work hard, have fun and move fast! Today we are a small, growing team of intelligent, genuine people, that value different perspectives and approaches to solving complex problems. We foster an environment that champions collaboration and prioritizes innovative solutions. Here, you are surrounded by the best.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

CoreWeave Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Typical time on-site: Flexible
HQRoseland, NJ
Bellevue, WA
Brooklyn, NY
London, UK
Philadelphia, PA
Sunnyvale, CA
Learn more

Similar Jobs

CoreWeave Logo CoreWeave

Senior Site Reliability Engineer, Developer Productivity

Cloud • Information Technology • Machine Learning
4 Locations
806 Employees

CoreWeave Logo CoreWeave

Hardware Engineer, Compute Infrastructure

Cloud • Information Technology • Machine Learning
4 Locations
806 Employees

CoreWeave Logo CoreWeave

HPC Network Engineer, InfiniBand

Cloud • Information Technology • Machine Learning
4 Locations
806 Employees

CoreWeave Logo CoreWeave

Senior Infrastructure Engineer, Metal Dev (NYC)

Cloud • Information Technology • Machine Learning
New York, NY, USA
806 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account