Principal Operations Engineer, Network

Posted 2 Days Ago
Be an Early Applicant
Hiring Remotely in Trivellini, Brescia, ITA
In-Office or Remote
258K-300K Annually
Entry level
Artificial Intelligence • Software
The Role
Serves as the senior technical authority for operational networks across hyperscale AI data centers. Leads site assessments, audits, network readiness, high-risk production changes, MOPs, and fleet-wide root cause analyses. Reviews network designs from an operational perspective, manages OEM and service-provider standards, and coordinates across network engineering, compute operations, facilities, supply chain, and customer teams. Requires extensive mission-critical network operations experience, IP and optical networking expertise, routing protocol knowledge, and 50–75% travel.
Summary Generated by Built In
About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.


We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate
  • Be a barrel. Full autonomy. Own things end to end, take on scope without being asked, no permission required to operate outside your core role.

  • Insane urgency. We drive everything forward as fast as possible.

  • Reason from first principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

  • Build something that actually matters. If you're going to spend your time, spend it on something that matters to the world.

The Data Center Operations Team

Examples of key problems the team is working on

  • Operate at the scale of a nation, not a building. Accelerating toward 100 GW by the end of the decade, roughly the entire electricity consumption of Japan. You won't just run a data center; you'll run infrastructure the size of a G7.

  • Fly the plane while it's being built. Running flawless operations inside a live construction zone, adapting in real time and turning the pace of build-out into our advantage. You will redefine operational excellence.

  • Write the playbook, don't inherit it. Most operators step into someone else's system. Here you build one, leaving your fingerprints on how the whole company runs. As we scale 100x, you'll set the standards, shape the operating model, and grow the team by the thousands.

     
Role Scope
  • Serve as the most senior technical authority for the operational network fleet across the hyperscale AI data center portfolio: switches, routers, cabling (copper and fiber), optics.

  • Lead site assessments and operational audits, and drive the technical readiness of the network ops team ahead of site activation.

  • Review network platforms and integration designs from an operational lens, and feed learnings back into network engineering, deployment, and supply chain as the build model becomes productionized.

  • Author, approve, and execute high-risk network MOPs and change records in live production, and lead fleet-wide root cause analysis on disruptions to closure.

  • Hold OEMs, ODMs, and service vendors accountable to a standard, including flawed RMA and integration processes, without burning the relationship.

  • Act as the connective tissue and force multiplier across network operations, network engineering, compute operations, facilities, supply chain, and customer-facing teams. Travel 50-75%.

     
What We're Looking For

The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would.

  • You've spent a career operating mission-critical network topologies at scale, with significant time as the senior technical voice on a site, campus, or fleet.

  • You have experience in IP network operations, as well as optical networking.

  • You have experience with TCP/IP and network routing protocols (OSPF, IS-IS, BGP, MPLS) and the devices of the physical infrastructure.

  • You've authored and executed high-risk MOPs and led root cause analysis on significant network events through closure.

  • You hold OEMs, ODMs, and deployment partners to a standard, and you enforce it without breaking the relationship.

  • You write clearly (health assessments, RCAs, design feedback) and you teach, lifting the technical readiness of every team around you.

  • Bonus: Hyperscale or large HPC fleets supporting thousands of end points. Linux and hardware management tooling. Standing up new sites from handover to steady state. Scripting for fleet-scale operations.

     
Salary & Benefits
  • Competitive total compensation package (salary + equity).

  • Retirement or pension plan, in line with local norms.

  • Health, dental, and vision insurance.

    • Generous PTO policy, in line with local norms.

      Total compensation may also include equity in the form of restricted stock units.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email [email protected] with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Skills Required

  • Career experience operating mission-critical network topologies at scale, including significant senior technical responsibility for a site, campus, or fleet
  • Experience in IP network operations and optical networking
  • Experience with TCP/IP and routing protocols including OSPF, IS-IS, BGP, and MPLS
  • Experience with physical network infrastructure devices
  • Experience authoring and executing high-risk MOPs
  • Experience leading root cause analysis for significant network events through closure
  • Experience holding OEMs, ODMs, and deployment partners to technical standards
  • Strong technical writing skills for health assessments, RCAs, and design feedback
  • Ability to teach and improve the technical readiness of surrounding teams
  • Experience with hyperscale or large HPC fleets supporting thousands of endpoints
  • Experience with Linux and hardware management tooling
  • Experience standing up new sites from handover to steady state
  • Scripting experience for fleet-scale operations
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: London
30 Employees
Year Founded: 2017

What We Do

Instantly reserve dedicated clusters of NVIDIA H200s and GB200s for any scale to supercharge your training and inference workflows.

Similar Jobs

Pfizer Logo Pfizer

Director R&D EHS Program Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees
177K-294K Annually

Capco Logo Capco

Artificial Intelligence Engineer

Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Remote or Hybrid
Italy
6000 Employees
65K-65K Annually

CrowdStrike Logo CrowdStrike

Regional Sales Manager

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
Italy
11000 Employees

Pfizer Logo Pfizer

Quality Assurance Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
28 Locations
121990 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account