IT Infrastructure Engineer

Posted 6 Days Ago
Be an Early Applicant
Prague, CZE
In-Office
Entry level
Artificial Intelligence • Hardware • Robotics • Software
The Role
Design, operate, and scale physical and computational infrastructure across networks, servers, GPU clusters, storage, backups, and lab systems. Responsibilities include network services, hardware management, Linux troubleshooting, provisioning, monitoring, capacity reporting, recovery planning, vendor coordination, and internal tooling. The role supports engineering, machine learning, content, capture, and robot-testing teams while maintaining reliable, secure, and efficient infrastructure.
Summary Generated by Built In
About REVEL
  • REVEL captures the skill of experienced workers, turns it into robot intelligence, and deploys that capability where work is hard to staff. Neural Gambit captures the force, motion and intent behind expert work. RAI learns from that data. Genesis robots perform the resulting tasks on site.

  • The company was founded in 2025 and is building its founding team in the Czech Republic. Capture sessions, model training, the print fleet and robot test rigs all depend on local network, storage and compute.

The role
  • You will design and operate the physical and computational infrastructure that the engineering, machine learning and content teams depend on: the network, the servers, the storage, and the compute those teams train and test on. You own everything from the cable in the wall to the job scheduler on the GPU nodes.

  • The work is hands-on and varied: racking and configuring hardware, designing network segmentation, diagnosing faults from firmware through to the kernel, and writing the provisioning and monitoring tooling that keeps it all visible. The systems are currently small enough for one person to administer directly, and are expected to grow quickly as capture volume and training load increase.

  • This role is paired with an IT Administrator who owns enterprise systems, identity and end-user support. You own the infrastructure layer; they own the accounts, devices and applications that sit on top of it. You will work closely together, and neither of you will be waiting on the other for the parts you own.

Responsibilities
  • Design, build and operate the network: switching, routing, VLAN segmentation, firewalls, site-to-site and remote-access VPN, and wireless coverage across office, lab and workshop.

  • Run the network services the sites depend on: DHCP, internal DNS, NTP, RADIUS and certificate-based network authentication, and the ISP and failover links behind them.

  • Manage the physical estate: servers, GPU nodes, NAS and backup storage, racks, structured cabling, PDUs, UPS units, and the compute and peripherals attached to capture and test rigs. Maintain an inventory, a firmware baseline and a replacement schedule.

  • Build the computational infrastructure on top of that: shared high-throughput storage for capture and training data, job scheduling and GPU allocation, environment provisioning, and container and image registries, so that available CPU, GPU and disk capacity is used efficiently and fairly.

  • Own backup and recovery for research data and infrastructure state, including offsite copies, restore testing and a documented recovery plan for the training and capture pipelines.

  • Write internal tooling: monitoring, alerting, inventory, capacity and usage reporting, and simple web interfaces for administration.

  • Diagnose hardware and network faults, from cabling and optics through to driver, firmware and kernel-level problems, including GPU, NVMe and high-throughput networking issues.

  • Work with vendors, integrators and support organisations on warranties, RMAs, escalations and infrastructure procurement.

  • Support other teams with their infrastructure: advise on what they build themselves, resolve the issues they report, and write documentation so common problems can be handled without you.

Requirements
  • Practical understanding of modern network infrastructure: L2/L3 switching, routing, VLANs, DHCP/DNS, firewalls, VPN and enterprise wireless.

  • Linux administration and bash. You can install, configure, automate and troubleshoot Linux systems from the command line, including storage, networking and boot-level problems.

  • Programming skills sufficient for internal tooling: scripts, small services and simple web interfaces for administration and monitoring. Python or an equivalent language.

  • Comfort working with physical hardware: racking, cabling, replacing components, and reading a vendor manual to the end.

  • Communication skills, and willingness to work on other people’s problems, including explaining causes and fixes to non-specialists.

Nice to have
  • Compute cluster administration (Slurm, Kubernetes or equivalent), including GPU scheduling and shared high-throughput storage.

  • Storage engineering: ZFS, Ceph, NFS or SMB at scale, tiering, and multi-terabyte dataset handling.

  • Configuration management and infrastructure-as-code (Ansible, Terraform, NixOS or similar).

  • Monitoring stacks such as Prometheus and Grafana.

  • Experience supporting a lab, workshop or manufacturing environment, including OT or machine networks kept separate from the office network.

  • Hybrid or cloud experience (AWS, GCP or Azure) for burst training capacity and offsite backup.

What we offer
  • Full ownership of the infrastructure layer, including architectural decisions.

  • Hardware and model development in the same building, so the results of infrastructure work are visible directly.

  • [Compensation range, equity, benefits, relocation support to be filled in.]

How to apply
  • Apply through the link on this posting. Include a short description of an infrastructure you built or took over: its scale, what you changed, and what problems you ran into.

Skills Required

  • Practical understanding of L2/L3 switching, routing, VLANs, DHCP/DNS, firewalls, VPN, and enterprise wireless
  • Linux administration and Bash, including installation, configuration, automation, and command-line troubleshooting
  • Programming skills for scripts, small services, and simple web interfaces; Python or equivalent
  • Hands-on hardware experience with racking, cabling, component replacement, and vendor documentation
  • Strong communication skills and willingness to support and explain issues to non-specialists
  • Compute cluster administration using Slurm, Kubernetes, or equivalent, including GPU scheduling and shared high-throughput storage
  • Storage engineering with ZFS, Ceph, NFS, or SMB at scale, including multi-terabyte datasets
  • Configuration management and infrastructure-as-code using Ansible, Terraform, NixOS, or similar
  • Experience with monitoring stacks such as Prometheus and Grafana
  • Experience supporting lab, workshop, manufacturing, OT, or machine networks
  • Hybrid or cloud experience with AWS, GCP, or Azure for burst training and offsite backup
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
10 Employees

What We Do

REVEL is an AI robotics company developing physical intelligence for general-purpose humanoid robots. The company captures the force, dexterity, and intent of human work through its Neural Gambit wearable to train RAI, the intelligence powering its robots. It serves as a modern robotics technology firm providing training data, infrastructure tools, and an AI data layer for the next generation of humanoid robots.

Similar Jobs

Pfizer Logo Pfizer

Director R&D EHS Program Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees
177K-294K Annually

SailPoint Logo SailPoint

Consultant

Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
2461 Employees

Mondelēz International Logo Mondelēz International

Program Manager

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
9 Locations
90000 Employees
4K-4K Annually

Pfizer Logo Pfizer

Quality Assurance Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
28 Locations
121990 Employees

Similar Companies Hiring

LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
200 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account