Infrastructure Support Engineer - Contract

Posted 9 Days Ago
Be an Early Applicant
London, Greater London, England, GBR
In-Office
Senior level
Artificial Intelligence • Semiconductor • Manufacturing
The Role
Provides day-to-day production infrastructure support across AWS, GitHub, Cloudflare, Terraform, Ansible, Linux, and Kubernetes. Owns the support queue, resolves access and infrastructure requests, manages configuration changes through pull requests, creates self-service automation, maintains incident runbooks, monitors capacity and licenses, and reports support metrics. The role excludes platform design, end-user IT support, and security incident response.
Summary Generated by Built In
About OLIX

AI is growing faster than any technology in history and the explosion in demand has created a massive infrastructure gap; we can no longer build chips or power stations fast enough to keep up. The industry is still leaning on a ten-year-old hardware blueprint that has reached its limit. A new paradigm that is faster and more efficient will be the biggest economic opportunity of the next century and create the most important company of the next decade. The OLIX Decode Accelerator 1 (DX-1) is the first accelerator architected specifically for decode. Rack-scale co-design of logic, data movement, packaging, optics and interconnect enables a step change in system level performance.

The Role

We're hiring one Infrastructure Support Engineer on a six-month contract to run day-to-day support for OLIX's production infrastructure: AWS, GitHub, Cloudflare, Terraform, Ansible, Linux and Kubernetes.

You'll be the first line for every infrastructure request across engineering — unblocking people fast, then removing the cause so the request stops coming back. The measure of success is simple: requests resolved in hours rather than days, and a queue that shrinks as automation takes over the repetitive work.

Start date: ASAP
Location: London, with working overlap for Austin and India
Contract: 6 months

Responsibilities
  • Own the infrastructure support queue, responding to every request within 15 minutes and aiming for resolution in hours, not days.

  • Resolve procedural requests end to end: access, permissions, capacity, licence installs, host rebuilds, DNS records, runner and repository administration.

  • Apply changes through Terraform and Ansible in the infra repository, by pull request.

  • Build self-service tooling and automations for common request types, removing blockers to engineering work.

  • Write and maintain runbooks for incidents affecting core infrastructure.

  • Escalate to infrastructure engineering where appropriate.

  • Monitor capacity and licence expiry, and resolve proactively before limits are hit.

  • Report weekly on volume, resolution times, top causes, and work removed by automation.


What this role does not cover: designing or building new platforms, owning the EDA platform build programme, end-user IT support (laptops, phones, Microsoft 365, identity provisioning — these stay with the IT support team), or security incident response.

Skills & Experience

Required:

  • Five or more years of Linux administration on Ubuntu or RHEL: users and permissions, filesystems, storage, networking, SSH, remote desktop.

  • AWS: EC2, S3, IAM, VPC, Secrets Manager.

  • Terraform and Ansible — reading and changing existing code, working through pull requests.

  • GitHub administration: organisations, repositories, permissions, Actions runners.

  • Experience working in a ticket queue against response and resolution targets.

  • Clear written runbooks and documentation.

Useful:

  • Cloudflare Zero Trust and WARP.

  • FlexLM licence administration.

  • SLURM and HPC or EDA compute environments.

  • Semiconductor or engineering environments using Cadence or Synopsys tools.

  • Kubernetes.

Due to U.S. export control regulations, candidates' eligibility to work at OLIX depends on their most recent citizenship or permanent residency status. We are generally unable to consider applicants whose most recent citizenship or permanent residence is in certain restricted countries (currently including Iran, North Korea, Syria, Cuba, Russia, Belarus, China, Hong Kong, Macau, and Venezuela). Applicants who have subsequently obtained citizenship or permanent residency in another country not subject to these restrictions may still be eligible.

Skills Required

  • Five or more years of Linux administration on Ubuntu or RHEL, including users, permissions, filesystems, storage, networking, SSH, and remote desktop
  • Experience with AWS, including EC2, S3, IAM, VPC, and Secrets Manager
  • Experience with Terraform and Ansible, including reading and modifying existing code through pull requests
  • GitHub administration experience, including organizations, repositories, permissions, and Actions runners
  • Experience working in a ticket queue with response and resolution targets
  • Ability to write clear runbooks and documentation
  • Experience with Cloudflare Zero Trust and WARP
  • FlexLM license administration experience
  • Experience with SLURM and HPC or EDA compute environments
  • Experience in semiconductor or engineering environments using Cadence or Synopsys tools
  • Kubernetes experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Durban
82 Employees
Year Founded: 2024

What We Do

The latest generation of AI models achieve breakthrough performance by using vastly more tokens to solve complex problems. As frontier models become more sophisticated, demand is compounding faster than today’s infrastructure can scale. Even the most dominant players, with full-stack control across silicon, software, and supply chains, are unable to solve this within the existing architecture. Inherent constraints in physical design and packaging mean a GPU-based approach is incapable of simultaneously delivering high throughput and high interactivity at low cost. Continuing AI’s advance and making it available to everyone requires a new compute paradigm. One that can overcome the fundamental limits of memory, energy, and speed that define today’s systems. If you like working on difficult and consequential problems, we want you at OLIX. We have offices in London, Austin, Toronto, San Francisco and Bristol. Check out our careers page at olix.com/careers

Similar Jobs

Teya Logo Teya

Business Development Representative

Fintech • Payments • Financial Services
In-Office
2 Locations
1000 Employees
50K-50K Annually

Braze Logo Braze

Solutions Engineer

Marketing Tech • Mobile • Software
Easy Apply
Hybrid
London, Greater London, England, GBR
2000 Employees

Hilton Logo Hilton

Office Manager

Software • Hospitality
In-Office
London, Greater London, England, GBR
121228 Employees

Hilton Logo Hilton

Receptionist

Software • Hospitality
In-Office
Middlesex, England, GBR
121228 Employees
40K-40K Hourly

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account