HPC / Linux System Administrator

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Saline, MI, USA
In-Office or Remote
Senior level
Information Technology • Consulting
The Role
Administer and optimize production HPC environments supporting CAE and engineering workloads. Responsibilities include managing Linux infrastructure, job schedulers, resource allocation, CAE applications, software licensing, provisioning, automation, networking, monitoring, capacity planning, troubleshooting, incident response, and root cause analysis. The role also supports hardware lifecycle activities, system health checks, documentation, reporting, and continuous improvements to cluster performance, reliability, and scalability.
Summary Generated by Built In
Description

Our client is an American company with more than 50 years of experience, dedicated to providing autonomy to the first line of retail/e-commerce, manufacturing, transportation and logistics, healthcare, the public sector, and other industries to achieve a competitive advantage.

Our client has more than 10,000 partners and 8,000 employees in 100 countries, offering better visibility through integrated and customized solutions for each industry, which allow connecting people, assets, and data with intelligence that help make critical decisions for the business.

The Challenge (Responsibilities)

  • Administer, configure, maintain, and optimize HPC job scheduling environments, including IBM Spectrum LSF, Slurm, OpenPBS/PBS, or comparable schedulers.
  • Design and tune job queues, resource allocation policies, and scheduling strategies for diverse CAE workloads.
  • Monitor HPC utilization, workload patterns, queue performance, and infrastructure capacity.
  • Optimize resource allocation to improve cluster efficiency, utilization, and workload throughput.
  • Troubleshoot job scheduling and execution issues affecting engineering workloads.
  • Install, configure, upgrade, test, and support CAE and simulation applications in production HPC environments.
  • Support integration between CAE applications and HPC job scheduling platforms.
  • Manage CAE software licensing infrastructure, including FlexLM, RLM, or comparable license-management technologies.
  • Monitor license availability and troubleshoot licensing-related issues.
  • Diagnose and resolve CAE application issues while minimizing disruption to engineering teams.
  • Administer and maintain Red Hat Enterprise Linux (RHEL) environments across HPC infrastructure.
  • Perform Linux OS provisioning, configuration, deployment, patching, and lifecycle management.
  • Utilize automated provisioning and configuration-management technologies such as PXE and comparable automation tools.
  • Develop and maintain scripts using Bash, KornShell, C Shell, Perl, Awk, or equivalent technologies to automate administration, monitoring, and health checks.
  • Maintain system logs, monitoring processes, operational documentation, and standard operating procedures.
  • Support enterprise HPC server, storage, and network infrastructure.
  • Troubleshoot hardware and infrastructure issues affecting cluster availability or performance.
  • Support high-performance networking technologies such as InfiniBand and high-speed Ethernet.
  • Participate in hardware installation, maintenance, upgrades, and lifecycle activities.
  • Perform capacity planning based on system utilization, workload trends, and anticipated engineering demand.
  • Perform regular system health checks across HPC and CAE environments.
  • Monitor infrastructure performance, availability, utilization, and workload execution.
  • Track system incidents and outages and perform structured root cause analysis.
  • Develop and implement preventive and corrective actions for recurring issues.
  • Follow established change-management processes for infrastructure updates and deployments.
  • Provide reporting on system utilization, incidents, capacity, and infrastructure performance.
  • Identify opportunities to improve HPC performance, reliability, scalability, and operational efficiency.

Your Profile (Requirements)

  • Bachelor's degree in Mechanical Engineering, Electrical Engineering, Computer Engineering, Computer Science, or a related field, and/or equivalent professional experience.
  • 7+ years of Linux system administration experience, preferably in Red Hat Enterprise Linux (RHEL) environments.
  • Hands-on experience administering High-Performance Computing (HPC) clusters.
  • Strong experience with HPC job schedulers such as IBM Spectrum LSF, Slurm, PBS/OpenPBS, or comparable enterprise HPC schedulers.
  • Experience supporting and integrating CAE applications within HPC environments.
  • Strong Linux scripting and automation skills using Bash, Shell, Perl, or comparable technologies.
  • Experience with Linux OS deployment, provisioning, patching, and system automation.
  • Solid understanding of enterprise server hardware, storage systems, and networking.
  • Strong troubleshooting and root cause analysis capabilities.
  • Experience supporting production-critical infrastructure environments.
  • High-Performance Mindset: Resilience, emotional intelligence, and a focus on agile delivery.
  • Technologist DNA: A deep understanding of the difference between "coding" and "engineering."

Desired

  • Hands-on experience with CAE applications such as Ansys, LS-DYNA, Nastran, or comparable engineering simulation platforms.
  • Experience managing CAE licensing platforms such as FlexLM or RLM.
  • Experience with InfiniBand or other high-performance networking technologies.
  • Experience with PXE-based provisioning and automated Linux deployments.
  • Experience developing internal administration tools or dashboards.
  • Familiarity with PHP or web-based internal tooling.
  • Experience supporting HPC environments used for automotive, manufacturing, simulation, or engineering workloads.
  • Experience with infrastructure capacity planning and performance optimization.
  • Familiarity with cloud-native foundations or AI coding assistants.

Languages

  • Advanced Oral English: For seamless collaboration with global teams.
  • Advanced Spanish.

Work Arrangement

We value flexibility to support your lifestyle. This position is available as:

  • Hybrid

If you meet these qualifications and are pursuing new challenges, start your application on our website to join an award-winning employer. Explore all our job openings | Sequoia Career’s Page: https://www.sequoia-connect.com/careers/

Requirements
  • Bachelor's degree in Mechanical Engineering, Electrical Engineering, Computer Engineering, Computer Science, or a related field, and/or equivalent professional experience.
  • 7+ years of Linux system administration experience, preferably in Red Hat Enterprise Linux (RHEL) environments.
  • Hands-on experience administering High-Performance Computing (HPC) clusters.
  • Strong experience with HPC job schedulers such as IBM Spectrum LSF, Slurm, PBS/OpenPBS, or comparable enterprise HPC schedulers.
  • Experience supporting and integrating CAE applications within HPC environments.
  • Strong Linux scripting and automation skills using Bash, Shell, Perl, or comparable technologies.
  • Experience with Linux OS deployment, provisioning, patching, and system automation.
  • Solid understanding of enterprise server hardware, storage systems, and networking.

Skills Required

  • Bachelor's degree in Mechanical Engineering, Electrical Engineering, Computer Engineering, Computer Science, or a related field, or equivalent professional experience
  • 7+ years of Linux system administration experience, preferably in Red Hat Enterprise Linux environments
  • Hands-on experience administering High-Performance Computing clusters
  • Strong experience with HPC job schedulers such as IBM Spectrum LSF, Slurm, PBS/OpenPBS, or comparable schedulers
  • Experience supporting and integrating CAE applications within HPC environments
  • Strong Linux scripting and automation skills using Bash, Shell, Perl, or comparable technologies
  • Experience with Linux OS deployment, provisioning, patching, and system automation
  • Solid understanding of enterprise server hardware, storage systems, and networking
  • Strong troubleshooting and root cause analysis capabilities
  • Experience supporting production-critical infrastructure environments
  • Advanced oral English
  • Advanced Spanish
  • Hands-on experience with CAE applications such as Ansys, LS-DYNA, Nastran, or comparable engineering simulation platforms
  • Experience managing CAE licensing platforms such as FlexLM or RLM
  • Experience with InfiniBand or other high-performance networking technologies
  • Experience with PXE-based provisioning and automated Linux deployments
  • Experience developing internal administration tools or dashboards
  • Familiarity with PHP or web-based internal tooling
  • Experience supporting HPC environments used for automotive, manufacturing, simulation, or engineering workloads
  • Experience with infrastructure capacity planning and performance optimization
  • Familiarity with cloud-native foundations or AI coding assistants
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Austin, TX
30 Employees
Year Founded: 2017

What We Do

From Technologist to Technologist. We are the catalysts of digital evolution. Our core expertise lies in connecting Top Technologists with Top Companies through unparalleled IT headhunting solutions. Our international expertise helps businesses of all sizes innovate and succeed. Our clients are global companies with over 100k employees serving 700+ clients in 50+ countries. Rooted in a profound grasp of technology, our premier IT Headhunting Services drive transformative growth for Fortune 500 corporations. We maintain the highest standards of technological innovation to meet global demands.

Similar Jobs

Wipfli Logo Wipfli

Consultant

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
88K-118K Annually

Wipfli Logo Wipfli

Senior Manager/Director, Tax - Insurance

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
142K-192K Annually

Comcast Logo Comcast

Account Executive

Digital Media • Information Technology • News + Entertainment
Remote or Hybrid
Michigan, USA
115000 Employees

Shield AI Logo Shield AI

Staff Engineer

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Remote
USA
200K-300K Annually

Similar Companies Hiring

Axle Health Thumbnail
Artificial Intelligence • Healthtech • Information Technology • Logistics
Santa Monica, CA
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account