As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production-ready platforms through standardized provisioning, automation, testing, and infrastructure validation activities. Working as part of a holistic team strategy, you will support large, complex customer deployments and ensure infrastructure environments are ready for operational handoff and long-term success.
Responsibilities:
- Provide technical expertise and engagement to support infrastructure readiness, platform engineering, and deployment activities across customer environments.
- Deploy, configure, and validate AI, GPU, and High Performance Computing (HPC) infrastructure solutions.
- Prepare and administer Kubernetes platforms, container runtimes, storage integrations, networking components, and cluster infrastructure.
- Install, configure, and validate NVIDIA technologies including GPU drivers, CUDA, GPU Operators, AI Enterprise prerequisites, and telemetry solutions.
- Validate accelerated networking technologies including InfiniBand, RoCE, RDMA, and GPU-to-GPU communications.
- Perform infrastructure readiness assessments, burn-in testing, operational acceptance testing, and performance validation activities.
- Configure and support server infrastructure including iDRAC, iLO, BMC, firmware, storage, and networking components.
- Deploy and administer Windows, Linux, VMware ESXi, Hyper-V, and KVM-based environments.
- Apply security hardening standards, compliance requirements, and operational best practices throughout deployment and validation activities.
- Develop and maintain automation workflows utilizing PowerShell, Python, Bash, and Infrastructure-as-Code methodologies.
- Create customer-facing deployment documentation, technical reports, readiness assessments, and operational validation deliverables.
- Troubleshoot complex hardware, operating system, virtualization, containerization, networking, and AI platform issues.
- Participate in advanced technical training and continued education to maintain expertise in cloud, infrastructure, AI, and platform technologies.
- Support technical engagements across customer environments and collaborate with internal engineering, architecture, and service delivery teams.
Qualifications:
- Associate degree (U.S.)/College Diploma (Canada) or equivalent combination of education and technical experience required.
- Bachelor's degree in Computer Science, Information Technology, Engineering, or related technical discipline preferred.
- 5+ years of experience in Infrastructure Engineering, Platform Engineering, Site Reliability Engineering (SRE), Systems Administration, or related technical roles.
- Experience deploying, supporting, or validating AI, GPU, HPC, or large-scale enterprise infrastructure environments.
- Experience with Kubernetes, container platforms, and enterprise Linux administration.
- Strong knowledge of server provisioning, virtualization, storage, networking, and infrastructure operations.
- Experience with VMware ESXi, Hyper-V, KVM, or related virtualization technologies.
- Experience developing automation and scripting solutions using PowerShell, Python, Bash, or similar tools.
- Knowledge of Infrastructure-as-Code and automated deployment methodologies.
- Experience with NVIDIA GPU technologies, CUDA, AI Enterprise, or related AI infrastructure platforms preferred.
- Knowledge of InfiniBand, RDMA, RoCE, or high-performance networking technologies preferred.
- Demonstrated troubleshooting, root-cause analysis, and problem-solving skills.
- Possess a customer-centric mindset and strong written and verbal communication skills.
- Possess intermediate computer skills, including proficiency with Microsoft Office applications.
- Ability to travel up to 25%.
Preferred Certifications
- Certified Kubernetes Administrator (CKA)
- Red Hat Certified System Administrator (RHCSA) or equivalent Linux certification
- NVIDIA certifications related to AI, GPU, or DGX platforms
- VMware Certified Professional (VCP) or equivalent
#LI-VR1 #Hybrid
In addition, Wesco offers a benefits program for eligible employees, which may include paid time off, medical, dental, and vision coverage, and retirement savings plans. Additional details about benefits are available here.
Skills Required
- Associate degree, college diploma, or equivalent combination of education and technical experience
- 5+ years of experience in Infrastructure Engineering, Platform Engineering, Site Reliability Engineering, Systems Administration, or related technical roles
- Experience deploying, supporting, or validating AI, GPU, HPC, or large-scale enterprise infrastructure environments
- Experience with Kubernetes, container platforms, and enterprise Linux administration
- Strong knowledge of server provisioning, virtualization, storage, networking, and infrastructure operations
- Experience with VMware ESXi, Hyper-V, KVM, or related virtualization technologies
- Experience developing automation and scripting solutions using PowerShell, Python, Bash, or similar tools
- Knowledge of Infrastructure as Code and automated deployment methodologies
- Demonstrated troubleshooting, root-cause analysis, and problem-solving skills
- Customer-centric mindset and strong written and verbal communication skills
- Intermediate computer skills, including proficiency with Microsoft Office applications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or related technical discipline
- Experience with NVIDIA GPU technologies, CUDA, AI Enterprise, or related AI infrastructure platforms
- Knowledge of InfiniBand, RDMA, RoCE, or high-performance networking technologies
- Certified Kubernetes Administrator certification
- Red Hat Certified System Administrator or equivalent Linux certification
- NVIDIA certifications related to AI, GPU, or DGX platforms
- VMware Certified Professional or equivalent certification
WESCO International Compensation & Benefits Highlights
The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about WESCO International and has not been reviewed or approved by WESCO International.
-
Retirement Support — Feedback suggests the 401(k) with company match is a relative strong point and is consistently part of the package. Retirement offerings are frequently cited as a solid element within the total rewards.
-
Leave & Time Off Breadth — Feedback suggests PTO policies are favorable, including self‑managed or generous banks in some roles along with paid holidays. Paid parental leave and related time‑off options add to the perceived strength of leave benefits.
-
Flexible Benefits — Feedback suggests employees have meaningful choice through medical plan options with HSA/FSA alongside ancillary programs like EAP, discounts, and supplemental coverages. The breadth of selectable add‑ons supports tailoring benefits to individual needs.
WESCO International Insights
What We Do
At Wesco, we believe life should run smoothly. As a leading provider of business-to-business distribution, logistics services and supply chain solutions, we create a world that you can depend on. Harnessing 100 years of ingenuity and expertise, we increase profitability, improve productivity and mitigate risk for approximately 150,000 customers worldwide. With nearly 1.5 million products and locations in more than 50 countries, Wesco is your partner in progress. Our company’s greatest asset is our people. From our corporate and field offices to our distribution sites, Wesco employs over 20,000 professionals around the globe. We’re committed to fostering diversity and inclusion across our workforce by embracing the unique perspectives, authenticity, and individuality our team members contribute to the company. Headquartered in Pittsburgh, Wesco is a publicly traded (NYSE: WCC) FORTUNE 500® company with 2022 net sales of $21.4 billion.









