Data Center Engineer

Sorry, this job was removed at 10:32 p.m. (CST) on Monday, Feb 03, 2025
Houston, TX
Hybrid
81K-95K Annually
Cloud • Greentech • Other • Energy
We're on a mission to eliminate flaring and emissions in the oil field.
The Role

Crusoe is building the World’s Favorite AI-first Cloud infrastructure company. We’re pioneering vertically integrated, purpose-built AI infrastructure solutions trusted by Fortune 500 companies to power their most advanced AI applications. Crusoe is redefining AI cloud infrastructure, with a mission to align the future of computing with the future of the climate. Our AI platform is recognized as the "gold standard" for reliability and performance. Our data centers are optimized for AI workloads and are powered by clean, renewable energy.

Be part of the AI revolution with sustainable technology at Crusoe. Here, you'll drive meaningful innovation, make a tangible impact, and join a team that’s setting the pace for responsible, transformative cloud infrastructure.

About This Role:

Crusoe is building the World’s Favorite AI-first Cloud infrastructure company, pioneering vertically integrated, purpose-built AI infrastructure solutions. We are seeking a Data Center Engineer to join our operations team based in Houston, TX. This role focuses on troubleshooting and automation within our High Performance Computing networks and server environments, requiring foundational team members who thrive in dynamic environments. This is a full-time, onsite position.

What You’ll Be Working On:

  • Hardware Build Execution: Execute hardware builds, including racking, cabling, power distribution, and provisioning based on detailed design requirements and rack elevation diagrams.

  • Hardware Testing and Validation: Support burn-in and stress testing of new GPU-based servers, ensuring hardware meets operational standards.

  • Hardware and Network Troubleshooting: Troubleshoot and resolve hardware and networking issues, including GPU failures, BIOS configurations, and firmware updates.

  • Inventory Management: Troubleshoot and update hardware inventory in NetBox, ensuring accurate records of equipment, connections, and power usage.

  • Automation Development: Develop and maintain Python scripts to automate hardware management and diagnostic processes for GPU hardware such as Nvidia H100’s and similar HPC components.

  • Deployment Pipeline Development: Collaborate with cross-functional teams to create automated deployment pipelines for hardware builds and configurations.

  • Data Center Layout Implementation: Work closely with design engineers to understand and implement data center layouts and infrastructure requirements.

  • Process Documentation: Document processes, including hardware diagnostics, build execution procedures, and scripting workflows.

  • Operational Efficiency Improvement: Provide insights into improving operational efficiency and suggest new automation opportunities.

  • Vendor Collaboration: Collaborate with vendors to resolve hardware issues, escalate support tickets, and track repairs or replacements.

  • Hardware Qualification and Spares Management: Assist in the qualification of new hardware and manage spares in inventory to support ongoing operations.

What You’ll Bring to the Team:

  • Python Scripting Proficiency: Professional proficiency in Python scripting, with a focus on automation and hardware interaction.

  • GPU Hardware Experience: Professional working experience with both air and water-cooled GPU-based hardware (e.g., Nvidia H100, A100, and similar) in data center environments.

  • DCIM Tool Familiarity: Familiarity with DCIM tools such as NetBox, including its API for scripting and integrations.

  • Server Hardware Knowledge: Knowledge of server hardware, including BMC management, BIOS configuration, and firmware deployment.

  • Safety-Sensitive Position Requirement: This position is designated a safety-sensitive position and/or is located in a safety-sensitive facility. Drug and alcohol program participation is required. Must be able to pass a background check.

  • Data Center Design Understanding: Understanding of data center design principles, including power distribution, cooling, and cabling.

  • Physical Capabilities: Ability to lift 50 lbs and work in a physically challenging (sound/vibration/thermal) environment.

Bonus Points:

  • Experience with other scripting/automation tools: Experience with other scripting languages (e.g., Bash, PowerShell) or automation frameworks (e.g., Ansible, Puppet, Chef).

  • Networking Knowledge: Basic networking knowledge (e.g., TCP/IP, subnetting, VLANs) relevant to data center environments.

  • Experience with Monitoring Tools: Experience with monitoring tools and systems (e.g., Prometheus, Grafana, Zabbix).

  • Specific GPU Knowledge: Deeper understanding of specific GPU architectures, drivers, and performance optimization techniques.

  • Data Center Certifications: Relevant data center certifications (e.g., CompTIA Data+, CDCP).

  • Version Control Experience: Experience using version control systems like Git.

Benefits:

  • Industry competitive pay

  • Restricted Stock Units in a fast growing, well-funded technology company

  • Health insurance package options that include HDHP and PPO, vision, and dental for you and your dependents

  • Employer contributions to HSA accounts

  • Paid Parental Leave

  • Paid life insurance, short-term and long-term disability

  • Teladoc

  • 401(k) with a 100% match up to 4% of salary

  • Generous paid time off and holiday schedule

  • Cell phone reimbursement

  • Tuition reimbursement

  • Subscription to the Calm app

  • MetLife Legal

  • Company paid commuter benefit; $100 per pay period

Compensation:

Compensation will be paid in the range of $80,750 - $95,000. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant’s knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Similar Jobs

CrowdStrike Logo CrowdStrike

Sales Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
7 Locations
10000 Employees
65K-90K Annually

DraftKings Logo DraftKings

New Business Executive, Market Expansion

Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Remote or Hybrid
Texas, USA
6400 Employees
105K-105K Annually

CrowdStrike Logo CrowdStrike

Sr. Business Partner, Field Alliance Operations (Remote)

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
2 Locations
10000 Employees
110K-160K Annually

CrowdStrike Logo CrowdStrike

Director, Product Marketing - NG-SIEM (Remote)

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
2 Locations
10000 Employees
170K-260K Annually
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Denver, CO
667 Employees
Year Founded: 2018

What We Do

Crusoe is on a mission to eliminate routine flaring of natural gas and reduce the cost of cloud computing. We are passionate about our goals to help the oil industry operate more efficiently, achieve better relationships with communities and regulators, and improve environmental performance. Crusoe repurposes otherwise wasted energy to fuel the growing demand for computational power in the expanding digital economy.

Why Work With Us

Crusoe has five core values with each value grounded in a set of actionable practices. The combination of philosophical values and actionable practices creates a decision-making framework for each employee to achieve success at Crusoe.

Gallery

Gallery

Similar Companies Hiring

Compa Thumbnail
Software • Other • HR Tech • Business Intelligence • Artificial Intelligence
Irvine, CA
60 Employees
Amplify Platform Thumbnail
Fintech • Financial Services • Consulting • Cloud • Business Intelligence • Big Data Analytics
Scottsdale, AZ
62 Employees
Milestone Systems Thumbnail
Software • Security • Other • Big Data Analytics • Artificial Intelligence • Analytics
Lake Oswego, OR
1500 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account