Director of IT Infrastructure & Operations

Posted Yesterday
Be an Early Applicant
Santa Clara, CA, USA
In-Office
250K-300K Annually
Expert/Leader
Artificial Intelligence • Hardware • Software • Semiconductor
The Role
Leads IT infrastructure and operations across hybrid cloud, on-premises, HPC, engineering compute, storage, networking, security, CI/CD, and employee technology. Owns capacity planning, disaster recovery, IAM, endpoint security, compliance readiness, service desk operations, budgets, vendors, and facilities technology. Manages and grows the IT team while supporting semiconductor and AI engineering workloads, protecting design IP, and ensuring infrastructure readiness for tape-outs and product releases.
Summary Generated by Built In
Role Overview

Velaura AI is seeking an experienced Director of IT Infrastructure & Operations to own IT and infrastructure end to end: the hybrid compute environment our designers depend on, the security program that protects our design IP, the systems that support our development, and the team that runs all of it. You'll lead our existing IT team and grow it.

We are looking for someone senior enough to set a 12–24 month architecture direction and hands-on enough to be in the middle of all major IT decisions such as storage scaling, COLO expansion, uptime while managing all vendor negotiation. The ideal candidate will be productive on day one, at cohort scale. Additionally they will:

  • Ensure Compute and storage capacity is planned ahead of demand rather than in reaction to it; queue wait times and storage performance are measured and trending the right way.
  • A new software engineer starts with repo access, a working build environment, and compute quota provisioned before they log in.
  • IAM, endpoint, and access controls are in place and documented well enough to pass a customer security audit.
  • A disaster-recovery plan exists for every system a tape-out depends on, and it has been tested.
  • Hands-on — you'd rather be inside the architecture and the troubleshooting than above it
  • Scale-minded — you build for the company we're becoming, not the one we are today
  • Engineering-oriented — you treat compute availability and storage performance as product-schedule items, because they are
  • Security-conscious — you treat design data as the crown jewels and still find ways to say yes to engineers
  • Pragmatic — you balance real controls against the speed a fast-growing company needs
  • Steady — outages, security incidents, hiring waves, lab problems, IT scale problems and hard deadlines don't rattle you

    Responsibilities
  • Engineering compute. Partner with engineering leadership to forecast and optimize infrastructure for compute-intensive EDA workloads — RTL, verification, physical design, simulation, tape-out. Manage Linux HPC environments, workload schedulers (LSF, Slurm, or equivalent), and large-scale shared storage. Make sure infrastructure is ready ahead of design freezes and tape-outs, silicon bring up, and customer SDK releases, not during them.
  • Software infrastructure. Own CI/CD at scale, self-hosted runners, and cross-compilation and build farms across multiple target architectures; container and artifact registries, package and dependency mirroring, and reproducible build environments; Kubernetes for internal services and developer tooling; and source control that holds up at 200+ people, with large artifacts and deep submodule trees.
  • Hybrid infrastructure. Architect and operate Azure, Google Cloud, AWS, on-prem servers, virtualization, storage, backups, and both corporate and lab networks. Build the capacity-planning discipline for compute, storage, bandwidth, licenses, and cloud resources — and make the build-vs-buy and cloud-vs-on-prem calls.
  • IP security. Stand up the program that protects RTL, physical-design databases, foundry data, and customer information: IAM with SSO, MFA, RBAC and privileged-access controls; endpoint security, DLP, encryption, vulnerability management; secure remote access; clear boundaries for employees, contractors, and partners. Lead us through customer security reviews and toward SOC 2 / ISO 27001 readiness — you own the outcome and decide the right mix of internal hires, MSPs, and specialist security firms to get there. Own the security controls around what we ship: open-source license compliance and SBOM generation for anything we redistribute, code signing and artifact provenance for SDK and OTA payloads, and secure key handling for signed images.
  • IT Infrastructure and facilities technology. Own the IT infrastructure — including secure segmentation from corporate and engineering environments, lab access controls, test systems, and data collection, and reliable secure remote access for a distributed engineering team — and coordinate with operations on power, cooling, rack space, cabling, and UPS.
  • Employee technology. Build onboarding and offboarding that survives large hiring waves: standard images, role-based access profiles, day-one readiness. Run a service desk with real SLAs that engineers, executives, and lab users all find responsive.
  • Budget and team. Own IT OpEx and CapEx, cloud FinOps (tagging, chargeback, commitments), vendor selection and contracts, and asset management. Lead and grow the existing IT team, and decide the right mix of internal staff, MSPs, and specialists.

    Required Qualifications
  • 10+ years in IT infrastructure, including 5+ leading infrastructure and operations teams.
  • You've scaled an IT environment through rapid organizational growth and can talk candidly about what worked and what didnt.
  • Strong Linux and engineering-compute background; real depth in at least one major cloud (Azure, GCP or AWS) and comfort operating all.
  • Hands on expertise in enterprise networking (firewalls, VPNs, segmentation) and enterprise storage, virtualization, and DR.
  • Hands-on experience implementing IAM, MFA, endpoint management, logging, and enterprise security controls.
  • A track record of building the operational layer: service desk, asset management, budgets, vendor contracts, capacity planning
  • Especially interesting to us: semiconductor, EDA, hardware, or AI-infrastructure experience; supporting a hardware organization and a software organization; software supply-chain security;  familiarity with Cadence / Synopsys / Siemens EDA flows and HPC environments; having supported real tape-out deadlines; securing design IP and foundry data access, or working inside SOC 2, ISO 27001, or NIST frameworks; hardware lab or high-availability data center experience; infrastructure-as-code (Terraform, Ansible, Python); CISSP, CISM, CCNP, or a cloud architect certification

    If you meet most of this and the rest looks learnable, we'd love to hear from you.

Why Velaura?

Velaura is building next-generation compute technology for cloud, edge, and Physical AI. Our solutions will enable robots, autonomous systems, drones, and other intelligent machines to operate efficiently in the physical world.
This is an opportunity to help build foundational technology at a time when the industry is undergoing fundamental change. You will work alongside experienced leaders, architects, engineers, and operators who have delivered industry-defining products across mobile, cloud, semiconductor, and AI platforms. If you enjoy solving difficult problems, working across disciplines, and helping shape the future of Physical AI, we would love to hear from you.

Equal Employment Opportunity and Accommodations

Velaura is an Equal Opportunity Employer that is committed to inclusion and diversity. Qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, gender, sexual orientation, gender identity, disability or protected veteran status. We also take affirmative action to offer employment
opportunities to minorities, women, individuals with disabilities, and protected veterans.

Velaura is committed to working with qualified individuals with physical or mental disabilities. Applicants who would like to contact us regarding the accessibility of our website or who need special assistance or a reasonable accommodation for any part of the application or hiring process may contact us at: [email protected]. This contact
information is for accommodation requests only. Evaluation of requests for reasonable accommodation will be determined on a case-by-case basis.

Skills Required

  • 10+ years of experience in IT infrastructure
  • 5+ years leading infrastructure and operations teams
  • Experience scaling IT environments through rapid organizational growth
  • Strong Linux and engineering-compute experience
  • Deep experience in at least one major cloud platform: Azure, Google Cloud, or AWS
  • Comfort operating across Azure, Google Cloud, and AWS
  • Hands-on enterprise networking experience, including firewalls, VPNs, and segmentation
  • Hands-on enterprise storage, virtualization, and disaster recovery experience
  • Hands-on experience implementing IAM, MFA, endpoint management, logging, and enterprise security controls
  • Experience building service desk, asset management, budgets, vendor contracts, and capacity planning operations
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
104 Employees
Year Founded: 2022

What We Do

Velaura AI develops ultra-low-power silicon and software for AI compute infrastructure, serving cloud, edge, and physical-AI applications. Its patented, energy-efficient digital-design technology and power-optimized architectures support next-generation AI accelerators, helping compute systems scale with lower energy use. Drawing on deep semiconductor expertise, the company aims to make AI computing more efficient across data centers and intelligent devices, with performance and sustainability central to its approach.

Similar Jobs

Nurix Therapeutics Logo Nurix Therapeutics

Director, IT Infrastructure & Operations

Healthtech • Biotech • Pharmaceutical
In-Office
Brisbane, CA, USA
288 Employees
260K-300K Annually

T-Mobile Logo T-Mobile

Designer

Other • Utilities
In-Office
8 Locations
89016 Employees
72K-150K Annually

Outset AI Logo Outset AI

Researcher

Artificial Intelligence • Software
Hybrid
San Francisco, CA, USA
30 Employees
120K-140K Annually

Cloudflare Logo Cloudflare

Infrastructure Engineer

Cloud • Information Technology • Security • Software • Cybersecurity
Hybrid
San Francisco, CA, USA
4400 Employees
194K-266K Annually

Similar Companies Hiring

Unusual Machines, Inc. Thumbnail
Hardware • Robotics
Orlando, FL
190 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account