Maximum of 25 job preferences reached.
Top Engineering Jobs
Cloud • Information Technology • Machine Learning
Develop, deploy, monitor, and improve services managing bare-metal infrastructure. Lead incident response, root-cause analysis, post-incident reviews, observability, automation, reliability engineering, disaster recovery, and on-call support. Build CI/CD pipelines, dashboards, alerts, and hardware lifecycle tooling using Go or Python, Kubernetes, Prometheus, Grafana, and Redfish-based services. Collaborate with engineering, fleet operations, vendors, and upstream communities to improve platform stability and reduce operational incidents.
Top Skills:
AWSCi/CdGCPGoGrafanaKubernetesPrometheusPythonRedfish
Cloud • Information Technology • Machine Learning
Lead customer-facing technical engagements focused on networking for HPC/cloud environments. Design, prototype, and deploy Kubernetes-based solutions, optimize customer workloads, contribute product feedback, run proofs-of-concept, and represent CoreWeave at events. Collaborate with engineering teams on product improvements and R&D for emerging solutions.
Top Skills:
Cloud ComputingHpcInfinibandKubernetesKubernetes CsiNcclNvidia Gpus
Cloud • Information Technology • Machine Learning
Leads technical due diligence for prospective data center sites, evaluating electrical, mechanical, civil, structural, architectural, utility, zoning, and infrastructure feasibility. Reviews provider proposals, capacity data, conceptual layouts, and building documentation; identifies risks, upgrades, cost and schedule impacts; prepares go/no-go recommendations and technical assessments. Coordinates with utilities, landlords, developers, internal subject matter experts, and design teams, then hands secured sites to Design Managers. Frequent travel is required.
Top Skills:
CdusChilled Water PlantsDashboardsIt/LvMepPower DistributionSubstationsTcs LoopsUtility Infrastructure
Cloud • Information Technology • Machine Learning
Design and review data center white space infrastructure for high-density computing environments. Coordinate electrical, mechanical, cooling, low-voltage, cabling, and network pathway systems; review drawings, submittals, RFIs, and field conditions; oversee consultants; resolve design and constructability issues; and support construction, commissioning, and deployment. Contribute to design standards, scalability, technical documentation, and mentorship across CoreWeave’s data center portfolio.
Top Skills:
AsanaAshraeAutocadBicsiBluebeamBranch CircuitsBuswaysCable TrayFan WallsGrounding SystemsHot Aisle ContainmentLadder RackLiquid CoolingNfpaOverhead ConveyancePdusRack Power DistributionRemote CdusRevitRppsSmartsheetStructured CablingTia-942Uptime Institute
Cloud • Information Technology • Machine Learning
Build and operate full-stack applications and AI-facing features for CoreWeave’s internal data platform. Responsibilities include developing TypeScript, React, Next.js, and Python services; designing APIs and relational data models; integrating analytical data platforms; deploying on Kubernetes; and owning production reliability, testing, observability, and on-call support. The role also involves collaborating with non-engineers to define processes, shipping LLM-backed applications, and evaluating AI output quality.
Top Skills:
BigQueryCi/CdDelta LakeDockerHelmHudiIcebergKafkaKubernetesLanggraphNext.JsPostgresPythonReactSnowflakeSparkSQLStarrocksTypescript
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Cloud • Information Technology • Machine Learning
Build and ship full-stack software for CoreWeave’s internal data infrastructure and operational applications. Develop TypeScript, React, and Next.js frontends; Python services on Kubernetes; SQL queries and data models; and AI-enabled interfaces such as text-to-SQL tools, retrieval systems, and agents. Own scoped features through production, including testing and post-release fixes, while collaborating with senior engineers and internal users.
Top Skills:
BigQueryCi/CdDockerGitHelmHttp ApisJavaScriptKubernetesLlmsNext.JsPythonReactSnowflakeSparkSQLStarrocksTypescript
Cloud • Information Technology • Machine Learning
Serves as a technical lead for customers using CoreWeave cloud infrastructure, with emphasis on Kubernetes in high-performance computing environments. Responsibilities include designing and deploying tailored solutions, leading proofs of concept, optimizing AI/ML workloads, conducting technical reviews, advising on product strategy, collaborating with engineering teams, and representing CoreWeave at industry events.
Top Skills:
Artificial IntelligenceAutomationCloud ComputingDistributed SystemsHigh-Performance ComputingInference FrameworksInfinibandKubernetesMachine LearningMulti-CloudNvidia Collective Communications Library (Nccl)Nvidia GpusScriptingSlurm
Cloud • Information Technology • Machine Learning
Leads systems engineering teams developing high-throughput file, block, and object storage for AI cloud infrastructure. Drives architecture, performance tuning, latency optimization, benchmarking, and 10x scaling across distributed storage systems. Provides technical mentorship, performance management, recruitment, and cross-functional leadership while maintaining a coding-focused culture and maximizing GPU utilization.
Top Skills:
CephCloud StorageDistributed SystemsGoGpudirect StorageHigh-Performance ComputingLustreMinioRust
Cloud • Information Technology • Machine Learning
The role involves leading the Observability Insights initiative at CoreWeave, focusing on developing APIs, agentic tools, and improving telemetry for AI systems. Candidates should possess significant experience in backend engineering and observability systems, especially for multi-tenant environments.
Top Skills:
ClickhouseGoGrafanaKubernetesLokiPrometheusPythonVictoria Metrics
Cloud • Information Technology • Machine Learning
Lead and develop an infrastructure engineering team responsible for internal Kubernetes platforms and foundational services. Set roadmaps, ownership, operating practices, and success metrics; guide reliable infrastructure and automation; establish testing, observability, SLO, incident response, change management, and on-call practices; resolve cross-team dependencies; and manage hiring, coaching, performance, priorities, and risk communication.
Top Skills:
Cloud InfrastructureDistributed SystemsGitopsInfrastructure AutomationKubernetesObservabilityProgressive DeliverySlos
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
All Filters
Total selected ()
No Results
No Results


