Top Tech Jobs & Startup Jobs

19 Days AgoSaved
Hybrid
2 Locations
184K-245K Annually
Senior level
184K-245K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Lead SOX testing and assurance across finance and operational processes, execute internal and advisory audits, evaluate control design and remediation, liaise with finance and business owners, support external auditors, and prepare audit reporting and ERM contributions.
Top Skills: AclArcherAuditboardPower BIPythonSQLWorkiva
19 Days AgoSaved
Hybrid
2 Locations
184K-245K Annually
Senior level
184K-245K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Lead and execute IT SOX testing and IT controls assessments across homegrown and third-party systems, review junior/co-sourced work, coordinate evidence and remediation, advise engineering/IT on control design, and prepare clear audit documentation and reporting for leadership and external auditors.
Top Skills: AclArcherAuditboardCi/CdCloud ComputingCobitGitIso 27001NistPower BIPythonRpaSQLWorkiva
19 Days AgoSaved
Remote or Hybrid
3 Locations
297K-440K Annually
Expert/Leader
297K-440K Annually
Expert/Leader
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Lead and grow Lambda's IAM engineering team to design, build, and operate secure, highly available identity, authorization, and key-management services (SSO, SCIM, RBAC, KMS/HSM, policy engines). Drive architecture, reliability, operations, cross-functional alignment, and enterprise adoption across Lambda Cloud.
Top Skills: APIsAudit/ObservabilityEncryptionHardware Security Module (Hsm)Key Management Service (Kms)KubernetesPolicy/Authorization EnginesRbacScimSecrets ManagementSecurity Token Service (Sts)Sso
20 Days AgoSaved
Hybrid
2 Locations
296K-346K Annually
Senior level
296K-346K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Build and operate core control-plane services for Lambda's GPU cloud: APIs, orchestration, schedulers, and host lifecycle workflows. Improve reliability, observability, deployment readiness, testing, and operational tooling. Debug production issues across distributed systems, partner with cross-functional teams, contribute to architecture and runbooks, and mentor engineers.
Top Skills: Ci/CdContainersGoKubernetesLinuxObservabilityPython
21 Days AgoSaved
Hybrid
2 Locations
255K-340K Annually
Senior level
255K-340K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Design and architect large-scale liquid-cooled HPC GPU clusters, define system requirements, develop testing and benchmarking frameworks, evaluate emerging technologies, produce architecture documentation, and provide technical leadership to engineering teams.
Top Skills: AnsibleCapacity PlanningCloud ComputingDirect-To-Chip Liquid CoolingDistributed StorageEthernetGpu ClustersHybrid CloudInfinibandKubernetesPerformance BenchmarkingSystem ValidationTerraform
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
Reposted 21 Days AgoSaved
Remote or Hybrid
3 Locations
271K-425K Annually
Expert/Leader
271K-425K Annually
Expert/Leader
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
The Account CTO will be responsible for strategic technology leadership, defining technical strategies, and driving cloud transformation initiatives at an enterprise level while mentoring technical talent and influencing product strategy based on customer needs.
Top Skills: Ai/Ml TechnologiesDistributed File SystemsHigh-Performance Parallel File SystemsInfinibandLarge-Scale Gpu Cluster DeploymentsNvidia Dgx/Hgx/Mgx SystemsNvme-Based SolutionsObject StorageRoce Networking
Reposted 21 Days AgoSaved
In-Office or Remote
2 Locations
125K-195K Annually
Senior level
125K-195K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
The Senior Incident Manager leads incident response for AI infrastructure, coordinating teams to resolve critical incidents, conducting post-incident analysis, and improving operational resilience across systems.
Top Skills: Cloud PlatformsDatadogGpu ClustersGrafanaJIRANetworkingPagerdutyPrometheusServicenow
23 Days AgoSaved
Hybrid
2 Locations
191K-255K Annually
Senior level
191K-255K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Lead end-to-end consolidation across all legal entities, manage intercompany eliminations and foreign currency, own NetSuite consolidation processes, design intercompany policy and SOX controls, coordinate consolidation close calendars, support audits and IPO-readiness, drive automation and system design for multi-entity growth, and mentor accounting staff.
Top Skills: BlacklineNetSuiteNetsuite OneworldOnestreamTrintech
Reposted 23 Days AgoSaved
Remote or Hybrid
3 Locations
266K-395K Annually
Senior level
266K-395K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Design, develop, and maintain high-performance distributed storage software and protocol APIs (file, block, object). Build scalable, resilient storage services, integrate with hardware (NVMe, GPU-direct), troubleshoot production data center issues, and participate in full SDLC for on-prem storage solutions.
Top Skills: CC++Ci/CdDockerDpuFibre ChannelGoGpuGpu-Direct StorageInfinibandIscsiKubernetesLinux Kernel InternalsLustreNfsNvmePythonRoceS3SmbSwift
Reposted 23 Days AgoSaved
Remote or Hybrid
3 Locations
314K-465K Annually
Expert/Leader
314K-465K Annually
Expert/Leader
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Lead design and implementation of high-performance distributed storage systems across object, block, and file paradigms. Drive architecture, mentor engineers, integrate storage with networking/compute/DPUs, optimize protocol performance, troubleshoot production data center issues, build benchmarking and observability tooling, and collaborate on cross-functional AI infrastructure deployments.
Top Skills: BlktraceBpftraceCC++CephDaosDpdkDpus (Nvidia Bluefield)EbpfFibre ChannelFioGoGpu-Direct StorageGrafanaInfiniband)IscsiKubernetesLustreMinioNfsNvmeNvme-OfPerfPrometheusRdma (RoceRustS3SmbSpdk
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account