Top Hybrid DevOps & Platform Engineering Jobs

14 Hours AgoSaved
Remote or Hybrid
2 Locations
Entry level
Entry level
Cloud • Information Technology • Internet of Things • Consulting
Designs and implements scalable Azure cloud architectures, automates infrastructure provisioning with Terraform, and oversees Azure DevOps CI/CD pipelines. Provides technical guidance, troubleshoots complex infrastructure issues, and optimizes cloud environments for performance, cost, security, and compliance. Collaborates with stakeholders and cross-functional teams, documents architecture and processes, conducts environment assessments, and delivers Azure training while driving cloud innovation and business objectives.
Top Skills: Azure DevopsCi/CdInfrastructure As CodeAzureTerraform
YesterdaySaved
Hybrid
San Francisco, CA, USA
350K-475K Annually
Entry level
350K-475K Annually
Entry level
Artificial Intelligence • Information Technology
Design, build, and operate large GPU supercomputing clusters supporting large-scale AI training and inference. Responsibilities include cluster provisioning, automation, scheduling, orchestration, capacity planning, storage and artifact management, monitoring, reliability improvements, and performance optimization. The engineer will partner with researchers on scaling runs and advise on distributed training and infrastructure trade-offs.
Top Skills: CudaJaxKubernetesLinuxNcclPythonPyTorchRustSlurmTensorFlow
YesterdaySaved
Hybrid
Austin, TX, USA
Senior level
Senior level
Information Technology • Consulting • Energy
Leads enterprise cloud modernization and data center operations for a Texas healthcare agency. Designs hybrid cloud architectures, network topologies, migration strategies, CI/CD pipelines, Kubernetes platforms, and Infrastructure as Code using Terraform and Ansible. Ensures cloud security, identity management, healthcare data governance, and HIPAA compliance. Troubleshoots distributed systems, leads Agile technical teams, establishes operating procedures, trains staff, and supports scalable AWS, Azure, or GCP environments.
Top Skills: Ai/Ml Cloud ServicesAnsibleAWSAws CloudformationAws CloudwatchAzure MonitorAzure Resource ManagerBashCi/CdCloud CliCnappData LakesDockerDynatraceEltETLGCPGcp Cloud OperationsHclHelmInfrastructure As CodeJSONKornshellKubernetesMicroservicesAzureNoSQLOraclePowershellRubySQLSysdig SecureTerraformYaml
Reposted 2 Days AgoSaved
Hybrid
Los Altos, CA, USA
Senior level
Senior level
Artificial Intelligence
Design and operate AWS cloud infrastructure, improve CI/CD and deployment workflows, build internal platform capabilities, strengthen observability, automate infrastructure, and reduce developer friction. Partner with product teams, participate in incident response and postmortems, maintain Terraform infrastructure as code, support security and compliance controls, and establish scalable operational standards.
Top Skills: AuroraAWSCi/CdEcsEksGoKubernetesLambdaMySQLPostgresPythonRdsTerraformTypescript
Entry level
Information Technology • Software
Deploy and sustain mission-critical production systems across multiple network environments. Manage Kubernetes-based containerized applications, build repeatable CI/CD deployment processes, modify Java Spring and Vue.js components, troubleshoot infrastructure through application layers, and collaborate with DevOps, software, and systems engineering teams. The role requires onsite work in Annapolis Junction with limited remote flexibility and an active TS/SCI clearance with full-scope polygraph.
Top Skills: AngularAnsibleAWSCi/CdDockerElasticsearchGitGitlab Ci/CdJavaKubernetesLinuxMongoDBNifiPythonReactSpring BootTerraformVue
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
3 Days AgoSaved
Hybrid
San Francisco, CA, USA
210K-300K Annually
Senior level
210K-300K Annually
Senior level
Artificial Intelligence • Healthtech • Other • Productivity • Telehealth • Conversational AI • Generative AI
Lead Assort Health’s infrastructure and platform engineering team. Own GCP infrastructure, Kubernetes, observability, CI/CD, developer tooling, data infrastructure, reliability, and cross-cutting systems. Hire and coach engineers, set technical direction, lead architecture reviews, support incident response, and partner with product teams to improve developer velocity, reliability, performance, cost, and operational excellence.
Top Skills: AWSCi/CdDistributed SystemsGCPGrafanaInfrastructure-As-CodeKubernetesObservabilityPrometheusSre
3 Days AgoSaved
Hybrid
3 Locations
140K-274K Annually
Senior level
140K-274K Annually
Senior level
Artificial Intelligence • Software • Generative AI
Build and operate highly available, scalable infrastructure for WRITER’s enterprise AI platform. Responsibilities include SRE, DevOps, platform engineering, cloud infrastructure, Kubernetes, Terraform, automation, observability, incident response, post-mortems, SLOs, on-call operations, and reliability improvements. The role requires cross-functional collaboration, end-to-end ownership, systemic problem-solving, and daily use of AI-assisted development and operational tooling.
Top Skills: AWSAzureClaude CodeCodexDroidElkGCPGoGrafanaHelmKubernetesPrometheusPulumiPythonTerraform
4 Days AgoSaved
Hybrid
2 Locations
14K-14K Annually
Internship
14K-14K Annually
Internship
Information Technology • Insurance • Professional Services • Software • Analytics
Paid, in-person summer internship offering placements in business systems analysis, client success, or site reliability engineering. Interns support client implementations, troubleshoot and configure software, maintain client relationships, analyze workflows, document processes, conduct incident investigations, improve system scalability, and develop automation. Candidates must be enrolled in a bachelor’s degree program and demonstrate strong communication, collaboration, attention to detail, and technical aptitude.
Top Skills: Ai ToolsC# .NetDatadogHTMLJavaScriptNew RelicSQLSumo LogicXML
Reposted 4 Days AgoSaved
Hybrid
Redmond, WA, USA
102K-261K Annually
Senior level
102K-261K Annually
Senior level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Design, build, automate, secure, and operate Microsoft’s global routing, switching, optical, and hybrid-cloud network services. Create low-level designs, automate provisioning and validation, manage optical deployments, troubleshoot Layer 2/3 and transport systems, participate in on-call/DRI rotations, drive reliability and security improvements, mentor engineers, and collaborate across product, security, and customer teams to deliver resilient, scalable enterprise services.
Top Skills: Agentic AiAnsibleAristaAzureBgpCiscoCwdmDwdmEvpnF5Http/SIpv4Ipv6Is-IsL2VpnL3VpnLinuxMlagMplsOspfOtdrOtnPalo Alto NetworksPanoramaPythonRoadmSegment RoutingTcpUdpVxlanWindows
4 Days AgoSaved
Hybrid
New York, NY, USA
183K-247K Annually
Senior level
183K-247K Annually
Senior level
Edtech • Machine Learning • Mobile • Other • Software
Designs, operates, and improves Duolingo’s large-scale distributed systems and core infrastructure. Responsibilities include diagnosing production issues, developing platforms and automation, conducting launch reviews and root cause analyses, maintaining incident response and postmortem practices, and improving reliability, scalability, and engineering velocity. The role partners with product and platform engineering teams to reduce operational toil and ensure high-quality service delivery.
Top Skills: DockerDynamoGoJavaKotlinKubernetesMesosMySQLNomadPostgresPython
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account