Top Hybrid DevOps & Platform Engineering Jobs

4 Days AgoSaved
Hybrid
Boulder, CO, USA
195K-220K Annually
Senior level
195K-220K Annually
Senior level
Information Technology • Software • Quantum Computing
Designs, builds, and operates infrastructure supporting quantum control systems and research workflows across on-premises laboratories and cloud platforms. Owns infrastructure-as-code, deployment automation, observability, security, compliance, incident response, and reliability improvements. Collaborates with software, hardware, control systems, and physics teams while evaluating infrastructure tools and balancing scalability, security, cost, and operational efficiency.
Top Skills: AnsibleAWSAzureCi/CdDatadogDockerGCPGrafanaKubernetesLinuxNist 800-53PodmanPrometheusProxmoxPulumiSoc 2Terraform
Reposted 5 Days AgoSaved
Remote or Hybrid
2 Locations
285K-340K Annually
Senior level
285K-340K Annually
Senior level
Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Generative AI
Build and scale ML-optimized HPC infrastructure, manage Kubernetes-based GPU/TPU superclusters, optimize for AI/ML training, and mentor teams while innovating in ML infrastructure.
Top Skills: GoGpuJaxKubernetesLinuxNcclPythonPyTorchRdmaTensorFlowTpu
5 Days AgoSaved
Hybrid
Austin, TX, USA
100K-115K Annually
Entry level
100K-115K Annually
Entry level
Security • Software
Supports cloud infrastructure maintenance and troubleshooting across AWS and GCP environments. Assists with Python and Bash automation, CI/CD pipelines using Jenkins, Docker, and Kubernetes, application and server monitoring, issue resolution, AI integration exploration, and infrastructure documentation. The role collaborates with global DevOps teams and is designed for a recent graduate developing foundational DevOps expertise.
Top Skills: AWSAzureBashDockerGCPGitJenkinsKubernetesLinuxNetworkingPython
5 Days AgoSaved
Hybrid
Austin, TX, USA
100K-115K Annually
Entry level
100K-115K Annually
Entry level
Aerospace • Defense • Manufacturing
Supports cloud infrastructure maintenance and troubleshooting across AWS and GCP, develops Python and Bash automation, assists with CI/CD pipelines using Jenkins, Docker, and Kubernetes, monitors application and server health, explores AI-enabled DevOps workflows, and maintains technical documentation.
Top Skills: AWSAzureBashDockerGCPGitJenkinsKubernetesLinuxNetworkingPython
5 Days AgoSaved
Hybrid
Fort Lauderdale, FL, USA
Senior level
Senior level
Manufacturing
Leads an SRE and DevOps team supporting high-volume web applications, mobile backends, APIs, cloud deployments, and edge infrastructure. Establishes SLOs, SLIs, error budgets, monitoring, tracing, and alerting strategies. Oversees 24/7 incident response, post-mortems, bot mitigation, DDoS defense, WAF rules, performance tuning, infrastructure automation, and reliability improvements. Requires cross-functional leadership, operational governance, and hands-on expertise with cloud platforms, containers, CI/CD, observability, and infrastructure as code.
Top Skills: AkamaiAWSAzureCdnCi/CdDatadogDdos DefenseDistributed TracingDockerDynatraceGitlabGrafanaKubernetesTerraformWaf
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
5 Days AgoSaved
Hybrid
Palo Alto, CA, USA
140K-165K Annually
Mid level
140K-165K Annually
Mid level
Other
Operate, improve, and scale an AWS-based SaaS platform through production operations and engineering. Responsibilities include infrastructure automation, CI/CD, observability, incident response, root cause analysis, on-call support, reliability improvements, and reducing operational toil. The role partners with software engineers to improve scalability, security, operational workflows, and production readiness.
Top Skills: AWSBashDatadogDockerEc2EcsGithub ActionsGitlab Ci/CdIamInfrastructure As CodeJenkinsKubernetesPythonRdsS3TerraformVpc
5 Days AgoSaved
Hybrid
Fort Lauderdale, FL, USA
Senior level
Senior level
Transportation • Travel • Hospitality
Leads an SRE and DevOps team supporting high-volume web applications, mobile backends, APIs, edge infrastructure, and cloud deployment pipelines. Establishes SLOs, SLIs, error budgets, observability, and reliability standards; manages 24/7 incident response, post-mortems, bot mitigation, DDoS defense, WAF rules, performance tuning, and infrastructure automation. Requires cross-functional leadership, operational governance, and resilience improvements for enterprise production systems.
Top Skills: AkamaiAWSAzureCdnCi/CdDatadogDdos DefenseDockerDynatraceGitlabGrafanaKubernetesTerraformWaf
5 Days AgoSaved
Hybrid
Fort Lauderdale, FL, USA
Senior level
Senior level
Travel
Leads an SRE and DevOps team supporting high-volume web, mobile backend, and API systems. Establishes SLOs, SLIs, error budgets, observability, and automated deployment practices. Oversees 24/7 incident response, post-mortems, performance improvements, edge infrastructure, bot mitigation, DDoS defense, WAF rules, and CDN caching. Manages operational readiness, cross-functional reliability initiatives, and infrastructure automation while ensuring availability, security, resilience, and performance.
Top Skills: AkamaiAWSAzureCdnCi/CdDatadogDdos DefenseDistributed TracingDockerDynatraceGitlabGrafanaKubernetesTerraformWaf
Reposted 5 Days AgoSaved
Hybrid
Palo Alto, CA, USA
165K-180K Annually
Senior level
165K-180K Annually
Senior level
Hardware • Information Technology • Design
Design, build, and operate backend services and infrastructure powering large-scale simulation and inference. Own Terraform-based IaC, CI/CD pipelines (Jenkins), build systems (CMake, Python packaging), and harden prototypes into production-ready systems while partnering with researchers and engineers.
Top Skills: AWSAzureC++CmakeCudaDockerGCPGpu InfrastructureJenkinsKubernetesNumpyPip/UvPythonPyTorchTerraform
5 Days AgoSaved
Hybrid
Irvine, CA, USA
Expert/Leader
Expert/Leader
Artificial Intelligence • Robotics
Designs and operates AWS and Kubernetes infrastructure, CI/CD systems, GitOps workflows, and internal developer platforms. The role improves reliability, scalability, observability, deployment velocity, and operational maturity across production systems and data pipelines. Responsibilities include infrastructure as code, monitoring, autoscaling, incident response, debugging across application and infrastructure layers, and partnering with engineering teams on platform standards and developer productivity.
Top Skills: ArgocdAWSAws CdkCloudFormationDatadogGithub ActionsGitopsGoGrafanaKafkaKedaKubernetesOpentelemetryPostgresPrometheusPythonRedisTerraformTypescript
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account