Top Hybrid DevOps & Platform Engineering Jobs

5 Days AgoSaved
Hybrid
Plano, TX, USA
Internship
Internship
Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
Supports telecom network operations through alarm and event monitoring, incident tracking, troubleshooting, log and performance analysis, maintenance coordination, and service assurance. Develops automation scripts and DevOps pipelines using Python, Bash, Groovy, and Java, while contributing to AI-driven predictive operations, documentation, dashboards, and process improvements. The intern works with engineering, security, and service teams under established operational procedures.
Top Skills: BashCloudDevops PipelinesDnsGitGroovyIpJavaJSONLinuxOauth2Oss/BssPythonRanRest ApisTcp/UdpVirtualization
6 Days AgoSaved
Easy Apply
Hybrid
Chicago, IL, USA
Easy Apply
140K-180K Annually
Senior level
140K-180K Annually
Senior level
Fintech • News + Entertainment • Software • Financial Services
Own and improve the production Linux infrastructure, including performance tuning, incident response, configuration management, container orchestration, networking, proxies, virtualization, secrets, observability, and data services. Build automation in Bash and Python, maintain infrastructure code through Git and peer review, troubleshoot complex system issues, write postmortems and runbooks, and support scalable, reliable systems in a regulated fintech environment.
Top Skills: AnsibleBashBpftraceCgroupsCheckmkChefDhcpDnsElastic StackGitHaproxyIcingaIostatKubernetesKvmLinuxNagiosNamespacesNginxNomadPerfPuppetPythonRabbitMQRedisSaltSsStraceSystemdTcp/IpTlsVaultVMwareXen
6 Days AgoSaved
Hybrid
San Francisco, CA, USA
230K-280K Annually
Senior level
230K-280K Annually
Senior level
Beauty • Enterprise Web • Fintech • Payments • Software
Lead Compute, Data, and DevEx teams while owning infrastructure and data-platform roadmaps, reliability, cloud costs, vendor strategy, and developer self-service. Establish SLI/SLO practices, incident response, post-mortems, CI standards, infrastructure-as-code tooling, and AI-enabled production operations. Manage managers and technical leads, scale the organization, develop engineers, and ensure systems support significantly higher traffic while maintaining 99.95% uptime across AWS regions.
Top Skills: ArgocdAWSCiDistributed SystemsInfrastructure As CodeKubernetesObservabilitySliSloTerraform
6 Days AgoSaved
Hybrid
4 Locations
137K-257K Annually
Senior level
137K-257K Annually
Senior level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Leads and manages a platform engineering team responsible for an enterprise-scale Kubernetes platform. Oversees hiring, coaching, performance, workforce planning, vendor management, platform reliability, security, incident management, infrastructure modernization, developer self-service, and delivery of strategic technology initiatives. Provides technical guidance, establishes operational metrics, manages dependencies and technical debt, and partners with product, architecture, engineering, and executive stakeholders.
Top Skills: Ai-Assisted EngineeringAmazon EksCi/CdCloud-Native InfrastructureDevOpsInfrastructure As CodeKubernetes
6 Days AgoSaved
Remote or Hybrid
United States
Mid level
Mid level
Artificial Intelligence • Cloud • Payments • Software • Business Intelligence • Generative AI • Automation
Administer hybrid Windows Server infrastructure across colocation, on-premises, Azure, and GCP environments. Manage virtual machines, cloud resources, patching, upgrades, backups, monitoring, alerts, incident response, vulnerability remediation, and root cause analysis. Develop automation using PowerShell, Python, Ansible, and infrastructure-as-code tools for provisioning, remediation, and self-healing workflows. Maintain documentation and runbooks while collaborating with engineering, IT support, security, and operations teams.
Top Skills: AnsibleAppdynamicsArm TemplatesBashDatadogEsxiGitlab Ci/CdGoogle Cloud Platform (Gcp)GrafanaAzureNew RelicOpentelemetryPatchmypcPowershellPythonSccmSignozSolarwindsTerraformVcenterVmware VsphereWindows ServerWsus
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
7 Days AgoSaved
Easy Apply
Hybrid
Atlanta, GA, USA
Easy Apply
98K-149K Annually
Entry level
98K-149K Annually
Entry level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Build and operate foundational infrastructure for PagerDuty’s real-time platform, including networking, compute, Kubernetes, and ingress systems. Improve reliability, scalability, and security; monitor system health through metrics, logs, and alerts; participate in 24/7 on-call rotations; support infrastructure rollouts; and contribute to agile planning and technical improvements.
Top Skills: Amazon EksAWSAzureCloudFormationDatadogDnsEnvoyGCPGoGrafanaIstioKubernetesLinuxNew RelicNginxPrometheusPythonRubySplunkSumo LogicTerraformTls
7 Days AgoSaved
Hybrid
Lake Forest, IL, USA
96K-173K Annually
Senior level
96K-173K Annually
Senior level
eCommerce • Information Technology • Retail • Industrial
Build, configure, and support hybrid cloud and on-premises infrastructure for enterprise SAP applications. Develop automation workflows, scripts, CI/CD processes, and configuration-management standards using tools such as Ansible and Terraform. Monitor systems, establish alert thresholds and service levels, document infrastructure solutions, and optimize performance. After onboarding, provide escalated SAP Basis support, including installations, upgrades, patching, troubleshooting, and administration across SAP workloads and HANA databases.
Top Skills: AnsibleAWSCi/CdCmdbConfiguration As CodeDevOpsGitInfrastructure As CodeLinuxRed Hat Enterprise LinuxSap ApoSap BobjSap BodsSap EccSap EwmSap GrcSap GtsSap HanaSap Pi/PoSap PortalSap S/4HanaSap Solution ManagerSuse LinuxTerraform
7 Days AgoSaved
Hybrid
Phoenix, AZ, USA
Senior level
Senior level
Aerospace • Artificial Intelligence • Cloud • Machine Learning • Software • Cybersecurity • Defense
Design and manage monitoring and alerting systems, improve service availability and performance, collaborate with development teams on reliable architectures and automation, and conduct root cause analysis to prevent recurring incidents. The role supports a global 24x7x365 team from Phoenix on a hybrid schedule and requires substantial site reliability engineering experience, cloud and orchestration knowledge, scripting proficiency, and monitoring expertise.
Top Skills: BashCloud ServicesContainerizationGrafanaOrchestrationPrometheusPython
7 Days AgoSaved
Easy Apply
Hybrid
Chicago, IL, USA
Easy Apply
239K-239K Annually
Senior level
239K-239K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Mobile • Wearables • Analytics
Leads the strategy, roadmap, and execution of Prolaio’s cloud-native platform engineering capabilities across GCP environments. Owns observability, infrastructure tooling, access controls, data lifecycle management, platform services, automation, governance, and AI-enabled features. Partners with engineering, data, security, quality, product, and clinical teams to deliver secure, compliant, auditable healthcare data platforms. Builds and mentors a high-performing platform engineering organization supporting clinical research, analytics, data science, and enterprise data services.
Top Skills: .NetAngularGCPJavaPythonReactSQL
8 Days AgoSaved
Easy Apply
Hybrid
3 Locations
Easy Apply
186K-232K Annually
Senior level
186K-232K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Biotech • Pharmaceutical
Lead the infrastructure and SRE team, setting technical direction for reliable, secure, and scalable engineering systems. Hire, coach, and develop engineers; oversee multi-cloud infrastructure, Kubernetes, IaC, CI/CD, observability, incident response, support rotations, and regulated workloads. Partner across engineering, data science, security, and operations to improve developer experience and enable safe AI-assisted software delivery. Manage infrastructure roadmaps, budgets, resourcing, documentation, and organizational reliability practices.
Top Skills: Ai Model Training WorkloadsAWSAzureCi/CdContainerized ApplicationsCots SoftwareDatabasesDockerFoss SoftwareGCPGitInfrastructure As CodeKubernetesLoad BalancersMl PipelinesMulti-Cloud InfrastructureOpentofuPythonSecrets ManagementSnowflakeTerraformVercelVirtual Networking
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account