The Role
Lead and scale the site reliability engineering function: set strategy, oversee SRE projects, improve incident management and monitoring, optimize resources and budgets, drive platform engineering and developer enablement, and mentor a high-performing SRE team to ensure system reliability and operational excellence.
Summary Generated by Built In
In this role, you will impact Honeywell’s ability to maintain high system reliability and operational excellence, supporting business continuity and customer satisfaction through effective site reliability engineering practices.
ResponsibilitiesKEY RESPONSIBILITIES
- Provide strategic direction and leadership to the site reliability engineering function.
- Oversee the planning, execution, and delivery of SRE projects and initiatives.
- Collaborate with cross-functional teams and stakeholders to enhance system reliability and performance.
- Ensure adherence to best practices in site reliability, incident management, and operational excellence.
- Manage and optimize SRE resources, budgets, and timelines.
- Identify and implement process improvements to enhance efficiency and productivity.
- Lead and develop a high-performing site reliability engineering team.
YOU MUST HAVE
- Extensive experience in site reliability engineering leadership roles with proven success in managing SRE teams and projects.
- Deep understanding of Azure/GCP and hybrid architectures; able to view designs for scalability, resilience and cost optimization.
- Hands-on experience with data platforms and MLOps; able to guide model deployment strategies
- Strong expertise in reliability engineering principles, incident management, monitoring, and automation to ensure system uptime and performance.
- Proficiency with cloud platforms, container orchestration (such as Kubernetes), and infrastructure as code tools.
- Experience with programming and scripting languages used in automation and tooling development.
- Experience with Platform Engineering & Developer Enablement. Build Internal development platform consisting of templates, guardrails & self-service CI/CD capabilities.
WE VALUE
- Bachelor’s degree in Computer Science, Engineering, or a related field.
- 8+ years of experience in site reliability engineering or related fields with leadership responsibilities.
- Strong problem-solving skills and ability to drive continuous improvement in complex systems.
- Experience with DevOps practices, CI/CD pipelines, and cloud-native technologies.
- Ability to foster a culture of collaboration, innovation, and operational excellence.
Skills Required
- Extensive experience in site reliability engineering leadership managing SRE teams and projects
- Deep understanding of Azure, GCP and hybrid architectures for scalability, resilience, and cost optimization
- Hands-on experience with data platforms and MLOps, including model deployment strategies
- Expertise in reliability engineering principles, incident management, monitoring, and automation
- Proficiency with cloud platforms, container orchestration (such as Kubernetes), and infrastructure as code tools
- Experience with programming and scripting languages used for automation and tooling development
- Experience with Platform Engineering and Developer Enablement, building internal dev platforms, templates, guardrails, and self-service CI/CD
- Bachelor's degree in Computer Science, Engineering, or a related field
- 8+ years of experience in site reliability engineering or related fields with leadership responsibilities
- Experience with DevOps practices, CI/CD pipelines, and cloud-native technologies
- Strong problem-solving skills and ability to drive continuous improvement in complex systems
- Ability to foster a culture of collaboration, innovation, and operational excellence
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company





