The Role
Leads a Site Reliability and Performance Engineering team supporting business-critical B2B and B2C eCommerce platforms across on-premises and public-cloud environments. Responsibilities include team management, technical mentoring, monitoring and observability, performance and load testing, CI/CD automation, infrastructure management, operational issue resolution, cloud adoption, documentation, compliance, and reliability improvements. The role requires 24x7 operational support, including weekend and public holiday coverage.
Summary Generated by Built In
Role Summary
As a Manager of Information Technology at Staples, you will collaborate with a business-critical team of engineers responsible for the B2B and B2C sites performance and availability of one of the top eCommerce companies in the United States. You will be a key contributor to the success of our Public Cloud Adoption initiative. This program will drive critical technology and tangible business value utilizing the latest cloud technologies. We are looking for a highly motivated and experienced Site Reliability and Performance Engineering leader who wants to grow their career and work with cutting-edge tools and technologies. The candidate must have a proven track record of supporting B2B, B2C sites and their integrations, both on-premises and in the public cloud, with demonstrated expertise in related technologies.
Duties & Responsibilities
Oversee the day-to-day operations of the Site Reliability and Performance Engineering team.
Set clear team goals, supervise, and manage the team.
Provide technical leadership and mentoring to team members.
Engage and collaborate with cross-functional Product, Engineering, Security, Operations, Infrastructure teams and Vendors to improve MTTD and MTTR
Design, develop, and implement infrastructure & application monitoring to ensure optimal platform availability and performance
Design and execute performance testing strategies including load, stress, and capacity planning using tools such as JMeter, Locust, and LoadRunner.
Automate performance testing within CI/CD pipelines to ensure continuous validation.
Research, analyze and recommend approaches for solving challenging operational issues
Develop and maintain robust knowledge documentation for the Site Reliability Engineering team and its partners
Proactively perform analysis and identify opportunities to innovate, automate, improve efficiency, and achieve cost savings
Foster innovation by encouraging new ideas and technologies within the team.
Ensure compliance with company standards and industry best practices.
Periodically review and assess the team's performance, providing feedback and facilitating professional growth.
Requirements
Basic Qualifications
- Bachelor’s degree in Computer Science or related field with continuous and progressive experience
- Minimum of 8 years of related experience working with these technologies:
- Application Performance Management and Monitoring tools such as New Relic, AppDynamics, SiteSpect, and Datadog
- Content Delivery: Akamai
- Infrastructure monitoring tools like Zabbix, and Prometheus
- Databases eg: MongoDB, Oracle, Couchbase, Redis, MySQL
- Frameworks such as Dust/Angular, Nodejs, Springboot
- Log Analytics tools like Splunk, and ELK/Elastic
- Digital experience tools like Fullstory
- Performance Testing tools such as JMeter, Loadrunner, etc.
- Performance tuning experience with Tomcat, Node.js and Spring Boot.
- Strong understanding of non-functional requirements, performance testing processes, and defect tracking.
- 8+ years of experience with Cloud Technologies, at least half of which should be on the Microsoft Azure platform
- Strong hands-on experience with infrastructure and services (systems, network, cloud technology, provisioning, storage, etc)
- Must have strong experience with programming in one or more scripting languages (Python, Azure CLI, or Powershell)
- Hands-on experience with tool sets related to automation, orchestration, and managing infrastructure (Terraform, Puppet, Ansible, or Jenkins)
- Experience with configuring, deploying, and administering infrastructure and application monitoring tools that assist in troubleshooting performance and stability issues in a cloud environment.
- As SRE and EIRE are global operational functions providing 24x7 support, weekend and public holiday coverage is an inherent expectation of these roles.
- Eligible coverage will be offset through compensatory time off, aligned with company policy.
Preferred Qualifications
Master’s degree in Computer Science Software Engineering or a related field.
Certifications in project management or specific software development methodologies.
Experience in working with cross-functional teams and stakeholders at high organizational levels.
Skills Required
- Bachelor's degree in Computer Science or a related field
- At least 8 years of related experience with application performance management and monitoring tools
- Experience with New Relic, AppDynamics, SiteSpect, and Datadog
- Experience with Akamai content delivery technology
- Experience with Zabbix and Prometheus infrastructure monitoring
- Experience with MongoDB, Oracle, Couchbase, Redis, and MySQL databases
- Experience with Dust, Angular, Node.js, and Spring Boot frameworks
- Experience with Splunk and ELK/Elastic log analytics
- Experience with FullStory digital experience tools
- Experience with JMeter, LoadRunner, or similar performance testing tools
- Performance tuning experience with Tomcat, Node.js, and Spring Boot
- Strong understanding of non-functional requirements, performance testing processes, and defect tracking
- At least 8 years of cloud technology experience, including at least four years on Microsoft Azure
- Strong hands-on experience with systems, networks, cloud technology, provisioning, and storage infrastructure
- Strong programming experience in one or more scripting languages, including Python, Azure CLI, or PowerShell
- Hands-on experience with Terraform, Puppet, Ansible, or Jenkins for automation, orchestration, and infrastructure management
- Experience configuring, deploying, and administering infrastructure and application monitoring tools in cloud environments
- Availability for 24x7 support, including weekends and public holidays
- Master's degree in Computer Science, Software Engineering, or a related field
- Certifications in project management or specific software development methodologies
- Experience working with cross-functional teams and high-level organizational stakeholders
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
Staples India is Staples’ technology and innovation hub in Chennai, building platforms, systems, and digital solutions that support the company’s global operations and future of work. Staples serves consumers and businesses with workplace products and services, including office supplies, janitorial products, technology, furniture, breakroom essentials, print and marketing, shipping, travel, and promotional offerings. Its India teams focus on engineering, eCommerce, process optimization, and enterprise solutions.

.png)






