About the Team
Our Product and Engineering team builds award-winning solutions on the Insight platform to help over 10,000 global organizations assess risk, detect threats, and automate security programs. Working with best-in-class technology, leading-edge research, and broad strategic expertise, our global teams continuously innovate to deliver a cohesive, highly reliable security platform.
About the Role
As a Senior Platform Operations Engineer, your primary responsibility will be to set the technical direction for how we operate, secure, and scale our platform, with a focus on our FedRAMP authorized environment. Specifically, your focus will be to:
- Define the architecture, standards, and tooling strategy for our platform infrastructure, deployment pipelines, and operational automation.
- Lead platform operations for our FedRAMP environment in-house, taking ownership of existing systems and determining what to retain, rebuild, or retire.
- Drive continuous improvements to service reliability, availability, and MTTR across services through incident response leadership and post-incident reviews.
- Design and build internal automation that improves the efficiency, reliability, and security of microservice operations and deployments.
- Partner across product engineering, security, and compliance teams to ensure platform scaling meets strict regulatory and availability requirements.
- Mentor engineers across the team to raise the technical bar through structured design reviews, code reviews, and knowledge sharing.
The skills and qualities you'll bring include
- Proven track record with 6+ years of experience in platform, infrastructure, or site reliability engineering, including hands-on experience operating cloud production infrastructure at scale across AWS, GCP, or Azure.
- Strong proficiency in Linux system administration, performance tuning, and troubleshooting alongside advanced container orchestration using Kubernetes and Docker.
- Demonstrated experience designing and maintaining infrastructure-as-code using Terraform, including module design and state management at scale.
- Hands-on experience with observability and monitoring tooling such as Grafana or Prometheus, including defining SLOs and alerting strategies.
- Familiarity with agentic development practices using modern tools, GitOps practices, Ansible, and automated CI/CD pipeline management.
- Direct experience working within or alongside FedRAMP authorized cloud environments and familiarity with continuous monitoring compliance obligations.
- Capability for Strategic Doing by taking complex, ambiguous platform challenges and returning a clear execution plan, sequencing, and evaluated trade-offs with minimal direction.
- Personal Accountability for establishing clear operational ownership, driving system reliability, and taking responsibility for platform availability outcomes.
- Ability for Navigating Change & Ambiguity while leading infrastructure migrations, platform transitions, or vendor-to-in-house system handovers.
- Commitment to Cross-Functional Collaboration by building global networks across engineering, security, and product teams to deliver sustainable platform improvements.
- Strong communication skills to convey technical objectives and rationale clearly across teams while placing customer needs at the forefront of decision-making.
- Embody our core values to foster a culture of excellence that drives meaningful impact and collective success.
We know that the best ideas and solutions come from multi-dimensional teams. That's because these teams reflect a variety of backgrounds and professional experiences. If you are excited about this role and feel your experience can make an impact, please don't be shy - apply today.
#LI-SM5
About Rapid7
At Rapid7, our vision is to create a secure digital world for our customers, our industry, and our communities. We do this by harnessing our collective expertise and passion to challenge what's possible and drive extraordinary impact. We're building a dynamic and collaborative workplace where new ideas are welcome.
Protecting 11,500+ customers against bad actors and threats means we're continuing to push the envelope just like we' ve been doing for the past 20 years. If you 're ready to solve some of the toughest challenges in cybersecurity, we're ready to help you take command of your career. Join us.
Skills Required
- 6+ years of experience in platform, infrastructure, or site reliability engineering
- Hands-on experience operating cloud production infrastructure at scale across AWS, GCP, or Azure
- Strong Linux system administration, performance tuning, and troubleshooting skills
- Advanced experience with Kubernetes and Docker container orchestration
- Experience designing and maintaining infrastructure-as-code with Terraform, including modules and state management
- Hands-on experience with observability and monitoring tools such as Grafana or Prometheus
- Experience defining service-level objectives and alerting strategies
- Familiarity with GitOps, Ansible, and automated CI/CD pipeline management
- Direct experience working within or alongside FedRAMP-authorized cloud environments
- Familiarity with continuous monitoring compliance obligations
- Ability to lead infrastructure migrations, platform transitions, or vendor-to-in-house handovers
- Strong communication and cross-functional collaboration skills
Rapid7 Compensation & Benefits Highlights
-
Leave & Time Off Breadth — Time off is highlighted by unlimited PTO for U.S. employees, plus 12 holidays and 5 global company days off within a hybrid model. These policies emphasize recharge opportunities beyond standard vacation allotments.
-
Equity Value & Accessibility — Ownership opportunities include an ESPP at a 15% discount with a lookback, and many roles also receive RSUs. This mix provides accessible paths to equity participation across functions.
-
Healthcare Strength — Core coverage includes comprehensive medical, dental, and vision plans alongside mental‑health resources. Competitive paid parental leave complements the health offering for families.
Rapid7 Insights
What We Do
At Rapid7, our vision is to create a secure digital world for our customers, our industry, and our communities. We do this by harnessing our collective expertise and passion to challenge what’s possible and drive extraordinary impact. We’re building a dynamic and collaborative workplace where new ideas are welcome. Protecting 11,000+ customers against bad actors and threats means we’re continuing to push the envelope - just like we’ve been doing for the past 20 years. If you’re ready to solve some of the toughest challenges in cybersecurity, we’re ready to help you take command of your career. Join us.
Why Work With Us
With our products, research, and open source communities, we’re building a secure digital future for everyone. This means constantly learning and evolving in an industry that’s anything but stagnant. You’ll be faced with tough challenges, and given the support to find creative solutions that drive our business, and your career forward.
Gallery
Rapid7 Offices
Hybrid Workspace
Employees engage in a combination of remote and on-site work.
Our default working model is hybrid, with employees working three days per week in the office. This approach underpins our commitment to flexibility and adaptability while supporting our dedication to development, teamwork and customer purpose.

.jpg)



















.jpg)
























