Staff Site Reliability Engineer

Posted Yesterday
Hiring Remotely in United States
Remote
180K-240K Annually
Senior level
Artificial Intelligence • Marketing Tech • Mobile • Software
Attentive is the AI marketing platform for leading brands.
The Role
Lead design and implementation of scalable, reliable platform systems; define SLIs/SLOs and observability; drive cross-team strategic initiatives; mentor engineers; own production standards, incident management, and cost/operational optimization to improve platform reliability and scalability.
Summary Generated by Built In
Attentive® is the AI marketing platform for 1:1 personalization redefining the way brands and people connect. We’re the only marketing platform that combines powerful technology with human expertise to build authentic customer relationships. By unifying SMS, RCS, email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue, and loyalty through real-time behavioral insights.
 
Recognized as the #1 provider in SMS Marketing by G2, Attentive partners with more than 8,000 customers across 70+ industries. Leading global brands like Crate and Barrel, Urban Outfitters, and Carter’s work with us to enable billions of interactions that power tens of billions in revenue for our customers.
 
With a distributed global workforce and employee hubs in New York City, San Francisco, London, and Sydney, Attentive’s team has been consistently recognized for its performance and culture. We’re proud to be included in Deloitte’s Fast 500 (four years running!), LinkedIn’s Top Startups, Forbes’ Cloud 100 (five years running!), Inc.’s Best Workplaces, and the Human Rights Campaign Foundation's Corporate Equality Index!

About the Role
Our Platform Infrastructure team is the backbone of everything we do at Attentive, providing a resilient and cost-effective platform that seamlessly handles billions of events from over 100 million customers daily. We own everything from compute, persistence, and networking to observability and deployments. Joining our team offers a high-growth career opportunity to collaborate with some of the world’s most talented engineers in a high-performance, high-impact culture.
As part of the Infrastructure and Platform organization, the Production Engineering Team is focused on delivering a fast and reliable platform that empowers Attentive engineers to deliver solutions quickly and safely. We build scalable systems that automate routine tasks so we can focus on other impactful efforts. Reliability, scalability, and security are our areas of expertise. We focus on release, observability, and cost optimization. Our mission is to create robust platforms and tools that allow stakeholders to concentrate on delivering exceptional products.
As a Staff Engineer, you will take a strategic role in designing and implementing solutions that enhance the reliability and scalability of our systems, while mentoring others and influencing technical roadmaps across the organization.

What You’ll Accomplish

  • Design and Deliver High-Impact Solutions: Design and implement systems that enhance reliability, observability, traceability, and incident management, ensuring the platform scales effectively
  • Lead Strategic Initiatives: Take ownership of cross-team collaborations and drive impactful projects by providing technical leadership and guidance
  • Partner Across Teams: Collaborate with engineers from AI/ML, Data, Platform, and Product teams to develop best-in-class services
  • Partner with engineers from AI/ML, Data, Platform, Product, and other groups to deliver best-in-class services
  • Establish Standards and Best Practices: Define and enforce production standards, processes, and tools to ensure operational excellence
  • Champion Reliability Goals: Advocate for and implement SLIs, SLOs, and other reliability-focused metrics across the engineering organization
  • Mentorship and Knowledge Sharing: Guide and mentor team members, fostering technical growth and helping to develop the next generation of engineering leaders
  • Innovate and Inspire: Drive continuous improvement by bringing creative ideas and challenging the status quo

Your Expertise

  • 7+ years of experience in Production Engineering, Backend Engineering, SRE, DevOps or similar role
  • Strategic visionary: Your strong technical background enables you to look beyond solving the immediate problem, planning for the future.
  • Proficient Problem-Solver: Strong coding ability in at least one language (e.g., Golang, Python, Java, Typescript) with the capability to solve complex issues through code
  • Track Record of Success: Demonstrated experience delivering medium to large-scale projects that drive meaningful improvements in platform reliability and scalability
  • Reliability Expertise: Deep understanding of production reliability concepts, including SLIs, SLOs, and incident management
  • Strong Communicator: Excellent verbal and written communication skills with the ability to influence and collaborate across technical and non-technical teams
  • Fast-Paced Experience: Familiarity with working in dynamic, reliability-focused production environments (preferred)

You'll get competitive perks and benefits, from health & wellness to equity, to help you bring your best self to work.

For US based applicants:

  • The US base salary range for this full-time position is $180,000 - $240,000 annually + equity + benefits
  • Our salary ranges are determined by role, level and location

#LI-HB1 

By applying for this position, your data will be processed as per Attentive's Privacy Policy.

Attentive Company Values
Default to Action - Move swiftly and with purpose
Be One Unstoppable Team - Rally as each other’s champions
Champion the Customer - Our success is defined by our customers' success
Act Like an Owner - Take responsibility for Attentive’s success
 
Learn more about AWAKE, Attentive’s collective of employee resource groups.
 
If you do not meet all the requirements listed here, we still encourage you to apply! No job description is perfect, and we may also have another opportunity that closely matches your skills and experience.
 
At Attentive, we know that our Company's strength lies in the diversity of our employees. Attentive is an Equal Opportunity Employer and we welcome applicants from all backgrounds. Our policy is to provide equal employment opportunities for all employees, applicants and covered individuals regardless of protected characteristics. We prioritize and maintain a fair, inclusive and equitable workplace free from discrimination, harassment, and retaliation. Attentive is also committed to providing reasonable accommodations for candidates with disabilities. If you need any assistance or reasonable accommodations, please let your recruiter know. 

Skills Required

  • 7+ years of experience in Production Engineering, Backend Engineering, SRE, DevOps, or similar role
  • Strong coding ability in at least one language (examples given: Golang, Python, Java, TypeScript)
  • Demonstrated experience delivering medium to large-scale projects that improve platform reliability and scalability
  • Deep understanding of production reliability concepts, including SLIs, SLOs, and incident management
  • Excellent verbal and written communication skills with ability to influence across technical and non-technical teams
  • Strategic systems thinking and ability to design forward-looking, scalable solutions
  • Experience mentoring and guiding engineers, fostering technical growth
  • Familiarity with working in dynamic, reliability-focused production environments

Attentive Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Attentive and has not been reviewed or approved by Attentive.

  • Healthcare Strength Health coverage includes comprehensive medical, dental, and vision plans, plus a fully covered One Medical membership and mental health resources. Employer-paid options and added wellness stipends indicate strong support for physical and mental wellbeing.
  • Equity Value & Accessibility Equity grants are described as significant and a meaningful component of total compensation. Market-value cash pay paired with stock helps position overall packages as competitive.
  • Parental & Family Support Generous paid parental leave, fertility and family-forming benefits, and supports such as Milk Stork and travel reimbursement for necessary medical care are offered. These provisions signal robust support for families across different needs.

Attentive Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
1,000 Employees
Year Founded: 2016

What We Do

Attentive® is the AI marketing platform for leading brands, designed to optimize message performance through 1:1 SMS and email interactions. Infusing intelligence at every stage of the consumer’s purchasing journey, Attentive empowers businesses to achieve hyper-personalized communication with their customers on a large scale. Leveraging AI-powered tools, a mobile-first approach, two-way conversations, and enterprise-grade technology, Attentive drives billions in online revenue for brands around the globe. Trusted by over 8,000 leading brands such as CB2, Urban Outfitters, GUESS, Dickey’s Barbeque Pit, and Wyndham Resort, Attentive is the go-to solution for delivering powerful commerce experiences for consumers with the brands they love. To learn more about Attentive or to request a demo, visit www.attentive.com or follow us on LinkedIn, X (formerly Twitter), or Instagram.

Why Work With Us

At Attentive, you'll connect with inspiring, high-caliber people, and be encouraged to take risks, get creative, and think bigger. We're solving big problems for our customers, through our innovative AI solutions, giving employees the opportunity to thrive along the journey. The sky's the limit when it comes to what's possible.

Gallery

Gallery

Similar Jobs

NBCUniversal Logo NBCUniversal

Site Reliability Engineer

AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Remote or Hybrid
New York, NY, USA

Circle Logo Circle

Site Reliability Engineer

Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
In-Office or Remote
San Francisco, CA, USA
1050 Employees
195K-258K Annually

Skydio Logo Skydio

Site Reliability Engineer

Artificial Intelligence • Hardware • Robotics • Software
In-Office or Remote
San Mateo, CA, USA
250 Employees
240K-300K Annually

Ping Identity Logo Ping Identity

Site Reliability Engineer

Cloud • Security • Software
Remote or Hybrid
USA
2300 Employees
170K-227K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account