Senior Cloud Platform Engineer (SMTS)

Posted Yesterday
Be an Early Applicant
2 Locations
In-Office
149K-246K Annually
Senior level
Cloud • Software
If you’re ready to build your future — and the future of technology — then you’re in the right place.
The Role
Own and operate Salesforce’s Monitoring Cloud infrastructure at scale. Design Terraform and Kubernetes automation across AWS and GCP, productize monitoring components such as Grafana, support air-gapped deployments, manage upgrades and performance, participate in on-call operations, conduct root-cause analysis, and deliver resilient platform features. Use AI-assisted development tools for infrastructure coding, testing, documentation, incident response, and automation while retaining accountability for architecture, security, reliability, and customer outcomes.
Summary Generated by Built In

To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts.

Job Category

Software Engineering

Job Details

About Salesforce

Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn’t a buzzword — it’s a way of life. The world of work as we know it is changing and we're looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce's core values at the heart of it all.

Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the right place! Agentforce is the future of AI, and you are the future of Salesforce.

Job Description: Senior Member of Technical Staff (SMTS) – Monitoring Cloud Infrastructure
Location: Bellevue / Seattle / San Francisco / Palo Alto/ Hybrid / On-Site
Role Level: Software Engineering Senior MTS
Team: Infrastructure Engineering / Monitoring Cloud

Position Overview
As a Senior Member of Technical Staff (SMTS) within our Monitoring Cloud team, you will be a key owner and operator of the systems that keep Salesforce reliable. You won't just be "using" tools; you will be productizing infrastructure to ensure our monitoring capabilities evolve at the scale of our multi-cloud footprint.
Your mission is to bridge the gap between high-level feature design and deep-system stability. From automating the "paved path" across AWS and GCP to securing air-gap environments for our most sensitive customers, you will ensure our monitoring stack is invisible, resilient, and intelligent.
This is an AI-first engineering role. You will use AI-assisted development tools (e.g., Claude Code) as the default for every inner-loop activity, code authoring, Terraform and Kubernetes scaffolding, test generation, refactoring, log/trace analysis, runbook drafting, and documentation. We expect AI to compound your throughput on routine implementation so you can focus your human judgment on architecture, security, on-call response, and customer outcomes.

Core Responsibilities
1. Infrastructure as Code (IaC) & Automation
Design and implement automation frameworks using Terraform and Kubernetes to manage monitoring infrastructure.
Standardize "paved path" deployments across AWS and GCP, eliminating manual configuration errors and ensuring global consistency.
Use AI-assisted tooling as the default for authoring, refactoring, and reviewing IaC modules, Helm charts, and automation scripts while directing intent, validating output, and owning the final result.

2. Infrastructure Upkeep & Productization
Own the lifecycle of the Monitoring Cloud stack, including version upgrades and performance tuning.
Productize core components (e.g., Grafana, custom Terraform providers) to make them consumable as reliable services by internal engineering teams.
Leverage AI for upgrade planning, release-note analysis, migration scaffolding, and boilerplate-heavy productization work (API wiring, schema plumbing, SDK generation), while retaining accountability for design and rollout. 
3. Secure & Air-Gapped Operations
Deploy and manage the full monitoring stack within highly isolated, air-gapped environments.
Ensure that our most secure customer segments receive the same level of observability and reliability as our public cloud offerings.
Apply AI assistance during development of the artifacts that ship into these environments; operate them in-network with the disciplined, human-driven workflows these environments require.
4. Operational Excellence & Health
Participate in the team’s on-call rotation, providing the deep technical expertise required to maintain strict SLAs and availability targets.
Conduct root-cause analysis (RCA) for complex system failures and implement long-term preventative fixes.
Address support requests with a “customer first” mindset
Use AI as a co-pilot during incident response and RCA: summarizing logs, correlating traces, proposing hypotheses, and drafting status updates and postmortem while the engineer remains the accountable responder and decision-maker.  
5. Next-Gen Feature Delivery
Design and deliver platform features that adhere to enterprise standards while pioneering AI-driven development practices to accelerate delivery and enhance system intelligence.
Contribute to and evolve the team's AI-assisted development playbook: prompts, agents, skills, evaluation harnesses, and guardrails that let the team ship faster without sacrificing quality or security. 

Required Qualifications
5+ years Proven track record in Distributed systems, API platforms, Infrastructure Engineering, Observability or DevOps at scale.
Proficiency with Kubernetes (K8s) and Terraform.
Hands-on experience managing infrastructure in AWS and/or GCP.
Proficiency in programming languages(eg:  java, python etc)
Experience managing or extending monitoring tools (e.g., Grafana), messaging systems (kafka etc), elastic search, caching frameworks
Security First: Understanding of authN/authZ security protocols, particularly in managing isolated or restricted network environments.
AI-assisted development fluency: demonstrated use of AI coding assistants (e.g., Claude Code) as part of a daily engineering workflow, able to prompt effectively, critically evaluate generated code, and integrate AI into IaC, testing, and automation pipelines.

Why Join This Team?
You will be at the heart of Salesforce’s "Stability First" mission. This role offers the unique challenge of operating at massive scale while solving the intricate security puzzles of air-gapped infrastructure.You'll also be on the leading edge of AI-augmented infrastructure engineering, using AI on every inner-loop activity to deliver more per engineer than has ever been possible, while still owning the human-critical work (on-call, security, and customer outcomes) that defines great infrastructure teams. If you enjoy building "infrastructure as a product" and thrive in a high-impact environment, this is your next step.

Unleash Your Potential

When you join Salesforce, you’ll be limitless in all areas of your life. Our benefits and resources support you to find balance and be your best, and our AI agents accelerate your impact so you can do your best. Together, we’ll bring the power of Agentforce to organizations of all sizes and deliver amazing experiences that customers love. Apply today to not only shape the future — but to redefine what’s possible — for yourself, for AI, and the world.

Accommodations

If you need a reasonable accommodation during the application or the recruiting process, please submit a request via this Accommodations Request Form.

Please note that Salesforce uses artificial intelligence (AI) tools to help our recruiters assess and evaluate candidates’ resumes and qualifications throughout the recruiting process. Humans will always make any candidate selection and hiring decisions. Please see our Candidate Privacy Statement for more information about how we use your personal data and your rights, including with regard to use of AI tools and opt out options.

Posting Statement

Salesforce is an equal opportunity employer and maintains a policy of non-discrimination with all employees and applicants for employment. What does that mean exactly? It means that at Salesforce, we believe in equality for all. And we believe we can lead the path to equality in part by creating a workplace that’s inclusive, and free from discrimination. Know your rights: workplace discrimination is illegal. Any employee or potential employee will be assessed on the basis of merit, competence and qualifications – without regard to race, religion, color, national origin, sex, sexual orientation, gender expression or identity, transgender status, age, disability, veteran or marital status, political viewpoint, or other classifications protected by law. This policy applies to current and prospective employees, no matter where they are in their Salesforce employment journey. It also applies to recruiting, hiring, job assignment, compensation, promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting, hiring, and promotion decisions at Salesforce are fair and based on merit. The same goes for compensation, benefits, promotions, transfers, reduction in workforce, recall, training, and education.

In the United States, compensation offered will be determined by factors such as location, job level, job-related knowledge, skills, and experience. Certain roles may be eligible for incentive compensation, equity, and benefits. Salesforce offers a variety of benefits to help you live well including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance, 401(k), and an employee stock purchasing program. More details about company benefits can be found at the following link: https://www.salesforcebenefits.com.Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, Salesforce will consider for employment qualified applicants with arrest and conviction records.

At Salesforce, we believe in equitable compensation practices that reflect the dynamic nature of labor markets across various regions. The typical base salary range for this position is $148,500 - $223,900 annually. In select cities within the San Francisco and New York City metropolitan area, the base salary range for this role is $178,900 - $246,000 annually. The range represents base salary only, and does not include company bonus, incentive for sales roles, equity or benefits, as applicable.

Skills Required

  • 5+ years of experience in distributed systems, API platforms, infrastructure engineering, observability, or DevOps at scale
  • Proficiency with Kubernetes
  • Proficiency with Terraform
  • Hands-on experience managing infrastructure in AWS and/or GCP
  • Proficiency in programming languages such as Java or Python
  • Experience managing or extending monitoring tools such as Grafana
  • Experience with messaging systems such as Kafka
  • Experience with Elasticsearch
  • Experience with caching frameworks
  • Understanding of authentication and authorization security protocols
  • Experience managing isolated or restricted network environments
  • Demonstrated fluency with AI-assisted development tools such as Claude Code
  • Ability to evaluate AI-generated code and integrate AI into infrastructure as code, testing, and automation pipelines

Salesforce Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Salesforce and has not been reviewed or approved by Salesforce.

  • Healthcare Strength Healthcare coverage is described as comprehensive, with medical, dental, and vision options, mental‑health programs, and low‑deductible or heavily subsidized tiers. Feedback suggests offerings like Lyra counseling, care navigation, and wellness reimbursements make access broad and user‑friendly.
  • Parental & Family Support Parental and family benefits are portrayed as robust, including paid leave for primary and secondary caregivers and extensive family‑building support such as fertility, adoption, and doula benefits. Backup childcare and caregiver resources provide practical help across different life stages.
  • Leave & Time Off Breadth Time‑off programs are highlighted as generous, with flexible PTO and hybrid options alongside seven days of paid Volunteer Time Off each year. Feedback suggests this combination supports work‑life balance and community engagement.

Salesforce Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
72,000 Employees

What We Do

Salesforce is the #1 AI CRM, where Humans with agents drive customer success together. Through Agentforce, our groundbreaking suite of customizable agents and tools, Salesforce brings autonomous AI agents, unified data from any source, and best-in-class Customer 360 apps together on one integrated platform to help companies connect with customers in a whole new way. Salesforce is democratizing AI agents for businesses of every size and industry so every company can embrace a workforce without limits. Our low code, open, and secure platform helps companies build and customize Salesforce fast so they can safely scale AI-powered work to every customer and employee experience and transform their business. Salesforce is proud to be the market leader, but we’re even more proud to lead in philanthropy, innovation and culture. Guided by core values of trust, customer success, innovation, equality, and sustainability, Salesforce is more than a business — we’re a platform for change.

Why Work With Us

There’s no typical day in the life of a Salesforce employee. You could be transforming our next AI innovation — or transforming your community. Closing deals — or closing your laptop for a day of Volunteer Time Off. Driving change for our customers — or driving change within one of our high-performing teams.

Gallery

Gallery

Similar Jobs

Cloudflare Logo Cloudflare

Data Center Engineer

Cloud • Information Technology • Security • Software • Cybersecurity
Hybrid
5 Locations
4400 Employees
98K-165K Annually

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Supervisor I

eCommerce • Fashion • Retail • Sales • Wearables • Design
Hybrid
North Bend, WA, USA
16000 Employees
16-25 Hourly

Onebrief Logo Onebrief

Engineering Manager

Software • Defense
Remote or Hybrid
3 Locations
350 Employees
205K-255K Annually

Sprout Social Logo Sprout Social

Account Executive

Marketing Tech • Social Media • Software • Analytics • Business Intelligence
Easy Apply
Remote or Hybrid
US
1400 Employees
150K-247K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account