AWS Platform Operations Lead, Vice President

Reposted 2 Days Ago
Be an Early Applicant
Quincy, MA, USA
In-Office
120K-218K Annually
Expert/Leader
Financial Services
The Role
Leads the strategy and transformation of an enterprise AWS platform toward AI-enabled, automated operations. Responsibilities include reliability engineering, observability, resiliency, SRE, automation, security governance, incident leadership, operational intelligence, and platform productivity. The role establishes cloud operating standards, drives AI-assisted remediation and analytics, leads disaster recovery and chaos engineering, and builds a high-performing organization across platform operations and engineering disciplines.
Summary Generated by Built In

State Street is seeking a visionary AWS Federated Platform Operations Lead to lead the evolution of cloud operations toward an intelligent, automated, and AI-enabled operating model.

This role is responsible for the operational strategy, reliability, resiliency, observability, security operations governance, and operational transformation of the AWS Federated platform. The successful candidate will partner closely with Platform Engineering, Architecture, Security, AI, and Enterprise Operations teams to deliver a secure, resilient, scalable, and highly automated cloud platform supporting the firm's most critical business workloads.

This is not a traditional infrastructure operations leadership role. Instead, the AWS Federated Platform Operations Lead will drive the next generation of cloud operations by leveraging AI, agentic workflows, automation, operational intelligence, reliability engineering, and platform-based support models to improve service quality, reduce operational risk, and accelerate innovation.

The role serves as a key member of the AWS Federated leadership team and will play a pivotal role in shaping the future multi-cloud operating model across State Street.

Key Responsibilities

Define and Execute the Cloud Operations Strategy

  • Develop and execute the long-term operational vision and roadmap for AWS Federated.
  • Establish operational standards, governance frameworks, and service management practices aligned with enterprise objectives.
  • Drive continuous improvement in operational maturity, platform reliability, customer experience, and service quality.
  • Partner with engineering leaders to integrate operability, observability, resiliency, and supportability into platform design decisions.
  • Influence enterprise cloud operating model strategy and operational transformation initiatives.

Lead the Transformation to AI-Powered Operations

  • Drive adoption of Agentic AI capabilities across cloud operations and platform management.
  • Establish a roadmap for autonomous cloud operations utilizing AI agents, automation, and operational intelligence platforms.
  • Implement AI-enabled capabilities for:
    • Incident correlation
    • Root cause analysis
    • Capacity forecasting
    • Operational insights
    • Risk identification
    • Automated remediation
  • Develop AI-powered operational assistants that improve engineering productivity and accelerate troubleshooting.
  • Partner with AI platform teams to operationalize emerging capabilities including foundation models, operational copilots, workflow orchestration, and intelligent automation.
  • Establish governance, controls, and guardrails for the responsible adoption of AI within operational processes.

Drive Reliability Engineering and Platform Resilience

  • Establish and mature Site Reliability Engineering (SRE) practices across AWS Federated.
  • Define service-level objectives (SLOs), error budgets, reliability scorecards, and resilience metrics.
  • Build platform reliability programs focused on prevention rather than reaction.
  • Lead resiliency engineering initiatives including:
    • Disaster recovery
    • Recovery automation
    • Chaos engineering
    • Failure testing
    • Cyber resiliency validation
  • Ensure operational readiness requirements are consistently incorporated into platform architecture and service enablement initiatives.

Define Next-Generation Observability and Operational Intelligence

  • Establish the strategic direction for observability across cloud, infrastructure, application, data, and security domains.
  • Create a unified operational intelligence framework that combines telemetry, operational events, security signals, and platform insights.
  • Drive adoption of modern observability capabilities including:
    • Distributed tracing
    • Metrics aggregation
    • Log analytics
    • AI-driven anomaly detection
    • Predictive analytics
  • Leverage machine learning and analytics to proactively identify capacity, performance, reliability, and operational risks.
  • Enable data-driven operational decision making through real-time dashboards and executive reporting.

Accelerate Automation and Platform Productivity

  • Establish Operations-as-Code and Automation-as-a-Product principles across the platform.
  • Lead initiatives to eliminate repetitive operational activities through automation and self-service capabilities.
  • Develop reusable automation frameworks that standardize platform operations and reduce operational complexity.
  • Drive adoption of event-driven automation and intelligent workflows.
  • Measure and improve operational efficiency through AI-assisted engineering, automation, and workflow optimization.

Strengthen Security, Compliance, and Operational Risk Management

  • Partner with Cyber Security, Risk, and Compliance teams to continuously improve platform security posture.
  • Drive operational governance for vulnerability remediation, patch strategy, cloud controls, and regulatory compliance requirements.
  • Implement automated compliance and continuous assurance capabilities wherever feasible.
  • Support cyber resilience and cyber-immunity initiatives through automation, monitoring, and operational controls.
  • Ensure cloud operational practices remain aligned with evolving enterprise security standards.

Provide Strategic Leadership During Critical Operational Events

  • Serve as the senior operational leader for AWS Federated during significant service-impacting events.
  • Provide strategic decision-making and executive communication during major incidents and platform disruptions.
  • Drive post-event learning, systemic improvement initiatives, and reliability investments.
  • Partner with engineering teams to ensure long-term corrective actions are implemented and measured.

Leadership and Organizational Development

  • Build and lead a high-performing organization spanning reliability engineering, observability, operational intelligence, automation, and platform operations disciplines.
  • Foster a culture of innovation, accountability, engineering excellence, and continuous improvement.
  • Develop future leaders with expertise in cloud engineering, AI-enabled operations, resiliency engineering, and platform management.
  • Champion adoption of emerging technologies and modern engineering practices across the organization.

Qualifications

Required Experience

  • Bachelors degree
  • 12+ years of experience in cloud platform engineering, reliability engineering, operations transformation, infrastructure engineering, or related disciplines.
  • 5+ years of leadership experience managing large-scale technology organizations.
  • Deep expertise operating enterprise-scale AWS environments in regulated industries.
  • Proven experience establishing operational strategy for complex cloud platforms.
  • Strong knowledge of cloud architecture, networking, security, identity, automation, and service management disciplines.
  • Experience leading large-scale operational transformation initiatives.
  • Demonstrated ability to influence senior executives and drive cross-functional change.

Preferred Experience

  • Financial services or highly regulated industry experience.
  • Experience implementing AIOps, operational analytics, or observability platforms.
  • Experience with platform engineering and internal developer platforms.
  • Familiarity with Agentic AI architectures, AI orchestration platforms, and intelligent automation technologies.
  • Experience with Terraform, Infrastructure as Code, GitOps, and modern CI/CD practices.
  • AWS Professional or Specialty Certifications.
  • Experience leveraging generative AI to improve engineering productivity and operational outcomes.

Leadership Attributes

The successful candidate will demonstrate:

  • Strategic thinking and enterprise leadership
  • Strong business and technology acumen
  • Customer-centric mindset
  • Data-driven decision making
  • Continuous learning mentality
  • Ability to lead through influence and collaboration
  • Passion for innovation, automation, and operational excellence

Success Measures

Success in this role will be measured through:

  • Improved platform reliability and resilience
  • Measurable reduction in operational risk

Salary Range:

$120,000 - $217,500 Annual

The range quoted above applies to the role in the location specified. If the candidate would ultimately work outside of the location above, the applicable range could differ.

Employees are eligible to participate in State Street’s comprehensive benefits program, which includes: our retirement savings plan (401K) with company match; insurance coverage including basic life, medical, dental, vision, long-term disability, and other optional additional coverages; paid-time off including vacation, sick leave, short term disability, and family care responsibilities; access to our Employee Assistance Program; incentive compensation including eligibility for annual performance-based awards (excluding certain sales roles subject to sales incentive plans); and, eligibility for certain tax advantaged savings plans.

For a full overview, visit https://hrportal.ehr.com/statestreet/Home.

About State Street

Across the globe, institutional investors rely on us to help them manage risk, respond to challenges, and drive performance and profitability. We keep our clients at the heart of everything we do, and smart, engaged employees are essential to our continued success.

We are committed to fostering an environment where every employee feels valued and empowered to reach their full potential. As an essential partner in our shared success, you’ll benefit from inclusive development opportunities, flexible work-life support, paid volunteer days, and vibrant employee networks that keep you connected to what matters most. Join us in shaping the future.

As an Equal Opportunity Employer, we consider all qualified applicants for all positions without regard to race, creed, color, religion, national origin, ancestry, ethnicity, age, disability, genetic information, sex, sexual orientation, gender identity or expression, citizenship, marital status, domestic partnership or civil union status, familial status, military and veteran status, and other characteristics protected by applicable law.

Discover more information on jobs at StateStreet.com/careers

Read our CEO Statement

Job Application Disclosure:

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Skills Required

  • Bachelor’s degree
  • 12+ years of experience in cloud platform engineering, reliability engineering, operations transformation, infrastructure engineering, or related disciplines
  • 5+ years of leadership experience managing large-scale technology organizations
  • Deep expertise operating enterprise-scale AWS environments in regulated industries
  • Experience establishing operational strategy for complex cloud platforms
  • Strong knowledge of cloud architecture, networking, security, identity, automation, and service management
  • Experience leading large-scale operational transformation initiatives
  • Ability to influence senior executives and drive cross-functional change
  • Financial services or highly regulated industry experience
  • Experience implementing AIOps, operational analytics, or observability platforms
  • Experience with platform engineering and internal developer platforms
  • Familiarity with Agentic AI architectures, AI orchestration platforms, and intelligent automation technologies
  • Experience with Terraform, Infrastructure as Code, GitOps, and modern CI/CD practices
  • AWS Professional or Specialty Certifications
  • Experience using generative AI to improve engineering productivity and operational outcomes

State Street Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about State Street and has not been reviewed or approved by State Street.

  • Retirement Support — Retirement support is framed as a standout component, highlighted by a 401(k) match described as 100% on the first 6% of base salary. This is positioned as a meaningful offset to less competitive cash compensation for some roles.
  • Leave & Time Off Breadth — Leave and time off are portrayed as relatively robust, with references to multi-week vacation, paid holidays, sick time, and additional days tied to wellness or volunteering. This breadth is repeatedly treated as a tangible part of total rewards beyond base pay.
  • Wellbeing & Lifestyle Benefits — Wellbeing and lifestyle benefits are presented as extensive, including the BeWell program, fitness discounts, onsite or supported health resources, and financial counseling. These offerings are depicted as strengthening the overall benefits proposition even when pay satisfaction is tepid.

State Street Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Boston, MA
39,782 Employees
Year Founded: 1792

What We Do

At State Street, we partner with institutional investors all over the world to provide comprehensive financial services, including investment management, investment research and trading, and investment servicing. Whether you are an asset manager, asset owner, alternative asset manager, insurance company, pension fund or official institution, you can rely on us to be focused on your challenges. We are committed to doing what it takes to help you perform better — now and in the future

Similar Jobs

Flywire Logo Flywire

Operations Specialist

Fintech • Payments • Software
Hybrid
Boston, MA, USA
1200 Employees
80K-100K Annually

Flywire Logo Flywire

Account Executive

Fintech • Payments • Software
Hybrid
Boston, MA, USA
1200 Employees
60K-80K Annually

Octus Logo Octus

Legal Workflow Strategist

Fintech • News + Entertainment • Software • Database • Financial Services
Easy Apply
Remote or Hybrid
United States
808 Employees

Liberty Mutual Insurance Logo Liberty Mutual Insurance

AVP, Senior Underwriting Officer, Large Construction

Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Remote or Hybrid
7 Locations
40000 Employees
108K-298K Annually

Similar Companies Hiring

Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account