Cloud Platform Engineering​ Manager

Posted 8 Days Ago
Be an Early Applicant
Buffalo, NY, USA
In-Office
140K-233K Annually
Expert/Leader
Fintech
The Role
Leads engineering and architecture teams responsible for M&T Bank’s enterprise AI platform. Owns platform strategy, Azure API Management and Azure AI Foundry operations, gateway traffic management, model onboarding, observability, reliability, security, auditability, and regulatory controls. Oversees vendors, budgets, cloud costs, production support, service levels, incident response, and transition of externally developed capabilities. Builds and develops platform engineering talent while partnering with technology, business, risk, cybersecurity, legal, compliance, and executive stakeholders.
Summary Generated by Built In

Overview

Manages the activities of several Engineering and/or Architecture Managers and/or Team Leaders responsible for building and operating M&T Bank’s enterprise AI platform. This team owns the Azure API Management gateway and Azure AI Foundry platform that enable the Bank to adopt and deliver AI solutions at scale.

Provides strategic and operational leadership for the AI platform, including gateway management, platform lifecycle environments, model onboarding, vendor oversight, observability, reliability and resilience. Oversees the infrastructure that supports AI solutions across the Bank and ensures the platform meets applicable security, risk, regulatory and operational requirements.

Responsible for the budget of cost center(s) spanning concurrent development streams and for providing long- and short-term strategic direction across multiple applications and technology domains. Serves as a primary contact for Engineering, Architecture and Technology leadership, business partners and third-party vendors.

This is a high-impact leadership role within a small, high-agency team. The position owns the platform operating model, oversees the transition of vendor-developed capabilities to internal teams and hires, develops and leads the engineering talent responsible for the platform’s long-term operation.

Primary Responsibilities

  • Own the strategy, engineering and ongoing operation of M&T Bank’s enterprise AI platform, including the Azure API Management gateway, Azure AI Foundry platform, lifecycle environments and supporting cloud infrastructure.
  • Oversee gateway traffic management, rate limiting, identity integration, model routing, capacity planning and platform performance.
  • Establish and maintain platform service levels, operating procedures, runbooks, on-call practices, business continuity capabilities and incident response processes.
  • Ensure the AI platform is reliable, resilient, scalable and capable of supporting the Bank’s growing portfolio of AI use cases.
  • Oversee model onboarding and lifecycle management, including routing, provider integration, multi-provider resiliency and inference cost management.
  • Build and manage platform observability capabilities, including OpenTelemetry, security information and event management integration, cost and token monitoring, and same-day queryable audit trails.
  • Ensure platform controls support applicable AI safety and risk requirements, including content filtering, personally identifiable information interception, prompt shielding, access management and traceability.
  • Partner with the teams responsible for the Bank’s agentic risk framework and AI delivery capability to ensure platform design and operations align with enterprise AI governance standards.
  • Manage and participate in consultations with client management to analyze short- and long-range business requirements and recommend innovations that anticipate the future impact of changing business and technology needs.
  • Build positive relationships with clients, business partners, Engineering and Architecture leadership, Technology leadership, Risk, Cybersecurity, Legal, Compliance and other key stakeholders.
  • Monitor technology and industry direction related to AI platforms, API gateway infrastructure, cloud services, large language model serving and emerging vendor capabilities.
  • Research and initiate changes to existing technologies, processes and operating models when needed. Lead vendor and product analysis and provide recommendations to senior leadership.
  • Manage third-party delivery relationships and oversee the transition of vendor-developed platform capabilities to internal teams. Review statements of work, monitor milestones, hold partners accountable to deliverables and escalate delivery risks when appropriate.
  • Oversee multiple application development, infrastructure, support, testing, implementation and project management efforts across the AI platform.
  • Organize and direct the activities of engineering teams, assign personnel to initiatives and ensure the completion of schedules and deliverables.
  • Develop short- and long-term staffing plans. Hire, retain and develop a strong platform engineering team capable of operating critical AI infrastructure in a regulated environment.
  • Implement technology consistent with Division standards and long-range plans. Ensure adherence to all Department and Technology standards, procedures and documentation requirements.
  • Oversee solution and platform designs based on business, technology, security, risk and regulatory requirements.
  • Manage escalated platform issues and remain current on activities outside the team that may affect the platform, client environment or broader AI ecosystem.
  • Develop and manage multiple cost center budgets, including vendor and cloud platform expenses. Support cloud cost management and FinOps practices at the platform level.
  • Recommend and initiate policies and procedures that improve the performance, reliability, efficiency and effectiveness of the Department.
  • Demonstrate a strong understanding of the business environment and needs within the area of responsibility. Understand the technical, business, operational and regulatory impacts of a project, platform decision or production issue.
  • Communicate platform architecture, request flows, authorization controls, audit trails and operational processes clearly to senior management, auditors, regulators and other governance stakeholders.
  • Support the job growth and career development of team members through coaching, training plans, feedback, guidance and knowledge sharing.
  • Exercise the usual authority of a manager concerning staffing, performance appraisals, promotions, salary recommendations, performance management and terminations.
  • Understand and adhere to the Company’s risk and regulatory standards, policies and controls in accordance with the Company’s Risk Appetite. Design, implement, maintain and enhance internal controls to mitigate risk on an ongoing basis. Identify risk-related issues requiring escalation to management.
  • Promote an environment that supports belonging and reflects the M&T Bank brand.
  • Maintain M&T internal control standards, including timely implementation of internal and external audit points and the resolution of issues raised by external regulators, as applicable.
  • Complete other related duties as assigned.

Scope of Responsibilities

Oversees a team where the majority of employees are engineering and/or architecture individual contributors, Engineering Supervisors, Team Leads and/or Engineering and Architecture Managers. The team is responsible for the infrastructure and platform capabilities that support enterprise AI solutions across M&T Bank.

The position has broad accountability for platform strategy, engineering, delivery and operations, including vendor oversight, production reliability, risk and regulatory controls, and the ongoing transition of externally developed capabilities to internal ownership.

Key Skills and Success Indicators

  • Platform ownership: Approaches the platform as an enduring product and operational capability, with a focus on service levels, reliability, runbooks, on-call coverage, capacity planning and continuous improvement.
  • Engineering leadership: Provides clear technical and organizational direction while creating accountability across platform engineering, infrastructure and architecture teams.
  • Vendor management: Reviews statements of work, monitors milestones, holds partners accountable to deliverables and escalates risks when commitments are not met.
  • Regulatory awareness: Clearly explains how requests flow through the gateway, how access is authorized, where activity is logged and how platform controls support regulatory and audit requirements.
  • Operational excellence: Anticipates production risks, establishes effective monitoring and incident response practices, and ensures the platform is resilient and supportable.
  • Team building: Attracts, hires, develops and retains strong engineers while creating meaningful career opportunities within infrastructure and platform engineering.
  • Strategic judgment: Balances speed, innovation, risk, cost and long-term platform sustainability when making decisions.
  • Stakeholder partnership: Builds credibility with Technology, business, risk, control and vendor partners and communicates complex platform topics in a clear, actionable manner.

Supervisory/Managerial Responsibilities

10 to 20

Education and Experience Required

  • A combined minimum of 11 years’ higher education and/or work experience, including a minimum of 4 years’ engineering and/or architecture experience and 5 years' leadership experience including people management
  • Direct experience in platform engineering, infrastructure engineering, cloud architecture or a related technology discipline.
  • Direct experience operating API gateway infrastructure at scale, such as Azure API Management, AWS API Gateway, Kong or an equivalent platform.
  • Experience managing production gateway capabilities, including traffic management, rate limiting, identity integration and routing.
  • Experience with Azure cloud services and technologies, including Azure AI Foundry, Microsoft Entra ID, Azure Monitor and Terraform or comparable infrastructure-as-code tooling.
  • Experience managing vendor delivery relationships, reviewing contractual deliverables and overseeing the transition of vendor-developed capabilities to internal teams.
  • Experience building and operating observability capabilities, including OpenTelemetry, security information and event management integration, and cloud consumption or cost monitoring.
  • Experience designing or operating infrastructure that produces timely, queryable audit trails and supports regulatory, risk and compliance reviews.
  • Experience working in a regulated industry, such as banking, insurance or healthcare.
  • Proficiency with pertinent project management, word processing and spreadsheet applications.
  • Capable of working on multiple projects and technology initiatives of a complex nature.
  • Experience with large system enhancements, conversions, platform implementations and production problem resolution.
  • Complete understanding of the system development life cycle.
  • Excellent problem-solving skills to assist in issue resolution.
  • Familiarity with application development software, cloud services, infrastructure technologies and hardware platforms.
  • Excellent verbal and written communication skills.
  • Excellent analytical and decision-making skills.
  • Experience encouraging teamwork and serving as a role model when leading and directing others.
  • Understanding of the technical, business, operational, risk and regulatory impacts of a project, platform decision or production problem.
  • Confidence leading multiple teams, including teams and vendor partners located across different geographic locations and time zones.
  • Experience defining and managing service levels, runbooks, on-call practices, capacity planning and production support processes.
  • Prior experience presenting complex technology topics, recommendations and risks to senior management.

Education and Experience Preferred

  • Bachelor’s degree.
  • Minimum of 12 years’ technology management or large program leadership experience.
  • Experience with large language model serving infrastructure, including model routing, multi-provider failover and inference cost management.
  • Familiarity with AI safety and guardrail enforcement at the infrastructure layer, including content filtering, personally identifiable information interception and prompt shielding.
  • Experience with FinOps or cloud cost management at the platform level.
  • Experience establishing a platform engineering team or capability from an early stage and defining its operating model.
  • Extensive application and product knowledge related to AI platforms, API management, cloud infrastructure and the technology area being led.
  • Subject matter expertise in supported applications and platforms, with advanced knowledge of interfacing and integrated applications.
  • Understanding of multiple business areas and their functions.
  • Proven mentoring, people leadership and organizational development capabilities.
  • Experience with the applications, technologies and functions of the area being led.
  • Good understanding of the Bank’s application framework, technology standards and control environment.
  • Awareness of the Bank’s business plan and strategic objectives, with the ability to influence and shape technology direction.
  • Advanced knowledge of complex systems and experience supporting strategic initiatives outside normal business-as-usual operations.
  • Self-motivated and able to motivate others.

Key Skills and Success Indicators

  • Platform ownership: Approaches the platform as an enduring product and operational capability, with a focus on service levels, reliability, runbooks, on-call coverage, capacity planning and continuous improvement.
  • Engineering leadership: Provides clear technical and organizational direction while creating accountability across platform engineering, infrastructure and architecture teams.
  • Vendor management: Reviews statements of work, monitors milestones, holds partners accountable to deliverables and escalates risks when commitments are not met.
  • Regulatory awareness: Clearly explains how requests flow through the gateway, how access is authorized, where activity is logged and how platform controls support regulatory and audit requirements.
  • Operational excellence: Anticipates production risks, establishes effective monitoring and incident response practices, and ensures the platform is resilient and supportable.
  • Team building: Attracts, hires, develops and retains strong engineers while creating meaningful career opportunities within infrastructure and platform engineering.
  • Strategic judgment: Balances speed, innovation, risk, cost and long-term platform sustainability when making decisions.
  • Stakeholder partnership: Builds credibility with Technology, business, risk, control and vendor partners and communicates complex platform topics in a clear, actionable manner.

#LI-JB3

M&T Bank is committed to fair, competitive, and market-informed pay for our employees. The pay range for this position is $139,700.00 - $232,900.00 Annual (USD). The successful candidate’s particular combination of knowledge, skills, and experience will inform their specific compensation.

LocationBuffalo, New York, United States of America

Skills Required

  • At least 11 years of combined higher education and/or work experience
  • At least 4 years of engineering and/or architecture experience
  • At least 5 years of leadership experience, including people management
  • Direct experience in platform engineering, infrastructure engineering, cloud architecture, or a related technology discipline
  • Experience operating API gateway infrastructure at scale, such as Azure API Management, AWS API Gateway, Kong, or equivalent
  • Experience managing production gateway capabilities, including traffic management, rate limiting, identity integration, and routing
  • Experience with Azure cloud services and technologies, including Azure AI Foundry, Microsoft Entra ID, Azure Monitor, and Terraform or comparable infrastructure-as-code tooling
  • Experience managing vendor delivery relationships, contractual deliverables, and transitions to internal teams
  • Experience building and operating observability capabilities, including OpenTelemetry, SIEM integration, and cloud consumption or cost monitoring
  • Experience designing or operating infrastructure that produces timely, queryable audit trails and supports regulatory, risk, and compliance reviews
  • Experience working in a regulated industry such as banking, insurance, or healthcare
  • Proficiency with project management, word processing, and spreadsheet applications
  • Ability to manage multiple complex projects and technology initiatives
  • Experience with large system enhancements, conversions, platform implementations, and production problem resolution
  • Complete understanding of the system development life cycle
  • Excellent problem-solving, verbal communication, written communication, analytical, and decision-making skills
  • Experience encouraging teamwork and serving as a role model while leading and directing others
  • Understanding of technical, business, operational, risk, and regulatory impacts of technology decisions and production problems
  • Confidence leading multiple teams and vendor partners across geographic locations and time zones
  • Experience defining and managing service levels, runbooks, on-call practices, capacity planning, and production support processes
  • Experience presenting complex technology topics, recommendations, and risks to senior management
  • Bachelor’s degree
  • At least 12 years of technology management or large program leadership experience
  • Experience with large language model serving infrastructure, model routing, multi-provider failover, and inference cost management
  • Familiarity with infrastructure-layer AI safety and guardrails, including content filtering, PII interception, and prompt shielding
  • Experience with FinOps or platform-level cloud cost management
  • Experience establishing a platform engineering team or capability and defining its operating model
  • Extensive knowledge of AI platforms, API management, cloud infrastructure, and related technologies
  • Subject matter expertise in supported applications and platforms, including interfacing and integrated applications
  • Understanding of multiple business areas and their functions
  • Proven mentoring, people leadership, and organizational development capabilities
  • Advanced knowledge of complex systems and experience supporting strategic initiatives
  • Self-motivated and able to motivate others

M&T Bank Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about M&T Bank and has not been reviewed or approved by M&T Bank.

  • Retirement Support Retirement benefits are positioned as a strong pillar, including a 401(k) match and the possibility of an additional employer contribution, plus access to an employee stock purchase plan.
  • Leave & Time Off Breadth Time-off offerings are framed as competitive, with a flexible PTO approach and paid volunteer time called out as a meaningful add-on to standard leave.
  • Wellbeing & Lifestyle Benefits Wellbeing support appears comparatively robust, highlighted by mental-health therapy/coaching sessions and broader wellness programming alongside community-oriented perks.

M&T Bank Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Buffalo, NY
21,590 Employees
Year Founded: 1856

What We Do

M&T Bank is a multi-state community-focused bank serving New York, Maryland, New Jersey, Pennsylvania, Delaware, Connecticut, Virginia, West Virginia and Washington, D.C. Founded in 1856, the company provides banking, investment, insurance and mortgage financial services to more than 3.6 million consumer, business and government clients.

Similar Jobs

Valon Logo Valon

Engineering Manager

Fintech • Real Estate
Remote or Hybrid
2 Locations
500 Employees
242K-284K Annually

Samsara Logo Samsara

Analytics Manager

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
United States
4000 Employees
119K-180K Annually

Optum Logo Optum

Care Team Associate

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Fishkill, NY, USA
160000 Employees
16-29 Hourly

Optum Logo Optum

Associate Patient Care Coordinator OBGYN

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Poughkeepsie, NY, USA
160000 Employees
16-29 Hourly

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account