Principal AI/ML Engineer

Posted 7 Days Ago
Be an Early Applicant
Toronto, ON, CAN
In-Office
Expert/Leader
Fintech
The Role
Leads the architecture, development, deployment, scalability, reliability, and operational excellence of enterprise AI and machine learning systems. Partners with researchers to productionize LLMs, agentic AI, and responsible AI capabilities; establishes MLOps and LLMOps practices; oversees observability, incident response, performance, security, and governance; and mentors engineers while providing senior technical leadership.
Summary Generated by Built In
As a Principal AI Engineer, you will serve as a senior technical leader responsible for transforming state-of-the-art AI research into scalable, production-ready capabilities that create measurable value for our clients. You will lead the architecture, engineering, operationalization, and ongoing reliability of advanced AI systems, ensuring they can scale across enterprise environments while meeting rigorous standards for performance, security, resilience, and responsible AI.
This role sits at the critical intersection of AI research, engineering, product development, and operations. You will partner closely with world-class AI researchers, product leaders, and engineering teams to accelerate the journey from prototype to production. Your work will span some of the most advanced areas of AI, including Large Language Models (LLMs), Trustworthy AI, agentic systems, and emerging AI technologies.
You will mentor engineers, shape architecture, guide production support strategy, and serve as a thought leader for scaling AI across the organization. In addition to building and scaling AI solutions, you will help establish an engineering culture that emphasizes operational excellence, ownership, reliability, and continuous improvement.

Key Responsibilities:


AI Architecture & Technical Leadership 

  • Define and lead the technical architecture for enterprise-scale AI and ML platforms. 
  • Design scalable, resilient, and reusable AI systems capable of supporting mission-critical workloads. 
  • Establish architectural standards, engineering patterns, and best practices for AI deployment and operations. 
  • Drive technical decisions around model serving, inference optimization, agent architectures, orchestration frameworks, observability, and AI infrastructure. 

Productize AI Research 

  • Partner closely with AI researchers to transform cutting-edge prototypes into production-grade solutions. 
  • Lead efforts to operationalize advanced AI capabilities across areas such as: 
  • Large Language Models (LLMs) 
  • Trustworthy and Responsible AI 
  • Agentic AI Systems 
  • Establish repeatable pathways that accelerate innovation-to-production cycles. 
  • Ensure production solutions maintain scientific rigor while meeting enterprise engineering standards. 
  • Bridge the gap between research breakthroughs and sustainable business value. 

Engineering Excellence & Scalability 

  • Solve the organization's most complex AI engineering and scalability challenges. 
  • Design systems that operate reliably at enterprise scale while balancing performance, latency, governance, security, and cost. 
  • Drive adoption of MLOps, LLMOps, and AI platform engineering best practices. 
  • Improve the robustness, maintainability, observability, and operational readiness of our AI products. 
  • Identify and eliminate architectural bottlenecks that impact scale, reliability, or client experience. 
  • Raise standards through coaching, architecture reviews, design guidance, and technical leadership. 

Production Reliability & Operational Leadership 

  • Own the operational excellence, reliability, performance and availability of our products. 
  • Lead technical response and resolution efforts for complex production incidents, performance degradation, model failures, and system outages. 
  • Serve as the senior technical escalation point for the team's most challenging production challenges. 
  • Establish best practices for AI system monitoring, observability, alerting, incident management, capacity planning, and service-level objectives (SLOs). 
  • Mentor and lead junior engineers in troubleshooting, root cause analysis, operational decision-making, and incident response. 
  • Drive post-incident reviews focused on learning, continuous improvement, and long-term corrective actions. 
  • Develop operational processes that ensure AI solutions remain secure, scalable, performant, and reliable for business-critical use cases. 
  • Partner with product, infrastructure, security, and support teams to proactively identify operational risks and continuously improve service reliability. 

Mentorship & Thought Leadership 

  • Mentor AI and ML engineers within the team. 
  • Foster a culture of technical excellence and operational ownership where engineers are accountable not only for building systems, but also for running and supporting them successfully in production. 
  • Represent our team as a thought leader in scalable AI deployment, operational excellence, and responsible AI practices. 

 

Required Qualifications 

  • 10+ years of experience in software engineering, machine learning engineering, AI engineering, or related technical disciplines. 
  • Deep expertise designing, deploying, and supporting large-scale AI and ML systems in production environments. 
  • Demonstrated success leading complex technical initiatives from concept through deployment and ongoing operations. 
  • Strong knowledge of software architecture, reliability engineering, observability, ML Ops, DevOps, and cloud technologies. 
  • Proven ability to mentor engineers and lead teams through highly complex technical and operational challenges. 

Preferred Qualifications 

  • Experience with foundation models, Large Language Models, and agentic AI architectures. 
  • Experience deploying agentic AI systems and multi-agent workflows. 
  • Experience with Trustworthy AI, Responsible AI, AI governance, or model risk management frameworks. 
  • Experience optimizing large-scale inference systems and AI infrastructure. 
  • Experience working in highly regulated environments and mission-critical production systems. 

 

What You'll Gain 

This role offers a unique opportunity to operate at the forefront of applied artificial intelligence and help bridge world-class research with real-world impact. 

You will: 

  • Work directly with world-class AI researchers on breakthrough technologies and next-generation AI capabilities. 
  • Own a critical position in the pipeline that transforms cutting-edge research into client value. 
  • Tackle some of the most difficult AI engineering, scalability, and operational challenges in the industry. 
  • Build AI capabilities that deliver meaningful business outcomes for clients. 
  • Develop deep expertise in operating advanced AI systems at scale while collaborating with leaders across research, product, and engineering. 

How We Work

Vanguard has implemented a hybrid working model for the majority of our crew members, designed to capture the benefits of enhanced flexibility while enabling in-person learning, collaboration, and connection. We believe our mission-driven and highly collaborative culture is a critical enabler to support long-term client outcomes and enrich the employee experience.

Skills Required

  • 10+ years of experience in software engineering, machine learning engineering, AI engineering, or related technical disciplines
  • Deep expertise designing, deploying, and supporting large-scale AI and ML systems in production environments
  • Demonstrated success leading complex technical initiatives from concept through deployment and ongoing operations
  • Strong knowledge of software architecture, reliability engineering, observability, MLOps, DevOps, and cloud technologies
  • Proven ability to mentor engineers and lead teams through complex technical and operational challenges
  • Experience with foundation models, large language models, and agentic AI architectures
  • Experience deploying agentic AI systems and multi-agent workflows
  • Experience with Trustworthy AI, Responsible AI, AI governance, or model risk management frameworks
  • Experience optimizing large-scale inference systems and AI infrastructure
  • Experience working in highly regulated environments and mission-critical production systems

Vanguard Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Vanguard and has not been reviewed or approved by Vanguard.

  • Retirement Support — Retirement support appears unusually strong through a 401(k) design that includes a match plus an additional employer contribution, which can materially lift long-term total rewards. HSA seeding and an enhanced employer match further strengthen the savings-and-benefits value of the package.
  • Wellbeing & Lifestyle Benefits — Wellbeing and lifestyle support is reinforced by a sizable annual FlexFund stipend that can be applied across many day-to-day categories such as fitness, childcare, and other personal expenses. On-site or virtual clinics and fitness options add practical health and wellness convenience.
  • Affordable Benefits — Healthcare and related benefits are positioned as comparatively affordable via heavily subsidized medical plans and broad coverage options. This affordability can offset moderate base pay for employees who place higher value on out-of-pocket cost reductions.

Vanguard Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Malvern, PA
20,252 Employees
Year Founded: 1975

What We Do

We are a community of 30 million who think – and feel – differently about investing. Together, we’re changing the way the world invests. Since our founding in 1975, helping our investors achieve their goals is our sole reason for existence. With no other parties to answer to and therefore no conflicting loyalties, we make every decision—like keeping investing costs as low as possible—with only your needs in mind. Vanguard is one of the world's largest investment companies, offering a large selection of high-quality low-cost mutual funds, ETFs, advice, and related services. Individual and institutional investors, financial professionals, and plan sponsors can benefit from the size, stability, and experience Vanguard offers. As of April 30, 2019, we managed more than $5.6 trillion in global assets. In addition, we have 189 funds in the United States and 225 funds in global markets. For Commenting Guidelines & Important information, visit here: http://vanguard.com/linkedin Vanguard Marketing Corporation, Distributor.

Similar Jobs

TransUnion Logo TransUnion

Consumer Operations Support Specialist

Big Data • Fintech • Information Technology • Business Intelligence • Financial Services • Cybersecurity • Big Data Analytics
Hybrid
2 Locations
13000 Employees
55K-69K Annually

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Sales Associate II

eCommerce • Fashion • Retail • Sales • Wearables • Design
Hybrid
Halton Hills, ON, CAN
16000 Employees
18-22 Hourly

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Cashier II

eCommerce • Fashion • Retail • Sales • Wearables • Design
Hybrid
Halton Hills, ON, CAN
16000 Employees
18-22 Hourly

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Sales Associate II

eCommerce • Fashion • Retail • Sales • Wearables • Design
Hybrid
Halton Hills, ON, CAN
16000 Employees
18-22 Hourly

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account