Lead Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Toronto, ON, CAN
Hybrid
113K-142K Annually
Senior level
Software
The Role
Leads reliability, scalability, performance, automation, onboarding, incident response, observability, and disaster recovery for Azure-hosted SaaS environments. Provides technical leadership for SRE practices, manages maintenance and deployments, improves operational processes, collaborates on platform architecture, engages stakeholders and vendors, and mentors engineers. The role includes production support, infrastructure readiness, configuration management, postmortems, root cause analysis, and occasional on-call or weekend support.
Summary Generated by Built In

WHAT MAKES US, US

Join some of the most innovative thinkers in FinTech as we lead the evolution of financial technology. If you are an innovative, curious, collaborative person who embraces challenges and wants to grow, learn and pursue outcomes with our prestigious financial clients, say Hello to SimCorp!

At its foundation, SimCorp is guided by our values – caring, customer success-driven, collaborative, curious, and courageous. Our people-centered organization focuses on skills development, relationship building, and client success. We take pride in cultivating an environment where all team members can grow, feel heard, valued, and empowered.

If you like what we’re saying, keep reading!

WHY THIS ROLE IS IMPORTANT TO US

As a Lead Site Reliability Engineer, you will be embedded within one of our Product Areas, taking ownership of specific responsibility domains where your experience, interests, and growth potential align best. You will work closely with engineers, clients, and stakeholders to ensure reliability, performance, and automation for both newly onboarded and long-running clients on the SimCorp SaaS offering.

Your contributions will drive stability, continuous improvement, and operational excellence in our Azure-based environments while supporting SimCorp’s transformation into a cloud-native SaaS provider. This role blends active engineering, incident response, platform configuration, and service quality, - guided by ITIL and SRE best practices.

WHAT YOU WILL BE RESPONSIBLE FOR

          Own the reliability, scalability, and performance of Azure-hosted environments

         Lead operational support and onboarding activities for new and running client platforms

         Provide technical leadership in SRE practices, tooling, and incident response

         Guide the transition of manual processes into automated, scalable solutions

         Lead solutions workshops, engage with vendors, and manage senior stakeholders

         Implement and evolve observability frameworks using SLOs, SLIs, and proactive alerting

         Oversee disaster recovery planning, incident postmortems, and root cause analysis

         Ensure high-quality onboarding delivery through reusable automation pipelines

         Mentor and support the growth of junior and senior engineers in your Product Area

         Keeping oversight and coordinating routine maintenance, deployments, refreshes, rollbacks, and release Application Upgrades

         Execute disaster recovery, configuration management, and infrastructure readiness tasks

         Collaborate with product owners and architects to influence platform design and future roadmaps

         Provide weekend or on-call support as needed 

         Contribute to incident, problem, and change management processes

WHAT WE VALUE

         Bachelor’s or Master’s degree in Computer Science or related field

         5–8+ years in Site Reliability Engineering or Cloud Infrastructure leadership roles

         Strong expertise in Microsoft Azure, including production-grade design and operation

         Proficiency with IaC tools like Terraform, Bicep, ARM, Ansible

         Deep understanding of cloud-native monitoring and incident management frameworks

         Hands-on experience in monitoring and logging tools (Azure Monitor, Application Insights, Log Analytics, Grafana)

         Familiarity with SimCorp Dimension is a strong plus

         Experience managing both onboarding projects and live production operations

         Broad knowledge of networking, virtualization, containerization (Kubernetes, Docker)

         Broad knowledge with Linux and Windows systems, APIs, scripting (PowerShell, Bash), and SQL

         Collaborative mindset and ability to work in cross-functional teams

         Proven ability to mentor and lead engineers, influence architecture decisions, and manage complexity

         Comfort balancing strategic priorities with hands-on execution

For Toronto only: The salary range for this position is $113,000 to 142,300 CAD. Base pay may vary based on factors such as years of experience, skills and qualifications. Additionally, employees are eligible for an annual discretionary bonus and benefits including health and dental care, time off and Group RRSP/TFSA.

Benefits

SimCorp offers several benefits that might play a significant factor in considering whether to accept a job offer. Since SimCorp operates in 30+ offices worldwide, the benefits package may vary from country to country.

Simcorp follows a global hybrid policy, asking employees to work from the office two days each week while allowing remote work on other days.

NEXT STEPS

Please send us your application in English via our career site as soon as possible, we process incoming applications continually. Please note that only applications sent through our system will be processed. At SimCorp, we recognize that bias can unintentionally occur in the recruitment process. To uphold fairness and equal opportunities for all applicants, we kindly ask you to exclude personal data such as photo, age, or any non-professional information from your application. Thank you for aiding us in our endeavor to mitigate biases in our recruitment process.

If you are interested in being a part of SimCorp but are not sure this role is suitable, submit your CV anyway. SimCorp is on an exciting growth journey, and our Talent Acquisition Team is ready to assist you discover the right role for you. The approximate time to consider your CV is three weeks.

We are eager to continually improve our talent acquisition process and make everyone’s experience positive and valuable. Therefore, during the process we will ask you to provide your feedback, which is highly appreciated.

WHO WE ARE

For over 50 years, we have worked closely with investment and asset managers to become the world’s leading provider of integrated investment management solutions. We are 3,000+ colleagues with a broad range of nationalities, educations, professional experiences, ages, and backgrounds in general.

SimCorp is an independent subsidiary of the Deutsche Börse Group. Following the recent merger with Axioma, we leverage the combined strength of our brands to provide an industry-leading, full, front-to-back offering for our clients.

SimCorp is an equal opportunity employer. We are committed to building a culture where diverse perspectives and expertise are integrated in our everyday work. We believe in the continual growth and development of our employees, so that we can provide best-in-class solutions to our clients.

SimCorp Canada welcomes and encourages applications from people with disabilities. Accommodations are available upon request for candidates taking part in all aspects of the selection process. Candidates who require accommodation during the recruitment process should contact the People & Culture team at [email protected].

This position is for an existing vacancy.

 Our hiring process uses AI‑enabled tools to support the screening and assessment of applications. All applicants are still reviewed by a skilled recruitment team. AI does not make hiring decisions. Human reviewers, who are trained to understand the limitations, potential risks, and possible biases associated with AI tools, evaluate all applications and AI‑assisted results before any decision is made.

If you have questions about how your information is processed in detail, feel free to contact us.

 #LI-Hybrid

Skills Required

  • Bachelor's or Master's degree in Computer Science or a related field
  • 5-8+ years of experience in Site Reliability Engineering or cloud infrastructure leadership roles
  • Strong expertise in Microsoft Azure, including production-grade design and operation
  • Proficiency with Infrastructure as Code tools such as Terraform, Bicep, ARM, or Ansible
  • Deep understanding of cloud-native monitoring and incident management frameworks
  • Hands-on experience with Azure Monitor, Application Insights, Log Analytics, and Grafana
  • Experience managing client onboarding projects and live production operations
  • Knowledge of networking, virtualization, and containerization technologies including Kubernetes and Docker
  • Knowledge of Linux and Windows systems, APIs, scripting with PowerShell and Bash, and SQL
  • Collaborative ability to work in cross-functional teams
  • Ability to mentor and lead engineers, influence architecture decisions, and manage complexity
  • Comfort balancing strategic priorities with hands-on execution
  • Familiarity with SimCorp Dimension
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Copenhagen
3,062 Employees
Year Founded: 1971

What We Do

SimCorp is a provider of industry-leading integrated investment management solutions for the global buy side. Founded in 1971, with more than 3,000 employees across five continents, we are a truly global technology leader who empowers 40 of the world’s top 100 financial companies through our integrated platform, services, and partner ecosystem. SimCorp is a subsidiary of Deutsche Boerse Group. For more information, see www.simcorp.com.

Similar Jobs

TD Bank Logo TD Bank

Site Reliability Engineer

Fintech • Insurance • Financial Services
In-Office
Toronto, ON, CAN
93823 Employees
126K-148K Annually
Hybrid
Toronto, ON, CAN
3062 Employees
114K-156K Annually

iManage Logo iManage

Senior Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Legal Tech • Productivity • Software
Hybrid
Toronto, ON, CAN
1100 Employees
In-Office
Toronto, ON, CAN
3062 Employees
100K-137K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account