Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
2 Locations
Hybrid
120K-145K Annually
Senior level
Fintech • Software • Financial Services
The Role
Build and operate reliable AWS infrastructure, Kubernetes clusters, Terraform automation, CI/CD pipelines, monitoring, and alerting. Design multi-region resilience, backups, disaster recovery, and failover strategies. Improve deployment practices, support on-call rotations, lead incident response, and perform root-cause analysis while collaborating with product engineering teams.
Summary Generated by Built In

About Luma Financial Technologies


Founded in 2018, Luma Financial Technologies (“Luma”) has pioneered a cutting-edge fintech software platform that has been adopted by broker/dealer firms, RIA offices, and private banks around the world. By using Luma, institutional and retail investors have a fully customizable, independent, buy-side technology platform that helps financial teams more efficiently learn about, research, purchase, and manage alternative investments as well as annuities. Luma gives these users the ability to oversee the full, end-to-end process lifecycle by offering a suite of solutions. These include education resources and training materials; creation and pricing of custom structured products; electronic order entry; and post-trade management. By prioritizing transparency and ease of use, Luma is a multi-issuer, multi-wholesaler, and multi-product option that advisors can utilize to best meet their clients’ specific portfolio needs. Headquartered in Cincinnati, OH, Luma also has offices in New York, NY, Miami, FL, Zurich, Switzerland and Lisbon, Portugal. For more information, please visit Luma’s website.

About the role

At Luma, our Site Reliability Engineer (SRE) team keeps our platform reliable, secure, and lightning fast. They own everything from AWS infrastructure and Kubernetes clusters to CI/CD pipelines, monitoring, and alerting. If you’re passionate about tackling big challenges, automating at scale, and making systems more resilient, we’d love to have you on the team.


What you'll do

  • Collaborate with product engineering teams to design and build the infrastructure their services run on.
  • Keep our Kubernetes clusters on AWS EKS running smoothly, secure, and ready to scale.
  • Design and deliver resilience strategies that cover multi-region architecture, backups, disaster recovery, and failover.
  • Automate infrastructure with Terraform and Infrastructure-as-Code, reducing manual effort and human error.
  • Help teams ship faster by improving CI/CD pipelines and deployment practices.
  • Monitor performance and reliability using modern observability tools.
  • Support on-call rotations and lead incident response with a focus on long-term fixes.


Qualifications

  • 5+ years of applicable experience in Site Reliability or Software Development Engineering required
  • Bachelor’s degree in Computer Science, Software Engineering or related concentration highly preferred
  • You code to solve problems and are comfortable in the following languages: Java, Java, Python, Bash and Go.
  • You have strong experience with AWS (RDS, CloudFront, IAM, VPCs), Terraform, and Kubernetes.
  • You are resilience focused, with experience designing and running systems that remain dependable during failures and recover seamlessly.
  • You have hands-on experience improving and operating CI/CD pipelines (e.g., CircleCI, GitHub Actions, or similar) to help teams ship faster with confidence.
  • You stay calm under pressure, bringing incident response expertise and strong root-cause analysis skills.
  • Most importantly, you are a team player who brings clear communication, strong collaboration, and a mindset of continuous improvement.

Luma Financial Technologies is unable to sponsor work visas or provide employment-based immigration sponsorship for this position, now or in the future. Applicants must be legally authorized to work in the United States without current or future sponsorship.

Any misrepresentation, including regarding work authorization or sponsorship needs during the application or interview process will result in disqualification from consideration for this role or termination of employment.


#LI-Hybrid, #LI-SS1

Skills Required

  • 5+ years of applicable experience in Site Reliability Engineering or Software Development Engineering
  • Bachelor's degree in Computer Science, Software Engineering, or a related concentration
  • Programming experience with Java, Python, Bash, and Go
  • Strong experience with AWS, including RDS, CloudFront, IAM, and VPCs
  • Strong experience with Terraform and Kubernetes
  • Experience designing and operating resilient systems that recover during failures
  • Hands-on experience improving and operating CI/CD pipelines, such as CircleCI or GitHub Actions
  • Incident response and root-cause analysis experience
  • Strong communication, collaboration, teamwork, and continuous-improvement mindset
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Cincinnati, Ohio
102 Employees
Year Founded: 2018

What We Do

Luma Financial Technologies provides the leading, independent, multi-issuer platform for structured products and annuities. Luma is an award-winning platform that has been used by broker dealers and their advisors nationwide for nearly a decade to more efficiently source, configure, compare and price structured products and annuities that meet their customer’s specific investment needs. Luma's advisor-centric design, renowned education and training capabilities, fully customizable deployment and complete product lifecycle support help drive the further adoption and growth of the structured products and annuities market.

Similar Jobs

MetLife Logo MetLife

Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Remote or Hybrid
United States
43000 Employees
111K-180K Annually

MetLife Logo MetLife

Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Remote or Hybrid
United States
43000 Employees
111K-180K Annually

MetLife Logo MetLife

Site Reliability Engineer

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Remote or Hybrid
United States
43000 Employees
111K-180K Annually

PwC Logo PwC

Site Reliability Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
58 Locations
370000 Employees
151K-187K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account