SRE-2

Posted Yesterday
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Mid level
Artificial Intelligence • Computer Vision • Software
The Role
Design, deploy, and operate reliable, scalable cloud and Kubernetes systems; automate infrastructure and CI/CD; build monitoring and operational tooling; participate in incident response and root-cause analysis; collaborate with engineering teams to improve operability for large-scale, data-intensive workloads.
Summary Generated by Built In
About Signzy
Signzy is an AI-powered RPA platform for financial services. No matter how complex your workflow or operational complexity, Signzy can completely automate your back-operations decision-making process into a real-time API. This is possible due to a combination of Nebula - Our no-code AI model builder and our Fintech API Marketplace of over 200+ APIs. Today we work with over 90+ FIs globally including the 4 largest banks in India and a Top 3 acquiring Bank in the US. Globally we have a strong partnership with MasterCard and offices in New York and Dubai to serve our customers in the 2 geographies. Our Product team of 120+ people is building a global AI product out of Bangalore.
Working at Signzy
At Signzy we breathe software and exploit the latest technologies to create the most amazing products. We comprise a tech-savvy team and are backed by investors who are enthusiastic about creating solutions using technology.
This is an invitation to be a part of the future!

Role Overview
We are looking for a Site Reliability Engineer (SRE-2) to help design, operate, and improve reliable, scalable systems in cloud and Kubernetes environments. This role involves close collaboration with engineering and platform teams to automate operations, improve observability, and ensure production systems remain stable and performant as they scale.
You will work on infrastructure, deployment pipelines, and operational tooling while actively participating in incident response and long-term reliability improvements.

Responsibilities
  • Design, deploy, and operate reliable and scalable systems across cloud and Kubernetes environments.

  • Automate infrastructure provisioning, deployments, and operational workflows.

  • Build and maintain tools for deployment, monitoring, and system operations.

  • Monitor system health and performance, and proactively identify areas for improvement.

  • Troubleshoot and resolve issues across development, test, and production environments.

  • Participate in incident response, root cause analysis, and reliability improvements.

  • Collaborate with engineering teams to improve system operability and deployment safety.

  • Support and operate large-scale systems, including data-intensive or AI-driven workloads.


Requirements
  • 3–5+ years of experience managing and operating production infrastructure and services in cloud environments such as AWS, Azure, or GCP.

  • Strong hands-on experience with Linux systems in production environments.

  • Experience working with containerized workloads and Kubernetes in real-world scenarios.

  • Working knowledge of Infrastructure as Code tools such as Terraform, Terragrunt, or Crossplane.

  • Experience designing and maintaining CI/CD pipelines using tools such as GitHub Actions, GitLab CI, Jenkins, Azure DevOps, or similar.

  • Familiarity with GitOps principles and tools such as Argo CD or Flux.

  • Solid understanding of cloud networking concepts, load balancing, and service connectivity.

  • Experience with monitoring, logging, and alerting systems such as Prometheus, Grafana, ELK/EFK, Datadog, or equivalent.

  • Proficiency in at least one scripting or programming language (e.g., Bash, Python).

  • Experience working with relational databases; exposure to NoSQL or data platforms is a plus.

  • Experience participating in on-call rotations, responding to production incidents, and performing root cause analysis.

  • Understanding of SLIs, SLOs, and error budgets, and how they are used to guide reliability and operational decisions.

  • Strong problem-solving skills and the ability to debug complex production issues.

  • Good verbal and written communication skills, especially during incidents and technical discussions.

Nice to Have
  • Experience operating systems at scale or in high-availability environments.

  • Exposure to on-prem or hybrid infrastructure.

  • Experience supporting data platforms, analytics, or AI/ML workloads.

What We Value
  • A strong sense of ownership and responsibility for production systems.

  • A focus on automation, reliability, and operational simplicity.

  • The ability to balance speed, stability, and long-term maintainability.

  • Curiosity and willingness to continuously improve systems and processes.


Skills Required

  • 3-5+ years managing and operating production infrastructure and services in cloud environments (AWS, Azure, or GCP)
  • Strong hands-on experience with Linux systems in production environments
  • Experience working with containerized workloads and Kubernetes
  • Working knowledge of Infrastructure as Code tools such as Terraform, Terragrunt, or Crossplane
  • Experience designing and maintaining CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, Azure DevOps, or similar
  • Familiarity with GitOps principles and tools such as Argo CD or Flux
  • Solid understanding of cloud networking concepts, load balancing, and service connectivity
  • Experience with monitoring, logging, and alerting systems such as Prometheus, Grafana, ELK/EFK, Datadog, or equivalent
  • Proficiency in at least one scripting or programming language (e.g., Bash, Python)
  • Experience working with relational databases
  • Exposure to NoSQL or data platforms
  • Experience participating in on-call rotations, responding to production incidents, and performing root cause analysis
  • Understanding of SLIs, SLOs, and error budgets
  • Strong problem-solving skills and ability to debug complex production issues
  • Good verbal and written communication skills
  • Experience operating systems at scale or in high-availability environments
  • Exposure to on-prem or hybrid infrastructure
  • Experience supporting data platforms, analytics, or AI/ML workloads
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Bengaluru, Karnataka
331 Employees
Year Founded: 2015

What We Do

Signzy is a market-leading platform that is redefining the speed, accuracy, and experience of how financial institutions are onboarding customers and businesses - using the digital medium. The company’s award-winning no-code GO platform delivers seamless, end-to-end, and multi-channel onboarding journeys while offering totally customizable workflows. It gives these players access to an aggregated marketplace of 240+ bespoke APIs that can be easily added to any workflow with simple widgets. Signzy is enabling 10 million+ end customer and business onboardings every month at a success rate of 99% while reducing the speed to market from 6 months to 3-4 weeks. It works with over 240+ FIs globally including the 4 largest banks in India, a Top 3 acquiring Bank in the US, and has a strong global partnership with Mastercard and Microsoft. The company’s product team is based out of Bengaluru and it has a strong presence in Mumbai, New York, and Dubai.

Similar Jobs

Kong Logo Kong

Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Software • Big Data Analytics
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
800 Employees

Kong Logo Kong

Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Software • Big Data Analytics
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
800 Employees

PhonePe Logo PhonePe

Site Reliability Engineer

Fintech • Financial Services
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
1000 Employees

PhonePe Logo PhonePe

Site Reliability Engineer

Fintech • Financial Services
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
1000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account