- Design, deploy, and operate reliable and scalable systems across cloud and Kubernetes environments.
- Automate infrastructure provisioning, deployments, and operational workflows.
- Build and maintain tools for deployment, monitoring, and system operations.
- Monitor system health and performance, and proactively identify areas for improvement.
- Troubleshoot and resolve issues across development, test, and production environments.
- Participate in incident response, root cause analysis, and reliability improvements.
- Collaborate with engineering teams to improve system operability and deployment safety.
- Support and operate large-scale systems, including data-intensive or AI-driven workloads.
- 3–5+ years of experience managing and operating production infrastructure and services in cloud environments such as AWS, Azure, or GCP.
- Strong hands-on experience with Linux systems in production environments.
- Experience working with containerized workloads and Kubernetes in real-world scenarios.
- Working knowledge of Infrastructure as Code tools such as Terraform, Terragrunt, or Crossplane.
- Experience designing and maintaining CI/CD pipelines using tools such as GitHub Actions, GitLab CI, Jenkins, Azure DevOps, or similar.
- Familiarity with GitOps principles and tools such as Argo CD or Flux.
- Solid understanding of cloud networking concepts, load balancing, and service connectivity.
- Experience with monitoring, logging, and alerting systems such as Prometheus, Grafana, ELK/EFK, Datadog, or equivalent.
- Proficiency in at least one scripting or programming language (e.g., Bash, Python).
- Experience working with relational databases; exposure to NoSQL or data platforms is a plus.
- Experience participating in on-call rotations, responding to production incidents, and performing root cause analysis.
- Understanding of SLIs, SLOs, and error budgets, and how they are used to guide reliability and operational decisions.
- Strong problem-solving skills and the ability to debug complex production issues.
- Good verbal and written communication skills, especially during incidents and technical discussions.
- Experience operating systems at scale or in high-availability environments.
- Exposure to on-prem or hybrid infrastructure.
- Experience supporting data platforms, analytics, or AI/ML workloads.
- A strong sense of ownership and responsibility for production systems.
- A focus on automation, reliability, and operational simplicity.
- The ability to balance speed, stability, and long-term maintainability.
- Curiosity and willingness to continuously improve systems and processes.
Skills Required
- 3-5+ years managing and operating production infrastructure and services in cloud environments (AWS, Azure, or GCP)
- Strong hands-on experience with Linux systems in production environments
- Experience working with containerized workloads and Kubernetes
- Working knowledge of Infrastructure as Code tools such as Terraform, Terragrunt, or Crossplane
- Experience designing and maintaining CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, Azure DevOps, or similar
- Familiarity with GitOps principles and tools such as Argo CD or Flux
- Solid understanding of cloud networking concepts, load balancing, and service connectivity
- Experience with monitoring, logging, and alerting systems such as Prometheus, Grafana, ELK/EFK, Datadog, or equivalent
- Proficiency in at least one scripting or programming language (e.g., Bash, Python)
- Experience working with relational databases
- Exposure to NoSQL or data platforms
- Experience participating in on-call rotations, responding to production incidents, and performing root cause analysis
- Understanding of SLIs, SLOs, and error budgets
- Strong problem-solving skills and ability to debug complex production issues
- Good verbal and written communication skills
- Experience operating systems at scale or in high-availability environments
- Exposure to on-prem or hybrid infrastructure
- Experience supporting data platforms, analytics, or AI/ML workloads
What We Do
Signzy is a market-leading platform that is redefining the speed, accuracy, and experience of how financial institutions are onboarding customers and businesses - using the digital medium. The company’s award-winning no-code GO platform delivers seamless, end-to-end, and multi-channel onboarding journeys while offering totally customizable workflows. It gives these players access to an aggregated marketplace of 240+ bespoke APIs that can be easily added to any workflow with simple widgets. Signzy is enabling 10 million+ end customer and business onboardings every month at a success rate of 99% while reducing the speed to market from 6 months to 3-4 weeks. It works with over 240+ FIs globally including the 4 largest banks in India, a Top 3 acquiring Bank in the US, and has a strong global partnership with Mastercard and Microsoft. The company’s product team is based out of Bengaluru and it has a strong presence in Mumbai, New York, and Dubai.







