Senior Site Reliability Engineer — CDN Infrastructure

Posted 14 Hours Ago
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Senior level
Cloud • Information Technology • Software • Cybersecurity
The Role
Operate and evolve CDN infrastructure across bare-metal edge servers, Kubernetes control planes, CI/CD, observability, data pipelines, and cloud infrastructure. Responsibilities include configuration management, Helm deployments, GitHub Actions pipelines, Prometheus and Grafana monitoring, ClickHouse performance tuning, Terraform-based provisioning, and troubleshooting DNS, BGP, routing, TCP/IP, kernel, and application issues.
Summary Generated by Built In

About the Role

We're looking for a Senior SRE with 5+ years of experience to help operate and evolve

the infrastructure behind our CDN — from the baremetal edge nodes serving traffic, to

the Kubernetes-based control plane, to the observability and data pipelines that tell us

what's actually happening on the network. You'll work across the full stack: provisioning

and configuration management, CI/CD, Kubernetes, and the logging/monitoring systems

that let us catch problems before customers do.

This is a hands-on, senior individual-contributor role for someone who's comfortable

moving between low-level networking debugging and higher-level platform/infra

automation, and who can drive decisions with less oversight.

What You'll Do

Operate and maintain our fleet of self-managed baremetal edge servers, using

SaltStack (or Ansible) for configuration management and automation

Manage and improve our managed Kubernetes cluster, deploying and maintaining

services via Helm

Build and maintain CI/CD pipelines (GitHub Actions) for infrastructure and service

deployments

Own and extend observability tooling — Prometheus, Grafana, Alertmanager, and

distributed tracing with Grafana Tempo

Maintain and query ClickHouse at scale — millions of rows ingested daily, with a

focus on schema design and low-latency query tuning

Manage cloud service configuration as code using Terraform

Diagnose and resolve issues across the stack — from DNS resolution and

BGP/routing anomalies to TCP/IP-level performance regressions and applicationWhat

You'll Need



Requirements

What You'll Need

5+ years of experience in an SRE, DevOps, or infrastructure/platform engineering

role

Deep understanding of networking fundamentals — DNS, TCP/IP, and CDN concepts

(caching, routing, anycast, edge delivery) — you should be comfortable reading a

packet capture or debugging a DNS resolution chain

Solid experience operating Linux systems in production, including self-managed

baremetal infrastructure

Hands-on experience with configuration management tools — SaltStack or Ansible

Experience with Kubernetes in production, including deploying and managing

services via Helm

Experience building CI/CD pipelines, ideally with GitHub Actions

Working knowledge of Terraform or similar IaC tools

Practical experience with Prometheus, Grafana, and Alertmanager for monitoring and

alerting; familiarity with distributed tracing (Grafana Tempo or similar)

Ability to write code in Go or Python for automation, tooling, or internal services

Strong debugging skills across layers — from kernel/network to application to

infrastructure automation

Good communication skills and comfort working in a small, high-ownership team

Nice to Have

Experience with BGP / anycast routing in a production CDN or network operator

context

Experience with ClickHouse or another columnar/analytical database at scale

Experience with Kafka/Redpanda or similar streaming systems for log/data pipelines

Prior experience at a CDN, ISP, hosting provider, or similar network-heavy operator

Experience with GitOps workflows (ArgoCD or similar)



Skills Required

  • 5+ years of experience in SRE, DevOps, infrastructure engineering, or platform engineering
  • Deep understanding of DNS, TCP/IP, CDN concepts, caching, routing, anycast, and edge delivery
  • Production Linux administration, including self-managed bare-metal infrastructure
  • Hands-on experience with SaltStack or Ansible
  • Production Kubernetes experience, including deploying and managing services with Helm
  • Experience building CI/CD pipelines, ideally with GitHub Actions
  • Working knowledge of Terraform or similar infrastructure-as-code tools
  • Practical experience with Prometheus, Grafana, and Alertmanager
  • Familiarity with distributed tracing such as Grafana Tempo or similar tools
  • Ability to write Go or Python for automation, tooling, or internal services
  • Strong debugging skills across kernel, networking, applications, and infrastructure automation
  • Good communication skills and comfort working in a small, high-ownership team
  • Production experience with BGP or anycast routing in a CDN or network operator context
  • ClickHouse or another columnar or analytical database experience at scale
  • Kafka, Redpanda, or similar streaming systems experience
  • Experience at a CDN, ISP, hosting provider, or similar network-heavy operator
  • Experience with GitOps workflows such as ArgoCD
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
10 Employees
Year Founded: 2019

What We Do

VergeCloud is an India-focused cloud infrastructure platform providing a Content Delivery Network (CDN), DNS, cloud security, and edge computing services. Its solutions are designed to accelerate websites and applications, reduce latency, improve performance, and protect digital assets through capabilities such as Web Application Firewall (WAF), DDoS mitigation, firewall rules, SSL/TLS, and real-time monitoring. The company serves businesses seeking scalable, secure, and reliable digital experiences.

Similar Jobs

Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees
Hybrid
2 Locations
289097 Employees
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account