Systems Development Engineer (SRE/DevOps)

Posted 12 Hours Ago
Be an Early Applicant
Hyderabad, Telangana, IND
Hybrid
Mid level
Information Technology • Business Intelligence
The Role
Build and maintain highly available, observable cloud-native systems. Implement CI/CD pipelines, IaC, and monitoring; operate AWS and Kubernetes; lead incident response, automation, and reliability improvements.
Summary Generated by Built In
As a Systems Development Engineer with an SRE focus, you’ll build and maintain highly available, observable, and efficient systems that support mission‑critical services. You will focus on CI/CD automation, infrastructure and configuration as code, and deep monitoring/observability, while owning incident response and reliability for cloud‑native applications on AWS and Kubernetes.

Job Responsibilities:

    •  Design, build, and maintain automated CI/CD pipelines using tools such as Harness, GitHub Actions, and ArgoCD.
    • Develop and maintain infrastructure and configuration as code using CloudFormation, Terraform, Ansible, and related automation tools.
    • Administer and optimize AWS environments, including core services, networking, security, and architecture for availability, performance, and cost.
    Manage and support Kubernetes clusters and containerized workloads, including configuration, scaling, and upgrades.
    • Design, implement, and evolve end‑to‑end monitoring and observability frameworks using tools such as Open Telemetry, Groundcover, CloudWatch, Datadog, Prometheus, New Relic, or similar platforms.
    • Create and maintain dashboards, logs, traces, SLIs/SLOs, and automated alerting systems to ensure reliability and rapid detection of anomalies.
    • Embed observability, CI/CD best practices, and operational readiness into all stages of the software development lifecycle in partnership with engineering teams.
    • Lead or participate in incident response, troubleshooting, and root cause analysis for production incidents, using observability data to drive fast resolution.
    • Automate operational tasks, runbooks, and incident remediation workflows to reduce toil and improve service reliability.
    • Contribute to risk mitigation, backup, and disaster recovery strategies, including periodic testing and continuous improvement.
    • Participate in shared after hours support and project work as needed.

Job Qualification:

    • 2-4 years of experience designing, implementing, and maintaining CI/CD pipelines (e.g., Harness, GitHub Actions, ArgoCD or similar tools).
    • Hands‑on experience with automation tools and Infrastructure as Code / Configuration as Code (CloudFormation, Terraform, Ansible).
    • Strong understanding of Infrastructure as Code and Configuration as Code principles and patterns.
    • Solid grasp of the software development lifecycle and modern SRE/DevOps practices.
    AWS administration and architecture experience, including networking, security, IAM, and core services.
    • Experience operating Kubernetes clusters (EKS or other distributions) and containerized workloads.
    • Deep experience with monitoring and observability tools such as OpenTelemetry, Groundcover, CloudWatch, Datadog, Prometheus, New Relic, or equivalent, including metrics, logs, and traces.
    • Ability to define and track SLIs/SLOs and use them to guide reliability improvements.
    • Proficiency in Linux administration, including system configuration, troubleshooting, and performance tuning.
    • Programming/scripting skills in at least one language such as Python, Go, or Rust for automation, tooling, and observability integrations.
    • Solid understanding of networking, load balancing, and performance tuning.
    • Experience troubleshooting complex distributed systems, supporting incident response, and driving root cause analysis.
    • Familiarity with risk mitigation, backup, and disaster recovery concepts.

    Preferred

    • Experience building unified observability platforms or standardized dashboards for multiple services/teams.
    • Experience with GitOps workflows and tools for declarative infrastructure and application delivery.
    • Background in incident command and post‑mortem frameworks.
    • Experience integrating observability and reliability practices into microservices and/or serverless architectures.
    • Experience integrating testing, security and compliance checks into CI/CD pipelines.

About Model N  
Model N is the leader in revenue optimization and compliance for pharmaceutical, medtech and high-tech innovators. For more than 25 years, we have helped customers maximize revenue, streamline operations, and maintain compliance through cloud-based software, value-add services, and data-driven insights. With a focus on innovation and customer success, Model N empowers life sciences and high-tech manufacturers to bring life-changing products to the world more efficiently and profitably. Model N is trusted by over 150 of the world’s leading companies across more than 120 countries. For more information, visit www.modeln.com.
 

Skills Required

  • 2-4 years designing, implementing, and maintaining CI/CD pipelines (Harness, GitHub Actions, ArgoCD or similar)
  • Hands-on experience with CloudFormation, Terraform, Ansible (Infrastructure as Code / Configuration as Code)
  • Strong understanding of Infrastructure as Code and Configuration as Code principles
  • AWS administration and architecture experience, including networking, security, IAM, and core services
  • Experience operating Kubernetes clusters (EKS or other distributions) and containerized workloads
  • Deep experience with monitoring and observability tools (OpenTelemetry, Groundcover, CloudWatch, Datadog, Prometheus, New Relic or equivalent)
  • Ability to define and track SLIs/SLOs and use them to guide reliability improvements
  • Proficiency in Linux administration, system configuration, troubleshooting, and performance tuning
  • Programming/scripting skills in at least one language such as Python, Go, or Rust for automation and tooling
  • Solid understanding of networking, load balancing, and performance tuning
  • Experience troubleshooting complex distributed systems, incident response, and root cause analysis
  • Familiarity with risk mitigation, backup, and disaster recovery concepts
  • Participation in shared after-hours support and on-call rotations as needed
  • Experience building unified observability platforms or standardized dashboards for multiple services/teams
  • Experience with GitOps workflows and tools for declarative infrastructure and application delivery
  • Background in incident command and post-mortem frameworks
  • Experience integrating observability and reliability practices into microservices and/or serverless architectures
  • Experience integrating testing, security, and compliance checks into CI/CD pipelines
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Mateo, CA
1,302 Employees
Year Founded: 1999

What We Do

Model N enables life sciences and high tech companies to drive growth and market share, minimizing revenue leakage throughout the revenue lifecycle. With deep industry expertise and solutions purpose-built for these industries, Model N delivers comprehensive visibility, insight and control over the complexities of commercial operations and compliance. Our integrated cloud solution is proven to automate pricing, incentive and contract decisions to scale business profitably and grow revenue. Model N is trusted across more than 120 countries by the world’s leading pharmaceutical, medical technology, semiconductor, and high tech companies, including Johnson & Johnson, AstraZeneca, Stryker, Seagate Technology, Broadcom and Microchip Technology. For more information, visit www.modeln.com.

Similar Jobs

CrowdStrike Logo CrowdStrike

Artificial Intelligence Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
India
11000 Employees

CrowdStrike Logo CrowdStrike

Account Manager

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
India
11000 Employees

CrowdStrike Logo CrowdStrike

Workday support analyst, 2PM -11PM IST (Remote, IND)

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
India
11000 Employees

CrowdStrike Logo CrowdStrike

Site Reliability Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
India
11000 Employees

Similar Companies Hiring

Amplify Platform Thumbnail
Fintech • Financial Services • Consulting • Cloud • Business Intelligence • Big Data Analytics
Scottsdale, AZ
62 Employees
Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account