Senior Cloud Infrastructure Engineer

Posted 7 Days Ago
Be an Early Applicant
Hyderabad, Telangana, IND
In-Office
Senior level
Healthtech
HealthEdge is on a mission to drive a digital revolution in healthcare.
The Role
Owns the design, operation, resilience, security, and migration of HealthEdge’s hybrid AWS, Azure, GCP, and on-premises infrastructure. Responsibilities include Infrastructure as Code, disaster recovery, Kubernetes/EKS, IAM, server administration, CI/CD automation, monitoring, incident response, vulnerability management, compliance, cost governance, and technical documentation. This senior individual contributor serves as a primary on-call and L2 escalation resource.
Summary Generated by Built In
Overview

Senior Cloud Infrastructure Engineer 


We're looking for a Senior Cloud Infrastructure Engineer to own the design, resilience, and day-to-day health of HealthEdge's infrastructure across AWS and our hybrid on-prem/Azure/GCP estate. This role sits at the intersection of cloud engineering and infrastructure engineering spanning AWS migration execution, disaster recovery, patching and platform currency, and the security/compliance controls that keep our security obligations intact. It's a hands-on senior IC role for someone who wants deep ownership of infrastructure resilience across a large, multi-account, multi-platform environment that's actively migrating off legacy on-prem infrastructure. 

Areas of Responsibility: 

Infrastructure and Hybrid Cloud Architecture


  • Own infrastructure across the environments including secure configuration baselines and patch management for OS images and on-prem hardware
  • Design, build, and maintain infrastructure across AWS (primary), with supporting work in Azure and GCP, plus the on-prem estate still in active retirement 
  • Build and maintain reusable, auditable Infrastructure as Code for cloud deployments, and for remaining on-prem server deployment. 
  • Support cloud networking execution, VPC provisioning, security group standards, and related connectivity work. 
  • Manage storage across cloud and legacy on-prem storage as workloads migrate; contribute to on-prem retirement and datacenter decommissioning efforts. 
  • Support AWS migration execution for in-flight waves, including server deployments and resource change requests via IaC. 

Disaster Recovery & Resilience 


  • Own disaster recovery architecture and execution across cloud environments. 
  • Maintain DR solutions and backup strategy and run DR drills on a regular cadence; document gaps and drive remediation. 
  • Design for resilience from the start and treat recoverability as a first-class requirement, not an afterthought. 

Security, Compliance & Vulnerability Management 


  • Contribute to vulnerability management triage across infrastructure teams, threat detection, and infra security findings review. 
  • Support PHI/PII classification scanning, penetration test coordination, and security exception approvals. 
  • Maintain compliance controls; support HIPAA and SOC 2 audit readiness, access review and recertification, evidence collection, and change freeze coordination. 
  • Maintain EKS container runtime security sensor coverage as part of ongoing platform hardening. 

Cloud Infrastructure Operations 


  • Design and manage roles and cross-account access controls following least-privilege principles across multi-account, multi-cloud environments. 
  • Own cloud execution: load balancers, DNS (), VPC provisioning, and security group standards. 
  • Manage compute resources at scale with an eye toward right-sizing and long-term maintainability. 
  • Administer cloud storage and database services with attention to cost, performance, and resilience. 
  • Own infrastructure health, cost, and performance monitoring using native and third-party tooling, building the observability that lets issues surface before they become incidents. 
  • Administer and harden Linux and Windows Server environments across cloud and on-prem, including patching, performance tuning, troubleshooting, Active Directory integration, Group Policy, DNS, and certificate services. 
  • Manage hybrid identity and authentication across on-prem and cloud workloads, and maintain OS-level security baselines and hardening standards across the estate. 

CI/CD, Automation & Delivery 


  • Build and evolve CI/CD pipelines for secure, repeatable infrastructure deployments. 
  • Write automation to reduce manual toil and enforce operational consistency across cloud and on-prem environments. 
  • Take solutions from proof-of-concept to production with an eye toward long-term maintainability, not just getting it working once. 

Reliability, Monitoring & Incident Response 


  • Monitor, scale, and maintain production infrastructure with availability, performance, and security as top priorities. 
  • Participate in on-call rotation; serve as L2 escalation point for cross-team infrastructure support; lead root cause analysis and drive incident retrospectives to closure. 
  • Author and maintain runbooks that hold up under pressure, not just at handoff. 

Cost & Tagging Governance 


  • Contribute to FinOps efforts: identify and remediate cost anomalies, own tagging remediation against enterprise tagging standards, and make pragmatic cost/performance/resilience tradeoffs. 

AI-Enabled Engineering 


  • Use AI coding assistants to accelerate IaC development, scripting, and troubleshooting. 
  • Use AI tooling to draft first-pass runbooks, DR documentation, and incident retrospectives to validate and refine before publishing. 

Collaboration & Documentation 


  • Document architecture, DR runbooks, and standard operating procedures others can actually follow under pressure. 
  • Provide technical guidance to product teams on infrastructure resilience, migration sequencing, and recovery design. 

Required Qualifications 


  • 5+ years of hands-on cloud infrastructure engineering experience, with deep expertise in AWS and in other cloud environments. 
  • Direct experience with disaster recovery design and execution. 
  • Strong Infrastructure as Code experience (CDK, Terraform, or CloudFormation). 
  • Experience with containerized environments and Kubernetes/EKS, including version upgrade and lifecycle management. 
  • Linux and Windows Server administration experience, including patching and OS lifecycle management at scale. 
  • Solid IAM design experience, including cross-account access and least-privilege enforcement. 
  • Strong scripting ability (Python, Bash, or PowerShell). 
  • Experience building and maintaining CI/CD pipelines. 
  • Comfortable being the primary on-call and L2 escalation point for infrastructure incidents. 

Preferred Qualifications 


  • Experience operating in regulated environments (FedRAMP, HIPAA, SOC 2) and understanding of what that means for infrastructure and DR design specifically. 
  • AWS certification (Solutions Architect or SysOps, Associate or Professional). 
  • Experience with hybrid infrastructure, bridging on-prem virtualization with cloud-native services during active migration. 
  • Familiarity with DISA STIG or CIS benchmark hardening, and vulnerability management/triage workflows. 
  • FinOps or cost governance experience, including tagging standards enforcement. 
  • Healthcare technology or digital health platform background. 
  • Experience with AI-assisted engineering workflows as part of daily practice. 

Behaviors & Traits 


  • Raises risk early rather than waiting for it to become an incident. 
  • Comfortable with ambiguity in a large, multi-account, evolving cloud environment. 
  • Strong sense of ownership; closes gaps rather than escalating and waiting. 
  • Communicates technical tradeoffs clearly to both engineers and non-technical stakeholders. 

Skills Required

  • 5+ years of hands-on cloud infrastructure engineering experience, including deep AWS expertise and experience with other cloud environments
  • Direct experience designing and executing disaster recovery
  • Strong Infrastructure as Code experience with CDK, Terraform, or CloudFormation
  • Experience with containerized environments and Kubernetes/EKS, including version upgrades and lifecycle management
  • Linux and Windows Server administration experience, including patching and OS lifecycle management at scale
  • IAM design experience, including cross-account access and least-privilege enforcement
  • Strong scripting ability in Python, Bash, or PowerShell
  • Experience building and maintaining CI/CD pipelines
  • Ability to serve as the primary on-call and L2 escalation point for infrastructure incidents
  • Experience operating in regulated environments such as FedRAMP, HIPAA, or SOC 2
  • AWS Solutions Architect or SysOps certification at the Associate or Professional level
  • Experience with hybrid infrastructure and active on-premises-to-cloud migrations
  • Familiarity with DISA STIG or CIS benchmark hardening and vulnerability management workflows
  • FinOps or cost governance experience, including tagging standards enforcement
  • Healthcare technology or digital health platform experience
  • Experience using AI-assisted engineering workflows

HealthEdge Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about HealthEdge and has not been reviewed or approved by HealthEdge.

  • Leave & Time Off Breadth Time off is positioned as relatively generous, with a large holiday calendar, vacation to start, unlimited sick time, and volunteer days. The overall package also includes flexibility elements that can increase the practical value of time off depending on role.
  • Retirement Support Retirement support stands out through a 401(k) match with immediate vesting, which strengthens the near-term value of the benefit. HSA/FSA options and employer contributions are also highlighted as part of the financial benefits mix.
  • Inclusive Benefits Coverage Medical coverage is described as inclusive, explicitly including infertility treatments and gender-affirming care alongside EAP and mental-health services. This breadth can improve perceived total rewards for employees with varied healthcare needs.

HealthEdge Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Boston, MA
1,600 Employees
Year Founded: 2004

What We Do

HealthEdge is on a mission to drive a digital revolution in healthcare. We’re connecting health plans, providers, and patients with end-to-end digital technology solutions to support new business models, reduce administrative costs and improve health outcomes. Our growing portfolio of products (HealthRules Payor, Source, GuidingCare, and Wellframe) provides talented and passionate professionals with opportunities to lead change and make a lasting, global impact in healthcare. Driving our mission are 2,000+ professionals worldwide. Together, we are committed to innovating a world where healthcare can focus on people.

Gallery

Gallery

Similar Jobs

In-Office
Hyderabad, Telangana, IND
4405 Employees
In-Office
Hyderabad, Telangana, IND
4405 Employees

Cisco Logo Cisco

Infrastructure Engineer

Cloud • Information Technology • Internet of Things • Professional Services • Software
In-Office
2 Locations
77500 Employees

Similar Companies Hiring

Sailor Health Thumbnail
Healthtech • Social Impact • Telehealth
New York City, NY
20 Employees
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account