The Role
Manage application migrations across environments and maintain production and staging reliability. Responsibilities include incident response, on-call support, root cause analysis, observability improvements, CI/CD and GitOps support, Kubernetes infrastructure management, containerization, infrastructure as code, automation, monitoring, and production release support. Weekend availability is required for migrations.
Summary Generated by Built In
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.
WHY JOIN US
If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!
ABOUT THE ROLE
We are looking for a DevOps Engineer to manage application migrations between environments and maintain production and staging system reliability. This person handles incident response and on-call support, while improving observability, automation, and Kubernetes-based infrastructure. Weekend availability for migrations and comfort with CI/CD pipelines and SLAs are essential.
WHAT YOU WILL DO
- Migrate applications between environments; these migrations typically take place on weekends, so weekend availability is required.
- Monitor and support production and staging environments in real time, ensuring high availability, performance, and stability.
- Respond to incidents, perform triage and root cause analysis, and contribute to post-incident reviews and remediation efforts.
- Participate in an on-call rotation with defined SLAs.
- Handle ad-hoc and unplanned operational requests from Product, Support, and internal teams.
- Maintain and enhance monitoring, alerting, dashboards, logs, and metrics; improve signal-to-noise ratio and standardize observability practices.
- Support CI/CD pipelines, production releases, and GitOps workflows.
- Contribute to automation efforts to reduce operational toil.
- Maintain and improve Kubernetes-based infrastructure and containerized workloads.
- Support Infrastructure as Code practices and ongoing environment improvements.
MUST HAVES
- At least 2 years of experience in Site Reliability Engineering, DevOps, or Production Operations.
- Demonstrable AWS experience supporting production environments.
- Experience supporting production SaaS applications.
- Strong understanding of CI/CD systems (GitHub Actions, Jenkins, CircleCI, or similar).
- GitOps experience and strong Git fundamentals.
- Experience using GitHub, Jira, and Confluence in collaborative engineering environments.
- Kubernetes experience (EKS, kOps, or similar).
- Docker/containerization experience.
- Observability stack experience (Grafana, Prometheus, Loki, PagerDuty, or similar).
- Scripting experience (Bash, Python, or Go).
- Infrastructure as Code experience (Terraform, Helm, or similar).
- Working knowledge of relational databases (e.g., PostgreSQL, MySQL) for troubleshooting and operational support.
- Comfortable working within structured operational processes and SLAs.
- Strong written and verbal English communication skills; able to clearly explain technical concepts.
- Self-driven with a growth mindset.
- Weekend availability for application migrations between environments when needed.
NICE TO HAVES
- AWS certifications (Solutions Architect, DevOps Engineer, SysOps Administrator, etc.).
- Experience in multi-tenant SaaS environments.
- Experience working in globally distributed teams.
- Familiarity with ChatOps practices.
- Experience improving monitoring quality and reducing alert fatigue.
- Background in operational cost optimization.
PERKS AND BENEFITS
- Professional growth: Accelerate your professional journey with mentorship, TechTalks, and personalized growth roadmaps.
- Competitive compensation: We match your ever-growing skills, talent, and contributions with competitive USD-based compensation and budgets for education, fitness, and team activities.
- A selection of exciting projects: Join projects with modern solutions development and top-tier clients that include Fortune 500 enterprises and leading product brands.
- Flextime: Tailor your schedule for an optimal work-life balance, by having the options of working from home and going to the office – whatever makes you the happiest and most productive.
Meet Our Recruitment Process
Application → Coding Challenge → Video Interview → Technical Interview or Hiring Manager Interview
Each step helps us understand your skills and overall fit.
If it’s a match, you’ll receive an offer.
Skills Required
- At least 2 years of experience in Site Reliability Engineering, DevOps, or Production Operations
- AWS experience supporting production environments
- Experience supporting production SaaS applications
- Strong understanding of CI/CD systems such as GitHub Actions, Jenkins, or CircleCI
- GitOps experience and strong Git fundamentals
- Experience using GitHub, Jira, and Confluence
- Kubernetes experience, including EKS, kOps, or similar
- Docker and containerization experience
- Experience with observability tools such as Grafana, Prometheus, Loki, or PagerDuty
- Scripting experience with Bash, Python, or Go
- Infrastructure as Code experience with Terraform, Helm, or similar
- Working knowledge of relational databases such as PostgreSQL or MySQL
- Comfort working within structured operational processes and SLAs
- Strong written and verbal English communication skills
- Self-driven attitude and growth mindset
- Weekend availability for application migrations
- AWS certification such as Solutions Architect, DevOps Engineer, or SysOps Administrator
- Experience in multi-tenant SaaS environments
- Experience working in globally distributed teams
- Familiarity with ChatOps practices
- Experience improving monitoring quality and reducing alert fatigue
- Background in operational cost optimization
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
AgileEngine is a privately held company established in 2010 that builds dedicated teams of designers and developers. We turn good ideas into awesome software that people actually want to use. Some of the biggest names and the hottest startups around the world chose us to build their tech.








