We are seeking an experienced AWS Platform Engineer / Site Reliability Engineer to join our client's platform team in Las Vegas, NV. This role blends release engineering, observability, and cloud infrastructure automation across a large AWS estate, with hands-on responsibility for EKS, networking, security, and monitoring tooling including Kibana, Dynatrace, and Apigee.
This is an onsite role, Monday through Friday, in Las Vegas, NV. Long-term engagement; FTE or contract.
WHAT YOU WILL DO
• Build and operate core AWS infrastructure: VPC, EC2, S3, IAM, Route 53, backups, EKS, SSO, MSK, Security Hub, GuardDuty, and related security/compliance tooling.
• Design and troubleshoot advanced cloud networking, including AWS Transit Gateway (TGW) and Direct Connect setup and hybrid connectivity.
• Own observability and monitoring across Amazon CloudWatch, Grafana, and OpenTelemetry (OTEL), including proactive alerting and telemetry configuration.
• Automate end to end with Terraform, GitLab CI/CD, and shell scripting.
• Operate Kubernetes (EKS) and Istio service mesh for traffic management, security, and observability.
• Drive patching and vulnerability remediation across cloud-native and containerized environments in line with compliance requirements.
• Develop Python tooling for automation, scripting, and infrastructure operations.
• Support release engineering and monitoring workflows across Kibana, Dynatrace, and Apigee.
REQUIRED QUALIFICATIONS
• Minimum 7+ years of relevant platform, SRE, or cloud infrastructure experience.
• Strong hands-on experience with core AWS services listed above.
• Advanced networking knowledge, including Transit Gateway and Direct Connect.
• Proficiency with Terraform, GitLab CI/CD, and shell scripting.
• Solid working knowledge of Kubernetes/EKS and Istio.
• Hands-on Python development for automation and infrastructure tooling.
• Experience with patching and vulnerability remediation in containerized environments.
• Problem-solving mindset and the ability to work effectively in fast-paced Agile/Scrum environments.
TECHNICAL SKILLS
AWS (VPC, EC2, S3, IAM, Route 53, EKS, SSO, MSK, Security Hub, GuardDuty), Kubernetes/EKS, Istio, Transit Gateway, Direct Connect, Terraform, GitLab CI/CD, Python, Shell, CloudWatch, Grafana, OpenTelemetry, Kibana, Dynatrace, Apigee
SKILLS MATRIX
Applicants will be asked to provide years of experience and a self-rating (1-10) for each of the following:
• AWS infrastructure
• Kubernetes / EKS
• Networking and security
• Terraform
• Python development for scripting, automation, and infrastructure tooling
• AWS services: VPC, EC2, S3, IAM, Route 53, backups, EKS, SSO, MSK, Security Hub, GuardDuty and related security/compliance tooling
• Release engineering and monitoring (Kibana, Dynatrace, Apigee)
Skills Required
- Minimum 7 years of relevant platform, SRE, or cloud infrastructure experience
- Hands-on experience with core AWS infrastructure and security services
- Advanced networking knowledge, including AWS Transit Gateway and Direct Connect
- Proficiency with Terraform, GitLab CI/CD, and shell scripting
- Working knowledge of Kubernetes/EKS and Istio
- Hands-on Python development for automation and infrastructure tooling
- Experience with patching and vulnerability remediation in containerized environments
- Problem-solving ability and effectiveness in fast-paced Agile/Scrum environments
What We Do
Ontrac Solutions helps organizations adopt emerging technologies to scale smarter. We build GenAI platforms, predictive analytics solutions, and drive cloud adoption. We're also a HubSpot partner, supporting landing page design, website development, CRM integration, workflows, and automation. From infrastructure to marketing ops, we deliver strategy and execution that drives growth.








