Principal DevOps Engineer — Platform Infrastructure (AWS)

Posted 8 Days Ago
Be an Early Applicant
Hiring Remotely in Poland
Remote
Expert/Leader
Information Technology • Software
The Role
Design and build a greenfield AWS platform for a high-volume gaming and sportsbook system. Own multi-account cloud architecture, Kubernetes, Terraform, GitOps, CI/CD, observability, reliability, disaster recovery, compliance, migrations, and production operations. Establish scalable infrastructure for traffic spikes and regulated money movement, lead incident readiness and resilience testing, support AI-assisted development infrastructure, and mentor engineers while setting organization-wide platform standards.
Summary Generated by Built In
Description

We are looking for Principal DevOps Engineer to design and build the AWS infrastructure and delivery platform the new gaming platform will run on. Greenfield: define the cloud foundation — account structure, networking, Kubernetes platform, CI/CD, observability — before the first production workload lands, then scale it to sportsbook peak traffic for millions of players. The platform must hold under live-betting spikes, strict uptime expectations for a money-moving system, and gaming-regulator audit and compliance requirements. Hands-on technical leader: writes infrastructure code daily, sets platform standards, and is the technical authority on how software is built, shipped and operated.

What you will do:

  • Design the AWS foundation from scratch: multi-account architecture (AWS Organizations), landing zone, VPC and networking, IAM strategy, cost governance
  • Build and operate the container platform — Amazon EKS, service mesh, autoscaling tuned for spiky sportsbook load, multi-AZ (and where justified multi-region) resilience
  • Define everything as code: Terraform for all infrastructure, GitOps delivery (Argo CD or similar), paved-road CI/CD pipelines so product teams ship safely and often
  • Establish the observability stack — metrics, logging, tracing, alerting (Prometheus/Grafana, OpenTelemetry, CloudWatch) — and drive an SLO-based reliability practice with error budgets
  • Own production readiness: incident response, on-call design, runbooks, chaos/load testing ahead of major sporting events, blameless postmortems
  • Build the infrastructure side of the migration off the current third-party platform: dual-running environments, data migration pipelines, cutover mechanics
  • Embed compliance into the platform: audit trails, environment segregation, backup/DR, controls that satisfy gaming regulators by construction
  • Partner with the AI coding platform team: provision and operate the infrastructure behind AI-assisted development, and bring AI tooling into DevOps workflows
  • Mentor engineers across teams on cloud-native and operational best practices; set organization-wide standards
Requirements

Must have:

  • 10+ years of DevOps / platform / infrastructure engineering, including 3+ years at staff/principal level with organization-wide influence
  • Deep hands-on AWS: EKS, EC2, RDS/Aurora, networking (VPC, Transit Gateway, Route 53), IAM at scale, multi-account architectures (AWS certifications such as SA Professional / DevOps Professional are a plus)
  • Expert-level infrastructure as code with Terraform
  • Strong Kubernetes operational depth: day-2 operations, upgrades, capacity, cost
  • Track record of building CI/CD and developer platforms engineering teams adopted willingly — golden paths, not gatekeeping
  • Observability and SLO/error-budget practice (Prometheus/Grafana, OpenTelemetry, CloudWatch)
  • Experience operating high-availability, high-throughput production systems with real traffic spikes, with documented playbooks from real incidents
  • Strong scripting/programming (Python, Go or Bash) and comfort reading application code
  • Experience designing backup, disaster recovery and business continuity for systems where data loss is not an option
  • Excellent written and spoken English; communicates platform decisions clearly to engineers and executives

Nice to have:

  • Hands-on use of AI coding tools (Claude Code, Codex) for infrastructure, pipelines and ops automation — a significant plus
  • Security engineering: cloud security posture management, secrets management (e.g. Vault), vulnerability management, SAST/DAST/SCA in CI/CD, ISO 27001 / SOC 2 / PCI DSS
  • iGaming, sports betting, fintech or another regulated, high-transaction-volume domain
  • Migration off a third-party vendor platform
  • Event-streaming infrastructure (Kafka/MSK) and database operations at scale
  • Polish and/or Spanish

Skills Required

  • 10+ years of DevOps, platform, or infrastructure engineering experience
  • 3+ years at staff or principal level with organization-wide influence
  • Deep hands-on AWS experience, including EKS, EC2, RDS/Aurora, VPC, Transit Gateway, Route 53, IAM, and multi-account architectures
  • Expert-level Terraform infrastructure-as-code experience
  • Strong Kubernetes operational experience, including day-two operations, upgrades, capacity, and cost management
  • Experience building CI/CD and developer platforms adopted by engineering teams
  • Observability and SLO/error-budget experience with Prometheus, Grafana, OpenTelemetry, or CloudWatch
  • Experience operating highly available, high-throughput production systems with real traffic spikes
  • Strong scripting or programming skills in Python, Go, or Bash, with ability to read application code
  • Experience designing backup, disaster recovery, and business continuity for systems where data loss is unacceptable
  • Excellent written and spoken English
  • AWS Solutions Architect Professional or DevOps Engineer Professional certification
  • Hands-on use of AI coding tools such as Claude Code or Codex
  • Security engineering experience, including cloud security posture management, secrets management, vulnerability management, SAST, DAST, SCA, ISO 27001, SOC 2, or PCI DSS
  • Experience in iGaming, sports betting, fintech, or another regulated high-transaction-volume domain
  • Experience migrating off a third-party vendor platform
  • Event-streaming infrastructure and database operations at scale, including Kafka or MSK
  • Polish and/or Spanish language skills
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Petaling Jaya
399 Employees
Year Founded: 2005

What We Do

Commit is a global tech services company with offices in Israel, US, Canada, UK, and Europe. The company was founded in 2005 and has over 700 multi-disciplinary innovation experts who serve a broad range of companies, from small startups to large enterprises in multiple business sectors. Commit specializes in advanced technologies and applications with dedicated practices in Cloud, GenAI, Software, IoT, Big Data, Cyber, Collaboration, Data center migration projects, and more. Commit offers innovative, end-to-end technology solutions by developing custom software and IoT platforms for clients looking to build their next-gen products within the modern ICT world. Commit’s complete and comprehensive engineering powerhouse of resources, and proprietary Flexible R&D methodology helps transform its clients’ technology visions into high-quality products while reducing costs and improving time-to-market.

Similar Jobs

Zapier Logo Zapier

People Business Partner

Artificial Intelligence • Productivity • Software • Automation
Remote
27 Locations
800 Employees

OpenX Technologies Logo OpenX Technologies

Senior Software Engineer

AdTech • Enterprise Web • Information Technology • Machine Learning • Marketing Tech • Sales
Easy Apply
Remote or Hybrid
Kraków, Małopolskie, POL
420 Employees
181-25K Hourly

UL Solutions Logo UL Solutions

Field Engineer

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Warsaw, Warszawa, Mazowieckie, POL
15000 Employees
102K-135K Annually

GitLab Logo GitLab

Senior Back-end Engineer

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
Poland
2500 Employees
272K-408K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account