Senior Infrastructure Engineer

Posted 10 Days Ago
Be an Early Applicant
Bellevue, WA, USA
In-Office
Senior level
Artificial Intelligence • Edtech • Gaming • Software
The Role
Architects, builds, and operates AWS cloud infrastructure and developer platforms, including Kubernetes, infrastructure as code, CI/CD, GitOps, networking, DNS, observability, reliability, and cloud security. Leads incident response, capacity planning, disaster recovery, cost optimization, compliance tooling, and infrastructure projects. Establishes platform standards, enables engineering self-service, and mentors team members.
Summary Generated by Built In

About Level
Level is a learning technology company dedicated to helping students build real academic and life skills with confidence and joy. We combine proven curriculum principles with world class interactive design to make meaningful practice something students want to come back to, not something they struggle through. We support what teachers, schools, and parents are already doing by increasing student engagement with high quality, standards aligned practice that reinforces classroom learning.

As an Senior Infrastructure Engineer on the Platform team, you will architect, build, and operate the cloud infrastructure and developer platform that every Level product runs on. You will own critical infrastructure end-to-end — the Kubernetes platform, infrastructure-as-code, CI/CD and GitOps delivery, networking (ingress and egress), DNS, observability, and cloud security posture — and provide the reliable, self-service foundations the rest of engineering builds on. You will work on a small, senior-leaning team where infrastructure decisions have direct, visible impact on reliability, performance, cost, and developer velocity.

What You'll Do

  • Cloud Infrastructure & IaC — Design, build, and operate secure, highly available AWS infrastructure using Terraform/OpenTofu with a GitOps workflow (Atlantis). Own capacity planning, DR, and cost optimization for the systems you run.

  • Kubernetes & Platform Operations — Operate and evolve EKS: autoscaling (Karpenter), upgrades, core add-ons, and Helm-based delivery (ArgoCD).

  • CI/CD & Developer Enablement — Build and maintain GitHub Actions pipelines that let platform and product teams ship fast and safely, with self-service tooling where it makes sense.

  • Networking, Ingress & DNS — Own ingress/egress (Traefik), service mesh and mTLS (Linkerd/Envoy), load balancing, edge TLS, and DNS (Route 53, Terraform-managed).

  • Observability & Reliability — Build observability with OpenTelemetry and SigNoz; use telemetry to drive reliability, performance, and cost decisions. Serve as an escalation point for complex incidents, leading troubleshooting and post-mortems.

  • Security & Compliance — Apply cloud security best practices across identity, secrets, and network boundaries, with particular care for student data and K-12 privacy. Operate posture/vulnerability tooling (Security Hub, GuardDuty, Inspector, Snyk) and org guardrails (Control Tower, SCPs).

  • Ownership & Mentorship — Set standards, mentor engineers, and leave the platform better than you found it.


What You Need

  • 5+ years operating large-scale cloud infrastructure (AWS strongly preferred)

  • Deep IaC experience (Terraform/OpenTofu; CloudFormation/Pulumi/CDK also relevant)

  • Strong Docker/Kubernetes (EKS) production experience

  • Scripting/automation proficiency (Python, Go, or Bash)

  • Solid cloud networking fundamentals (VPC, DNS, load balancing, ingress, firewalls/WAF, VPNs) and security best practices

  • Proven CI/CD and GitOps experience (GitHub Actions or similar)

  • Observability experience (metrics/logs/traces) used to drive real decisions

  • Track record leading infrastructure projects independently, end to end

  • Strong communication across technical and non-technical audiences

Nice to Have

  • ArgoCD, Atlantis, Linkerd/Envoy, Traefik, Karpenter, Helm

  • OpenTelemetry, SigNoz (our stack), Grafana, Datadog, or Prometheus

  • Backstage or other internal developer platform experience

  • AI/ML infra experience (GPU scheduling, model/agent hosting, inference gateways)

  • Rust service CI/CD, CloudFront/CDN experience

  • AWS Solutions Architect / DevOps Engineer – Professional certification

  • Distributed-systems background, OSS infrastructure contributions

Skills Required

  • 5+ years operating large-scale cloud infrastructure, preferably AWS
  • Deep infrastructure-as-code experience with Terraform or OpenTofu
  • Strong production experience with Docker and Kubernetes, including EKS
  • Proficiency in Python, Go, or Bash scripting and automation
  • Strong cloud networking fundamentals, including VPC, DNS, load balancing, ingress, firewalls/WAF, and VPNs
  • Strong cloud security best-practices knowledge
  • Proven CI/CD and GitOps experience, such as GitHub Actions
  • Experience using metrics, logs, and traces for observability-driven decisions
  • Track record of independently leading infrastructure projects end to end
  • Strong communication across technical and non-technical audiences
  • Experience with ArgoCD, Atlantis, Linkerd or Envoy, Traefik, Karpenter, or Helm
  • Experience with OpenTelemetry, SigNoz, Grafana, Datadog, or Prometheus
  • Backstage or other internal developer platform experience
  • AI/ML infrastructure experience, including GPU scheduling, model or agent hosting, or inference gateways
  • Rust service CI/CD experience
  • CloudFront or CDN experience
  • AWS Solutions Architect or DevOps Engineer Professional certification
  • Distributed-systems background or open-source infrastructure contributions
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
218 Employees
Year Founded: 2018

What We Do

Level is an educational technology company that develops learning science products and academic content, focusing on ELA and Math. The company integrates game systems engineering, animation, and AI into its tools, employing a multidisciplinary team of designers, engineers, and content specialists to create its educational platform.

Similar Jobs

Airwallex Logo Airwallex

Senior Software Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
Hybrid
Seattle, WA, USA
2300 Employees
150K-245K Annually
Hybrid
2 Locations
289097 Employees

Shield AI Logo Shield AI

Infrastructure Engineer

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
In-Office
Seattle, WA, USA
110K-210K Annually
Hybrid
2 Locations
289097 Employees

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account