Principal Site Reliability Engineer

Posted 2 Days Ago
Be an Early Applicant
Boston, MA, USA
Hybrid
200K-250K Annually
Senior level
Consumer Web • Gaming • Mobile • News + Entertainment • Software
Jackpocket gives lottery fans an easy, secure way to order lottery tickets from their phone.
The Role
Leads the strategy, architecture, and operation of Kubernetes-based cloud and on-premise infrastructure. Drives reliability engineering, Infrastructure as Code, GitOps, observability, automation, incident response, capacity planning, and cost optimization. Leads cross-functional platform initiatives, establishes SLOs and error budgets, mentors engineers, and advances AI-enabled engineering practices across the organization.
Summary Generated by Built In

At DraftKings, AI is becoming an integral part of both our present and future, powering how work gets done today, guiding smarter decisions, and sparking bold ideas. It’s transforming how we enhance customer experiences, streamline operations, and unlock new possibilities. Our teams are energized by innovation and readily embrace emerging technology. We’re not waiting for the future to arrive. We’re shaping it, one bold step at a time. To those who see AI as a driver of progress, come build the future together.

The Crown Is Yours

As a Principal Site Reliability Engineer, you'll shape the long-term strategy for the infrastructure behind one of the most demanding platforms in sports betting and gaming. You'll drive the architectural direction of our cloud and on-premise platforms, helping engineering teams build, deploy, and operate highly reliable systems at scale. Working across Platform Engineering and Site Reliability Engineering, you'll influence how we modernize our infrastructure, strengthen operational excellence, and prepare our platform for the next generation of growth.

What you'll do
  • Define and execute the long-term strategy for our Kubernetes platform across Google Kubernetes Engine, Amazon Elastic Kubernetes Service, RKE2, and on-premise environments, ensuring reliability, scalability, and operational consistency.

  • Drive architectural decisions across critical infrastructure, including cluster lifecycle management, networking, identity and access management, observability, autoscaling, capacity planning, and cost optimization.

  • Lead large-scale platform initiatives across multiple engineering teams, establishing technical direction, engineering standards, and measurable outcomes that improve platform reliability and developer experience.

  • Establish and evolve reliability practices by defining service level objectives, service level indicators, and error budget frameworks that align platform performance with business priorities.

  • Build automation-first infrastructure through Infrastructure as Code, GitOps workflows, self-healing systems, and internal platform tooling that improve engineering velocity and reduce operational overhead.

  • Champion the responsible adoption of AI-powered engineering capabilities that improve operational efficiency, accelerate incident response, and enhance developer productivity.

  • Lead critical platform incidents, drive post-incident improvements, and strengthen platform resilience through automation, capacity planning, and operational excellence.

  • Mentor senior engineers, influence technical strategy across the organization, and elevate engineering excellence through architecture reviews, coaching, and technical leadership.

What you'll bring
  • A Bachelor's Degree in Computer Science or a related technical field.

  • At least 8 years of experience designing, operating, and scaling distributed cloud and on-premise infrastructure, including at least 3 years operating at the Staff, Principal, or equivalent technical leadership level.

  • Proven experience leading large-scale infrastructure or platform initiatives that require cross-functional alignment and long-term technical ownership.

  • Deep expertise with Kubernetes, including cluster architecture, networking, storage, security, operators, lifecycle management, and large-scale production operations.

  • Extensive experience building and operating production infrastructure in AWS and Google Cloud Platform using Infrastructure as Code technologies such as Terraform, Pulumi, or similar tools.

  • Strong software development experience in Go, Python, or both, with expertise in GitOps, continuous integration and continuous delivery, observability, distributed systems, Linux, and reliability engineering principles.

  • Experience incorporating AI-powered tools into engineering workflows while applying sound judgment around reliability, security, and operational risk.

  • Exceptional communication and leadership skills with a proven ability to mentor engineers, influence technical strategy, and drive engineering excellence. Experience working in regulated industries, hybrid cloud environments, contributing to open-source projects, or holding cloud certifications is preferred.

Join Our Team

We’re a publicly traded (NASDAQ: DKNG) technology company headquartered in Boston. As a regulated gaming company, you may be required to obtain a gaming license issued by the appropriate state agency as a condition of employment. Don’t worry, we’ll guide you through the process if this is relevant to your role.

The US base salary range for this full-time position is 200,000.00 USD - 250,000.00 USD, plus bonus, equity, and benefits as applicable. Our ranges are determined by role, level, and location. The compensation information displayed on each job posting reflects the range for new hire pay rates for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific pay range and how that was determined during the hiring process. It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Skills Required

  • Bachelor's degree in Computer Science or a related technical field
  • At least 8 years designing, operating, and scaling distributed cloud and on-premise infrastructure
  • At least 3 years operating at the Staff, Principal, or equivalent technical leadership level
  • Experience leading large-scale infrastructure or platform initiatives requiring cross-functional alignment and long-term technical ownership
  • Deep expertise with Kubernetes, including cluster architecture, networking, storage, security, operators, lifecycle management, and production operations
  • Experience building and operating production infrastructure in AWS and Google Cloud Platform
  • Experience with Infrastructure as Code technologies such as Terraform or Pulumi
  • Strong software development experience in Go, Python, or both
  • Expertise in GitOps, continuous integration and delivery, observability, distributed systems, Linux, and reliability engineering principles
  • Experience incorporating AI-powered tools into engineering workflows
  • Exceptional communication and leadership skills, including mentoring engineers and influencing technical strategy
  • Experience working in regulated industries, hybrid cloud environments, contributing to open-source projects, or holding cloud certifications
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
330 Employees
Year Founded: 2013

What We Do

Jackpocket is creating a more convenient, fun, and responsible way to take part in the lottery. The first licensed third-party lottery app in the United States, Jackpocket offers players a secure way to order official state lottery tickets, such as Powerball, Mega Millions, and more. Lottery players use Jackpocket to place ticket orders for their favorite games, check lottery results, join lottery pools with other Jackpocket players, and turn on Autoplay so they never miss a drawing. Jackpocket is available in Arizona, Arkansas, Colorado, Idaho, Maine, Massachusetts, Minnesota, Montana, Nebraska, New Hampshire, New Jersey, New Mexico, New York, Ohio, Oregon, Texas, Washington, D.C., and West Virginia with many new markets on the horizon.

Why Work With Us

Our dedicated team of superhumans (and an adorable office pup or two) are making waves in an $80 billion industry. We believe transparency, teamwork, and a healthy dose of daily fun are key to creating Worktopia. Jackpocketeers are supportive, forward-thinking, and truly enjoy spending time together. Come join us for the long haul!

Gallery

Gallery

Similar Jobs

SOPHiA GENETICS Logo SOPHiA GENETICS

Principal Software Engineer

Artificial Intelligence • Big Data • Healthtech • Software • Biotech
Hybrid
Boston, MA, USA
450 Employees
88K-168K Annually

Akamai Technologies Logo Akamai Technologies

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
169K-305K Annually

PTC Logo PTC

Principal Software Engineer

Information Technology • Internet of Things • Software • Virtual Reality
In-Office
Boston, MA, USA
7347 Employees
131K-200K Annually

DraftKings Logo DraftKings

Site Reliability Engineer

Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Hybrid
Boston, MA, USA
6400 Employees
200K-250K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account