Senior Infrastructure Engineer

Posted Yesterday
Be an Early Applicant
Seattle, WA, USA
In-Office
130K-200K Annually
Senior level
Artificial Intelligence • Software • Generative AI
The Role
Lead design and operation of scalable, secure infrastructure for an AI-driven platform. Own Kubernetes clusters, CI/CD, observability, IaC, and real-time compute services. Drive automation, reliability best practices, disaster recovery, and long-term infrastructure strategy to support rapid product iteration and growth.
Summary Generated by Built In

Gradial is the marketing operations system of work that helps marketers and creatives move from idea to execution faster. Our platform orchestrates across martech stacks, workflows, and people to automate marketing execution, cutting execution time, so marketers focus on the work only humans can do: shaping brands, understanding customers, and creating the work that moves people.

Backed by leading investors, we’re building software that adapts to the user, not the other way around. We move with urgency, operate with ownership, and solve hard problems from first principles. If you want to do ambitious work, take real responsibility, and help define the future of AI-native content operations, you’ll do your best work here.

The Role 

As a Senior Infrastructure Engineer at Gradial, you will architect and evolve the systems that power our AI-driven content operations platform. This role is ideal for someone who thrives in startup-to-scale up environments and brings a deep understanding of how to make infrastructure reliable, secure, and scalable.

You’ll play a critical role in building on our core systems, supporting rapid product iteration and ensuring the platform is built for growth. If you’ve owned infrastructure in production, guided system evolution, and want to shape the future AI, we’d love to meet you.

What You'll Own

  • Design and maintain scalable, secure, and resilient infrastructure to support Gradial’s AI platform.
  • Lead Kubernetes cluster management, CI/CD pipelines, observability tooling, and infrastructure-as-code efforts.
  • Anticipate scaling needs and proactively evolve infrastructure architecture to support growth and reliability.
  • Take full ownership of real-time, compute-intensive services: designing, deploying and maintaining to meet high performance standards with minimal oversight.
  • Establish and enforce best practices for system reliability, performance monitoring, and disaster recovery.
  • Evaluate and implement infrastructure automation tools to improve deployment velocity and reduce operational burden.
  • Act as a strategic voice on infrastructure investment, technical debt management, and long-term scalability planning.

What We're Looking For

  • 5+ years of experience in DevOps, SRE or platform engineering roles.
  • Proven track record designing and operating large-scale, production-grade infrastructure.
  • Deep expertise in Kubernetes, cloud-native architecture, and container orchestration.
  • Proficiency with infrastructure-as-code (e.g., Terraform, GitOps ), CI/CD tooling, and monitoring stacks (e.g., Prometheus, Grafana).
  • Experience in high-growth environments, especially scaling infrastructure from early product-market fit to maturity.
  • Strong communication skills and a collaborative, ownership-driven mindset.

Nice to Have

  • Familiarity with AI/ML infrastructure, including GPU provisioning and model deployment.
  • Prior experience supporting cloud or multi-cloud architectures.
  • Comfort with TypeScript or Python to support tooling and operational scripts.

Compensation

The salary range for this position is $130,000 – $200,000 annually. Final compensation will be determined based on factors such as experience, skills, and qualifications. In addition to base salary, this role may be eligible for performance-based bonuses and equity awards. Gradial offers a comprehensive benefits package, including medical, dental & vision insurance, 401K retirement plan, paid time off, paid sick leave and other employee wellness programs. 

What we offer
  • Competitive salary and meaningful equity
  • Comprehensive health, dental and vision coverage
  • Fast-paced environment with flexibility and ownership
  • Real impact, zero bureaucracy
  • A front-row seat to building category-defining AI infrastructure
You'll thrive here if you...
  • Learn quickly, actively seek out new challenges, and regularly reconsider “how it’s always been done.”
  • Have an innate drive for being 1% better than the day before, building towards greatness.
  • Embrace AI as a core tool for problem-solving, innovation, and scale.
  • Show customer-obsession (internal or external), high ownership/accountability, and bias for action.
  • Communicate clearly, directly, with curiosity, and assuming good intentions.
  • Thrive in fast-paced, hyper-growth environments where building better > maintaining status quo.
AI Literacy & Interviewing Tools

As an AI-first company, we prioritize AI literacy as a core competency in our hiring decisions. We’re excited by candidates who thoughtfully apply AI tools in their work, but during interviews we’re focused on you. This is your opportunity to show how you think, communicate, and solve problems. Over-reliance on AI-generated responses during the interview process (especially when it obscures your own voice) will result in disqualification. We want to understand your unique perspective and how you approach challenges, both with and without AI.

Gradial is dedicated to creating an environment where diverse perspectives are valued and all team members can grow. We offer competitive compensation, equity, flexible work hours, comprehensive benefits, and a collaborative culture focused on learning and impact.

Skills Required

  • 5+ years of experience in DevOps, SRE or platform engineering roles.
  • Proven track record designing and operating large-scale, production-grade infrastructure.
  • Deep expertise in Kubernetes, cloud-native architecture, and container orchestration.
  • Proficiency with infrastructure-as-code (e.g., Terraform, GitOps), CI/CD tooling, and monitoring stacks (e.g., Prometheus, Grafana).
  • Experience in high-growth environments, especially scaling infrastructure from early product-market fit to maturity.
  • Strong communication skills and a collaborative, ownership-driven mindset.
  • Familiarity with AI/ML infrastructure, including GPU provisioning and model deployment.
  • Prior experience supporting cloud or multi-cloud architectures.
  • Comfort with TypeScript or Python to support tooling and operational scripts.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Seattle, WA
6 Employees
Year Founded: 2023

What We Do

Gradial is the enterprise generative AI platform for transforming content management and customer journey intelligence. Gradial leverages frontier generative image and text models to enable personalized digital experiences at scale. We empower creative teams to bring their vision to life using words and truly differentiate their brand's voice in a sea of content.

Similar Jobs

CoreWeave Logo CoreWeave

Senior Software Engineer

Cloud • Information Technology • Machine Learning
In-Office
4 Locations
1450 Employees
182K-242K Annually

DigitalOcean Logo DigitalOcean

Senior Software Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
In-Office
Seattle, WA, USA
1400 Employees
184K-231K Annually

CrowdStrike Logo CrowdStrike

Infrastructure Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
USA
11000 Employees
140K-215K Annually

Microsoft Logo Microsoft

Senior Software Engineer

Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
In-Office
Redmond, WA, USA
206870 Employees
120K-261K Annually

Similar Companies Hiring

Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
LTX Thumbnail
Robotics • Conversational AI • Generative AI
Jerusalem, Israel
300 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account