Sr. DevOps Engineer

Reposted 9 Days Ago
Be an Early Applicant
Portland, OR, USA
In-Office
75-85 Hourly
Senior level
Consulting
The Role
Own and improve release pipelines, deployment automation, fleet management, CI/CD, and observability for a distributed computer vision platform running on Linux edge systems. Build Ansible and Docker-based infrastructure, maintain Go update tooling, and develop LLM-powered testing and failure-analysis systems. Harden offline deployments with signed artifacts and deterministic builds, validate post-deployment health, and mentor engineers on release hygiene, reproducibility, and infrastructure-as-code.
Summary Generated by Built In
Position Overview
We're seeking an experienced senior development operations (DevOps) engineer to own the build, test, and deployment pipeline for our distributed computer vision platform running at edge sites across customer manufacturing floors. This is a full-time position working directly with our engineering team to harden our release process, drive deployment automation, and pioneer LLM-driven test generation and validation.
What You'll Do
  • Own and evolve the end-to-end release pipeline — branching strategy, build orchestration, artifact promotion, and rollback — across our Bazel monorepo and Python deployable units
  • Design and maintain Ansible-driven fleet automation for heterogeneous Linux edge nodes (Ubuntu LTS, NVIDIA driver stacks, Docker with NVIDIA runtime)
  • Manage all update tooling, currently written in Golang
  • Build LLM-powered automated testing systems: test generation from specs, flake triage, log/failure analysis, regression diffing, and release-note synthesis from commit and ticket history
  • Harden CI/CD for offline and bandwidth-constrained deployment targets (airgap wheel distribution, signed artifacts, deterministic builds)
  • Drive observability for releases — deployment telemetry, version drift detection, and post-deploy health validation across the fleet
  • Mentor engineers on release hygiene, reproducible builds, and infrastructure-as-code practices
What We're Looking For
  • 10+ years of professional experience in release engineering, DevOps, or SRE roles shipping production Linux systems
  • Deep curiosity for software, infrastructure, and applied AI — particularly using LLMs as production engineering tools, not just chat assistants
  • Expert-level Python (3.8+) with a strong grasp of packaging, dependency resolution, and PEP 440 versioning discipline
  • Demonstrated ownership of Linux fleets at scale — kernel, systemd, networking, package management
  • Excellence in technical communication, runbook authorship, and post-incident documentation
  • Strong systems thinking — comfortable reasoning about failure modes across hardware, OS, container, and application layers
Required Technical Skills
  • Expert proficiency with Ansible (roles, dynamic inventory, idempotent design); working knowledge of Terraform
  • Expert proficiency with Docker, including creation and lifecycle management of containers, image hardening, registry management and installing & configuring the NVIDIA container runtime
  • Production experience with Linux administration: systemd, networking (VLANs, DHCP, DNS), kernel/driver management (especially NVIDIA/DKMS), package and APT internals
  • Strong Python skills focused on tooling, automation, packaging (wheels, pip, private indexes), and subprocess/CI integration
  • Proficiency with Git workflows, branching strategies, and modern CI/CD systems (GitHub Actions, GitLab CI, or equivalent)
  • Experience designing and operating automated test infrastructure — unit, integration, hardware-in-the-loop, and end-to-end
  • Practical experience using LLMs (Anthropic, OpenAI, or local) as part of engineering workflows — test generation, code review augmentation, log analysis, or agentic tooling
Nice to Have
  • Bazel or similar monorepo build systems
  • Edge or embedded deployment experience
  • Tailscale, WireGuard, or zero-trust networking in production
  • gRPC/protobuf service ecosystems
  • Vault, PKI, or secrets management at fleet scale
  • Background in regulated or compliance-driven environments (CMMC, ISO 27001, SOC 2)
Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Range $75/hr – $85/hr.

Skills Required

  • 10+ years of professional experience in release engineering, DevOps, or SRE roles shipping production Linux systems
  • Expert-level Python 3.8+ skills, including packaging, dependency resolution, PEP 440 versioning, tooling, automation, wheels, pip, private indexes, subprocesses, and CI integration
  • Experience managing Linux fleets at scale, including kernel, systemd, networking, package management, and NVIDIA driver or DKMS management
  • Strong technical communication, runbook authorship, and post-incident documentation skills
  • Systems-thinking ability across hardware, operating system, container, and application layers
  • Expert proficiency with Ansible, including roles, dynamic inventory, and idempotent design
  • Working knowledge of Terraform
  • Expert proficiency with Docker, including container creation and lifecycle management, image hardening, registry management, and NVIDIA container runtime configuration
  • Production Linux administration experience, including systemd, VLANs, DHCP, DNS, kernel and driver management, package management, and APT internals
  • Proficiency with Git workflows, branching strategies, and modern CI/CD systems such as GitHub Actions or GitLab CI
  • Experience designing and operating automated unit, integration, hardware-in-the-loop, and end-to-end test infrastructure
  • Practical experience using LLMs such as Anthropic, OpenAI, or local models in engineering workflows
  • Bazel or similar monorepo build system experience
  • Edge or embedded deployment experience
  • Production experience with Tailscale, WireGuard, or zero-trust networking
  • Experience with gRPC and Protocol Buffers service ecosystems
  • Experience with Vault, PKI, or fleet-scale secrets management
  • Background in regulated or compliance-driven environments such as CMMC, ISO 27001, or SOC 2
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Bellevue, WA
251 Employees
Year Founded: 2007

What We Do

Fresh is an integrated consulting team of designers, developers, and engineers that build fresh experiences people love. Powered by our multi-talented teams, our winning combination of innovative thinking, creative design, sophisticated development, and advanced engineering ensure we’re delivering end-to-end experiences to help you grow. From creative brands and websites to intuitive apps and robots, we work with you from strategy to execution to create what's next.

Similar Jobs

In-Office or Remote
4 Locations
913 Employees

Hearst Magazines International Logo Hearst Magazines International

Senior Devops Engineer

Digital Media • Events • News + Entertainment
In-Office or Remote
2 Locations
10008 Employees
133K-161K Annually

Keeper Security, Inc. Logo Keeper Security, Inc.

Test Engineer

Mobile • Security • Software • Cybersecurity
Remote or Hybrid
US
350 Employees
In-Office
11 Locations

Similar Companies Hiring

Energy CX Thumbnail
Greentech • Professional Services • Business Intelligence • Consulting • Energy • Financial Services • Utilities
Chicago, IL
108 Employees
Quantum Rise Thumbnail
Software • Professional Services • Natural Language Processing • Machine Learning • Consulting • Automation • Artificial Intelligence
Chicago, Illinois
20 Employees
Northslope Thumbnail
Artificial Intelligence • Information Technology • Software • Analytics • Consulting • Generative AI
London, GB
100 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account