Senior DevOps Engineer, Infrastructure & Reliability

Posted 3 Hours Ago
Be an Early Applicant
4 Locations
In-Office or Remote
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Worth is the underwriting & onboarding platform that helps financial institutions say yes to small businesses, faster.
The Role
Build and operate scalable, reliable infrastructure using Terraform, Kubernetes, AWS, and CI/CD automation. Responsibilities include managing EKS platforms, improving deployment pipelines, securing networking and IAM, enhancing observability, optimizing cloud costs, implementing disaster recovery, reducing operational toil, and modernizing legacy infrastructure. The role also leads incident response, reliability initiatives, platform adoption, and cross-team infrastructure projects.
Summary Generated by Built In

Worth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier for engineers shipping software. 

This is a hands-on build role. You will spend most of your time writing Terraform, tuning Kubernetes workloads, automating things that are currently manual, and shipping infrastructure changes to production. You'll join a small platform team with an established roadmap and existing patterns, and a strong voice in how the work gets built.

  • Implement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift.
  • Own and evolve our Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default.
  • Optimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases.
  • Design and enforce secure networking, IAM, and secrets management strategies across environments.
  • Improve observability by refining metrics, logs, and tracing using tools like DataDog, ensuring actionable insight into system health.
  • Optimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements.
  • Implement disaster recovery planning, backup strategies, and multi-region resilience initiatives.
  • Refactor brittle or manually managed infrastructure into automated, testable, and reproducible systems.
  • Introduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support.
  • Partner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments.
  • Communicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety.
Technology Stack
  • Cloud & Infrastructure: AWS (EKS, RDS, MSK, S3, Lambda, IAM, VPC)
    Containerization & Orchestration: Kubernetes, ArgoCD
    Infrastructure-as-Code: Terraform
    CI/CD: GitHub Actions
    Monitoring & Observability: DataDog
    Data & Messaging: PostgreSQL, Kafka, Redis
    Languages (as needed): Bash, Python, TypeScript, JavaScript

Requirements
  • 8+ years in DevOps, SRE, or infrastructure engineering.
  • Proven experience designing and operating production Kubernetes environments at scale.
  • Deep hands-on expertise with AWS infrastructure and cloud networking.
  • Strong experience building and maintaining Terraform modules across large cloud environments.
  • Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.
  • Experience leading incident response processes and driving meaningful postmortem outcomes.
  • Strong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).
  • Proven ability to modernize legacy infrastructure and eliminate manual operational toil.
  • Track record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction.
  • Demonstrated ability to build trust across teams while raising the reliability bar.
Success Metrics
  • System Reliability: Maintain or exceed defined SLO/SLA targets with reduced incident frequency and duration.
  • Infrastructure Stability: Reduce production incidents caused by misconfiguration, manual processes, or infrastructure drift.
  • Operational Efficiency: Increase the percentage of infrastructure managed through code and automation.
  • Cost Optimization: Improve cloud cost efficiency without sacrificing reliability or performance.
Bonus Points (Nice to Have)
  • Experience coding applications
  • Experience operating high-throughput Kafka clusters (MSK or self-managed).
  • Strong background in database performance tuning (PostgreSQL, Redis).
  • Experience implementing autoscaling strategies for high-traffic systems.
  • Familiarity with service mesh technologies.
  • Experience building internal developer platforms (IDP).
  • Background in security best practices (zero-trust networking, policy-as-code).
  • Experience with multi-region or globally distributed systems.
  • Experience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing).

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.


Benefits
  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources

Skills Required

  • 8+ years of experience in DevOps, SRE, or infrastructure engineering
  • Experience designing and operating production Kubernetes environments at scale
  • Hands-on expertise with AWS infrastructure and cloud networking
  • Experience building and maintaining Terraform modules across large cloud environments
  • Ownership of CI/CD systems and measurable improvement of DORA metrics
  • Experience leading incident response processes and driving postmortem outcomes
  • Understanding of distributed systems, event-driven architectures, Kafka, and PostgreSQL performance
  • Experience modernizing legacy infrastructure and eliminating manual operational toil
  • Ability to take infrastructure projects from ambiguity through production independently
  • Ability to build trust across teams while improving reliability
  • Application coding experience
  • Experience operating high-throughput Kafka clusters
  • Database performance tuning experience with PostgreSQL and Redis
  • Experience implementing autoscaling strategies for high-traffic systems
  • Familiarity with service mesh technologies
  • Experience building internal developer platforms
  • Background in zero-trust networking or policy-as-code
  • Experience with multi-region or globally distributed systems
  • Experience implementing SLOs, error budgets, or chaos testing

Worth Compensation & Benefits Highlights

  • Healthcare Strength Core medical, dental, and vision coverage are complemented by HSA/FSA options, life insurance, an EAP, and wellness resources, creating a comprehensive health package. Individual coverage is characterized as solid, reinforcing perceived strength for single enrollees.
  • Leave & Time Off Breadth Unlimited/flexible PTO, paid holidays, and bereavement leave are included as part of the offering. This breadth points to flexible time-away options beyond the basics.
  • Parental & Family Support Generous parental leave is highlighted as part of the package. New parents can expect paid time away in addition to standard PTO and holidays.

Worth Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Winter Park, Florida
70 Employees
Year Founded: 2023

What We Do

Worth is the AI-powered platform that consolidates onboarding, underwriting, and risk monitoring for fintechs, lenders, payment processors, and financial institutions. Founded in 2023, we built Worth to replace slow, manual underwriting with a single system that verifies, scores, and monitors small and medium-sized businesses (SMBs) in real time. At the center of the platform is Crosswalking Technology. Our proprietary AI/ML models intelligently match businesses across disparate data sources, ensuring the highest level of accuracy and reliability in SMB entity resolution. By integrating multiple first- and third-party authoritative data sources into our crosswalk-matching logic, Worth ensures that businesses are correctly identified, even in cases of duplicate addresses, name variations, or incomplete records. This data moat spans 186 integrations and 25 global and local partners across 200+ countries and territories, resolving fragmented SMB signals into a database of 350M+ SMBs with a 98% data match rate. Our product suite — Worth Pre-Fill, Custom Onboarding, Case Management, Decisioning Engine, Perpetual Risk Monitoring, and Worth Wallet — is available via API, SDK, or fully white-labeled, enabling financial institutions to consolidate their entire onboarding and underwriting stack into one platform. Customers using Worth have increased approval rates by 37%+, reduced application abandonment by 43%+, cut vendor costs by 25%, and reduced time to revenue by 55%+. We're SOC 2 Type II certified and GDPR and CCPA compliant, and have raised $55M in funding to date. Today, 50+ customers rely on Worth to onboard and underwrite their SMB customers faster and more accurately.

Why Work With Us

We're solving a genuinely hard problem: turning fragmented SMB data into one durable, explainable identity that banks and lenders can trust. Backed by $55M in funding and already live with 50+ customers, we're a tight-knit team with real traction, where your work would help shape the roadmap.

Gallery

Gallery
Gallery
Gallery

Worth Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Typical time on-site: Not Specified
HQOrlando

Similar Jobs

Worth Logo Worth

Director Of Sales

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Worth Logo Worth

Security Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Worth Logo Worth

Business Development Representative

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Worth Logo Worth

Senior Software Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account