Staff Distributed Systems Engineer

Posted 4 Days Ago
Be an Early Applicant
New York City, NY, USA
In-Office
Senior level
Fintech • Mobile • Social Media • Cryptocurrency
The Role
Own the reliability, scalability, and performance of a multi-region backend platform. Design and operate high-throughput services, optimize PostgreSQL and Redis-compatible datastores, implement resilience controls, reduce cross-region latency, and build tested failover and disaster-recovery capabilities. This hands-on role directly operates production infrastructure, improves observability, and helps architect scalable features.
Summary Generated by Built In
About the role

fomo is a trading app for the rest of us. With fomo you can sign up in seconds and have instant access to any asset on-chain without the need for external wallets, bridges or prior knowledge. fomo has social features that are actually useful - follow top traders (with full access to their portfolio and trades) and find tokens early. fomo does all of this while providing the best-in-class execution and data for experienced traders.

https://x.com/fomo

About the role

We are looking for a Staff Distributed Systems Engineer to own the reliability, scalability, and performance of our multi-region backend platform.

You will own critical shared infrastructure, including datastores, caches, messaging systems, and regional application services, and design systems that remain predictable during traffic surges, dependency failures, infrastructure changes, and partial regional outages.

You will also establish new failover and disaster-recovery capabilities within our stack, including defining recovery objectives and implementing the systems and testing required to recover services and data safely.

This is a hands-on engineering role with direct ownership of production systems. You will build and operate application and infrastructure components while improving data systems and observability.

Responsibilities

  • Design and operate high-throughput, multi-region services.

  • Improve datastore and cache performance, capacity, replication, and failure handling.

  • Implement backpressure, concurrency limits, load shedding, rate limiting, circuit breakers, and bounded retries.

  • Reduce cross-region latency and improve data locality.

  • Design and test service, datastore, and regional failover procedures.

  • Help architect new features to operate at scale from day one.

  • Level-up the team on how to think about scale.

Qualifications

  • 8 or more years of backend, platform, or infrastructure engineering experience, or equivalent practical experience.

  • Experience designing and debugging distributed, high-throughput production systems.

  • Strong PostgreSQL experience, including query performance, indexing, connection pooling, replication, transaction contention, and failure modes.

  • Strong experience with Redis-compatible systems such as Redis, Valkey, Dragonfly, or KeyDB, including sharding, replication, memory management, hot keys, and failure handling.

  • Experience operating services on AWS, ideally using ECS, RDS, and ElastiCache.

  • Experience with infrastructure as code, preferably Terraform.

  • Proficiency in Go, TypeScript/Node.js, or a comparable systems-oriented language.

  • Hands-on experience designing and testing failover and disaster-recovery systems, including backup restoration, replication, regional failover, and RTO/RPO validation.

Nice to have

  • Experience with NATS JetStream, Kafka, or another durable messaging system.

  • Familiarity with Datadog APM and AWS Performance Insights.

  • Experience performing live datastore or cache topology migrations.

  • Experience operating systems with bursty or unpredictable traffic.

  • Experience with financial, trading, cryptocurrency, gaming, or other high-throughput systems.

Skills Required

  • 8 or more years of backend, platform, or infrastructure engineering experience, or equivalent practical experience
  • Experience designing and debugging distributed, high-throughput production systems
  • Strong PostgreSQL experience, including query performance, indexing, connection pooling, replication, transaction contention, and failure modes
  • Strong experience with Redis-compatible systems such as Redis, Valkey, Dragonfly, or KeyDB, including sharding, replication, memory management, hot keys, and failure handling
  • Experience operating services on AWS, ideally using ECS, RDS, and ElastiCache
  • Experience with infrastructure as code, preferably Terraform
  • Proficiency in Go, TypeScript/Node.js, or a comparable systems-oriented language
  • Hands-on experience designing and testing failover and disaster-recovery systems, including backup restoration, replication, regional failover, and RTO/RPO validation
  • Experience with NATS JetStream, Kafka, or another durable messaging system
  • Familiarity with Datadog APM and AWS Performance Insights
  • Experience performing live datastore or cache topology migrations
  • Experience operating systems with bursty or unpredictable traffic
  • Experience with financial, trading, cryptocurrency, gaming, or other high-throughput systems
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
6 Employees
Year Founded: 2025

What We Do

Fomo Labs develops fomo, a social-first, cross-chain cryptocurrency trading platform available on web and mobile. The app lets users trade memecoins, altcoins, stablecoins, and other supported assets, follow traders, copy trades, and swap across supported chains without bridging. Its mission is to make onchain trading more accessible and social, helping users discover trends and participate in them quickly through a streamlined user experience.

Similar Jobs

Anthropic Logo Anthropic

Software Engineer

Artificial Intelligence • Natural Language Processing • Generative AI
In-Office
3 Locations
2500 Employees
320K-485K Annually

LiveKit Logo LiveKit

Staff Software Engineer

Artificial Intelligence • Cloud • Information Technology • Software
In-Office or Remote
30 Locations
83 Employees
135K-300K Annually

LiveKit Logo LiveKit

Staff Software Engineer

Artificial Intelligence • Information Technology • Internet of Things
In-Office or Remote
30 Locations
34 Employees
135K-300K Annually

Tubi Logo Tubi

Staff Software Engineer

News + Entertainment
In-Office or Remote
3 Locations
504 Employees
227K-325K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account