Software Engineer - Data Platform

Posted 2 Days Ago
Be an Early Applicant
Palo Alto, CA, USA
In-Office
180K-440K Annually
Senior level
Information Technology
The Role
Build, operate, and scale distributed data infrastructure (Kafka, Spark, Flink, Trino, HDFS) to enable real-time ML pipelines and analytics at petabyte scale. Design high-throughput, low-latency ingestion and transport, optimize performance, debug distributed systems, and collaborate with ML and product teams to ensure reliable, production-grade data movement and compute.
Summary Generated by Built In

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

The Data Platform team builds and operates the infrastructure responsible for all large-scale data transport and processing across the company. We own and manage core systems including Apache Kafka, HDFS, Spark, Flink, and Trino, enabling real-time ML pipelines, feed ranking, experimentation, analytics, and observability at petabyte scale. Our team deals with latency-critical workloads, high-throughput streaming, and distributed compute systems that require fault tolerance, performance, and absolute reliability.

As a software engineer on the Data Platform team, you will design, build, and operate the distributed systems powering SpaceXAI's data movement and compute. You will take ownership of infrastructure components that process trillions of events daily, driving the scalability, performance, and reliability of the systems that power product and ML workloads across the company.

RESPONSIBILITIES:
  • Design and implement high-throughput, low-latency data ingestion and transport systems.
  • Scale and optimize multi-tenant Kafka infrastructure supporting real-time workloads.
  • Extend and tune Spark, Flink, and Trino for demanding production pipelines.
  • Build interfaces, APIs, and pipelines enabling teams to query, process, and move data at petabyte scale.
  • Debug and optimize distributed systems, with a focus on reliability and performance under load.
  • Collaborate with ML, product, and infrastructure teams to unblock critical data workflows.
BASIC QUALIFICATIONS:
  • Proven expertise in distributed systems, stream processing, or large-scale data platforms.
  • Proficiency in  Rust, Go, Scala  or similar systems languages.
  • Hands-on experience with  Kafka, Flink, Spark, Trino, or Hadoop*in production.
  • Strong debugging, profiling, and performance optimization skills.
  • Track record of shipping and maintaining critical infrastructure.
  • Comfortable working in fast-moving, high-stakes environments with minimal guardrails.
COMPENSATION AND BENEFITS:

$180,000 - $440,000 USD

Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

Skills Required

  • Proven expertise in distributed systems, stream processing, or large-scale data platforms
  • Proficiency in Rust, Go, Scala or similar systems languages
  • Hands-on experience with Kafka, Flink, Spark, Trino, or Hadoop in production
  • Strong debugging, profiling, and performance optimization skills
  • Track record of shipping and maintaining critical infrastructure
  • Strong communication skills
  • Comfortable working in fast-moving, high-stakes environments with minimal guardrails
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Palo Alto, CA
96 Employees

What We Do

Understand the Universe

Similar Jobs

CrowdStrike Logo CrowdStrike

Principal Software Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Hybrid
4 Locations
11000 Employees
195K-290K Annually

ServiceNow Logo ServiceNow

Software Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Santa Clara, CA, USA
29000 Employees
176K-308K Annually

Plenful Logo Plenful

Senior Software Engineer

Artificial Intelligence • Healthtech
Hybrid
San Francisco, CA, USA
120 Employees

Guidewire Software Logo Guidewire Software

Software Engineer

Cloud • Information Technology • Insurance • Software • Analytics
In-Office
San Mateo, CA, USA
3400 Employees
124K-210K Annually

Similar Companies Hiring

Scrunch  Thumbnail
Artificial Intelligence • Information Technology • Marketing Tech • Software • SEO
Salt Lake City, Utah
Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account