Senior Data Engineer

Posted 4 Days Ago
3 Locations
Hybrid
165K-220K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Software
Regard is empowering the future of medicine.
The Role
Design, build, and operate data services and pipelines to ingest, standardize, and deliver clinical data for product, analytics, and ML. Tune Spark workloads, enforce data quality, monitor production systems, and partner with product and engineering. Participate in on-call support.
Summary Generated by Built In

As a Senior Data Engineer at Regard, you will own the design, development, and production deployment of the data services that power the Regard platform. From ingesting and standardizing clinical data across health systems to making it reliably available for downstream product, analytics, and machine learning workflows, you'll build and evolve the infrastructure that enables the platform. This includes analyzing and tuning Spark workloads and partitioning strategies to control costs, adapting to upstream breaking changes, and enforcing rigorous data quality standards so our analytics are as dependable as our application code. We prioritize transparent, code-driven systems over black-box services, and you'll help architect the data platform that supports that philosophy.

 

About Regard

Our mission is to bring world-class healthcare to everyone. Regard is an AI-powered Proactive Documentation platform that advances how care is delivered by reviewing all patient data in the EHR to recommend diagnoses and surface clinical evidence. Regard drafts a note even before the physician sees the patient, enabling an approach that gets  documentation right at the point of care - we call it Proactive Documentation. This improves quality of care, reduces physician burden, and improves hospital finances. We are excited by challenges, mission-oriented work, and meaningful relationships. We work closely with some of the top health systems in the country and are leading the change that healthcare - one of the largest and most inefficient industries in the world - needs. We want you to join us.

Our Tech Stack:

  • Data: S3, Apache Iceberg, EMR, PySpark, Dagster, Kubernetes, Clickhouse, PostgreSQL, FastAPI, Metabase

 

Responsibilities:

  • Collect, model, and consolidate data into the data platform to support analytics, ML development, and research initiatives

  • Design, build, and evolve data models and pipelines that reliably transform and deliver data to downstream consumers

  • Own data quality in collaboration with engineering teams, ensuring datasets are trustworthy and production-ready

  • Partner closely with product to deliver analytics and actionable insights to internal and external stakeholders

  • Own the reliability and day-to-day operation of the data platform and its pipelines through proactive monitoring, alerting, and operational management

Minimum Qualifications:

  • Bachelors degree in Computer Science, Mathematics, Statistics, or a related field, or equivalent practical experience

  • 5+ years of experience in data engineering roles

  • 3+ years of experience using PySpark to build data pipelines

  • 3+ years of experience in public cloud provider technologies (AWS tooling such as S3, EMR, or Athena)

  • Strong proficiency in Python and SQL

  • Hands-on experience across the full data stack, with particular depth in data modeling and pipeline design

  • Practical experience with LLM-assisted development, with an understanding of its capabilities and limitations

  • Willingness to participate in on-call operational support for owned systems

Preferred Qualifications:

  • Experience with one or more of the following technologies: Apache Iceberg, Dagster, Clickhouse, PostgreSQL, FastAPI, Metabase

  • Experience working with healthcare data, including HIPAA compliance, data de-identification, and familiarity with open data standards such as OMOP CDM

  • Experience building and supporting data pipelines for ML workflows, including model training, validation, deployment, and ongoing performance evaluation

Hybrid Work | Location | Work Authorization

  • For this role, Regard is currently only considering candidates who are authorized to work in the US without visa sponsorship, and are within the New York City, Los Angeles, or San Francisco metro areas

  • We expect our Engineers to be in the office on Tuesdays and Thursdays. We also require more frequent in-office work during the onboarding period and team onsite weeks up to once per month

  • We will provide relocation assistance to anyone who does not already reside in the NYC metro area

  • We prefer hiring people within commuting distance of our offices because we value getting together in person regularly

  • For those who enjoy working from our LA or Manhattan offices on a more regular basis, we offer catered lunches and other fun perks

  • Additionally, hybrid employees have the flexibility to work from locations outside of their home office from up to 6 weeks per year

Comp | Perks | Benefits

  • Eligible for equity

  • 99% employer paid health benefits (Medical, Dental, and Vision) + One Medical subscription

  • 18 PTO days/yr + 1 week holiday break

  • Monthly health & wellness budget

  • Company-sponsored team retreat + social events

  • A sabbatical program

Our goal at Regard is to provide and maintain a work environment that fosters mutual respect, professionalism and cooperation. Regard is proud to be an equal opportunity employer that does not discriminate on the basis of actual or perceived race, creed, color, religion, national origin, ancestry, alienage or citizenship status, age, disability or handicap, sex, gender identity, marital status, familial status, veteran status, sexual orientation or any other characteristic protected by applicable federal, state or local laws. We celebrate diversity and are proud of our supportive, inclusive workplace.

 

All candidates must successfully complete a background check as part of the hiring process.

Skills Required

  • Bachelor's degree in Computer Science, Mathematics, Statistics, or related field, or equivalent practical experience
  • 5+ years of experience in data engineering roles
  • 3+ years using PySpark to build data pipelines
  • 3+ years working with public cloud provider technologies (AWS tooling such as S3, EMR, or Athena)
  • Strong proficiency in Python and SQL
  • Hands-on experience across the full data stack, with particular depth in data modeling and pipeline design
  • Practical experience with LLM-assisted development
  • Willingness to participate in on-call operational support for owned systems
  • Experience with Apache Iceberg, Dagster, Clickhouse, PostgreSQL, FastAPI, Metabase
  • Experience working with healthcare data, HIPAA compliance, data de-identification, and standards such as OMOP CDM
  • Experience building and supporting data pipelines for ML workflows (training, validation, deployment, evaluation)
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Los Angeles, CA
75 Employees
Year Founded: 2017

What We Do

Regard is a health-tech start-up designed with physician workflows in mind. Regard synthesizes historical patient data, automates note-taking, and recommends diagnoses, empowering physicians to practice at the top of their license and administrators to spend their energy advancing the industry.

Why Work With Us

At Regard, we are building a product that is saving lives. We have an open feedback/growth minded culture where you have direct access to executive team. Being a remote/hybrid company, we offer an outstanding work/life balance with opportunities for the team to come together in person for work and fun!

Gallery

Gallery

Similar Jobs

Samsara Logo Samsara

Senior Data Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
United States
4000 Employees
134K-203K Annually

CoreWeave Logo CoreWeave

Senior Data Engineer

Cloud • Information Technology • Machine Learning
In-Office
5 Locations
1450 Employees
153K-204K Annually

Fusion Risk Management Logo Fusion Risk Management

Senior Data Engineer

Professional Services • Software
Remote or Hybrid
United States
258 Employees
135K-155K Annually

Babylist Logo Babylist

Senior Data Engineer

eCommerce • Healthtech • Kids + Family • Retail • Social Media
Easy Apply
Remote or Hybrid
United States
300 Employees
186K-222K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account