Sr. Data Engineer

Posted 2 Hours Ago
Hiring Remotely in United States
Remote
100K-157K Annually
Senior level
Information Technology
The Role
Manages and evolves enterprise data lakes and warehouses, designing scalable AWS-based batch, CDC, and near-real-time ETL/ELT pipelines. Builds data quality, governance, lineage, security, monitoring, and recovery controls; optimizes Spark, Glue, Athena, Redshift, and S3 workloads. Develops reusable Python, SQL, and PySpark frameworks, supports schema evolution and analytical data modeling, and collaborates with business and technical stakeholders to deliver reliable data solutions.
Summary Generated by Built In
Overview

Manages and evolves the enterprise data lake and data warehouse while ensuring the reliable, secure, and efficient flow of high-quality data. Implements data processes, managing data architecture, designing ETL processes, and analyzing data for business insights.


The base salary range for this position is $99,937-$157,044.


Actual pay will be determined based upon a candidate’s job-related knowledge, skills, education, experience, geographic location, and may include other job-related factors such as certification(s), professional licensure, or internal equity considerations.

Responsibilities
  • Design, implement and maintain scalable data pipnes on WS using S3, DMS, Glue, lambda, step function/MWAA & Redshift.
  • Develop robust batch and near-real-time ETL/ELT workflow to ingest, cleanse, transform and load data from databases, legacy applications and event streams using Python & Pyspark.
  • Design incremental/CDC mechanism, including restart ability, idempotency, duplicate handling and recovery.
  • Implement automated controls for completeness, accuracy, reconciliation, schema changes & lineage.
  • Optimize Glue/Spark, Athena, Redshift & S3 workload through partitioning, columnar formats, query tuning and appropriate storage/compute design.
  • Design near real time/event-driven pipelines using Kinesis/Kafka where required, covering ordering, retry, idempotency and failure recovery.
  • Implement AWS data security, least privilege access, data classification and governance controls.
  • Monitor pipelines such as CloudWatch, troubleshoot failure and resolving production data incidents.
  • Enforce Git/version control, code review, automated testing and CI/CD practices.
  • Work with product owners, architect, reporting and business stakeholders to translate requirements into scalable data solutions.
  • Document pipelines and operational procedures.

Qualifications

Qualifications Required

  • Bachelor's degree in a computer-related field from an accredited college or university and five (5) or more years of experience in data engineering, building scalable and distributed ETL data pipelines in enterprise environments.
  • Experience building and operating scalable AWS-based data platforms and pipelines using services including Lambda, Glue, Athena, S3, Redshift, DMS, MWAA (Airflow), and Step Functions, supporting batch, CDC, and near real-time data processing.
  • Advanced proficiency in Python, SQL, and PySpark with hands-on experience developing reusable ETL/ELT frameworks, data warehouses, data marts, and integrations across databases, APIs, event streams, and analytics environments.
  • Experience implementing data quality, governance, and optimization best practices, including automated validation frameworks, Lake Formation and Glue Data Catalog, performance tuning, and cost optimization across AWS data services.
  • Strong communication skills with the ability to translate complex data concepts for business stakeholders; experience in healthcare, life sciences, and other highly regulated environments with HIPAA, GDPR, FDA, or similar compliance requirements preferred.
  • Experience with metadata management, data lineage, data observability, master data management, or enterprise data catalog solutions.
  • Knowledge with data modeling & analytical data model, schema design, schema evolution, and data structure optimized for reporting and analytics.
  • Knowledge of data lake and data warehouse architecture include data partitioning and columnar storage format such as Parquet.
  • Relevant AWS certification, such as AWS Certified Data Engineer – Associate, or an equivalent cloud or data engineering certification.

WORKING CONDITIONS

  • Flexible work hours in fun collaborative environment
  • Working remote requires a reliable internet connection
  • Must have the ability to travel, as needed for company meetings

Skills Required

  • Bachelor's degree in a computer-related field from an accredited college or university
  • Five or more years of experience in data engineering and building scalable, distributed ETL data pipelines in enterprise environments
  • Experience building and operating scalable AWS-based data platforms and pipelines using Lambda, Glue, Athena, S3, Redshift, DMS, MWAA or Airflow, and Step Functions
  • Experience supporting batch, change data capture, and near-real-time data processing
  • Advanced proficiency in Python, SQL, and PySpark
  • Hands-on experience developing reusable ETL/ELT frameworks, data warehouses, data marts, and integrations across databases, APIs, event streams, and analytics environments
  • Experience implementing data quality, governance, and optimization practices, including automated validation frameworks, Lake Formation, Glue Data Catalog, performance tuning, and AWS cost optimization
  • Strong communication skills and ability to translate complex data concepts for business stakeholders
  • Experience in healthcare, life sciences, or other highly regulated environments with HIPAA, GDPR, FDA, or similar compliance requirements
  • Experience with metadata management, data lineage, data observability, master data management, or enterprise data catalog solutions
  • Knowledge of data modeling, analytical data models, schema design, schema evolution, and data structures optimized for reporting and analytics
  • Knowledge of data lake and data warehouse architecture, including data partitioning and columnar formats such as Parquet
  • Relevant AWS certification, such as AWS Certified Data Engineer Associate, or an equivalent cloud or data engineering certification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Miramar, FL
62 Employees
Year Founded: 2020

What We Do

With an emphasis on quality management and customer service, we provide superior solutions which reinvent the user experience.

Similar Jobs

Centerfield Logo Centerfield

Senior Data Engineer

AdTech • Consumer Web • Digital Media • eCommerce • Marketing Tech • SEO
Remote or Hybrid
United States
890 Employees

Doximity Logo Doximity

Senior Software Engineer

Healthtech • Information Technology • Mobile • Productivity • Software • Analytics • Telehealth
Easy Apply
In-Office or Remote
2 Locations
740 Employees
165K-221K Annually

CDW Logo CDW

Data Engineer

Information Technology
Remote or Hybrid
US
15100 Employees
105K-150K Annually

Airwallex Logo Airwallex

Senior Software Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office or Remote
Seattle, WA, USA
2300 Employees
180K-240K Annually

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account