AWS Data Engineer

Posted 10 Hours Ago
Be an Early Applicant
2 Locations
Remote
Mid level
Artificial Intelligence • Energy • Renewable Energy
The Role
Build and maintain AWS-based data lakehouse infrastructure, including batch and streaming ETL/ELT pipelines, high-frequency time-series ingestion, orchestration, analytical datasets, governance, security, monitoring, alerting, and data quality testing. The role requires strong Python, PySpark, advanced SQL, open table formats, AWS data services, lakehouse modeling, and infrastructure-as-code experience.
Summary Generated by Built In

ON.energy is building the backbone of energy and AI infrastructure powering grid-safe data centers and mission-critical facilities. The company supplies and operates hyperscale power systems that solve the toughest resilience challenges, delivering custom solutions for AI data centers, mission-critical facilities, and front-of-the-meter assets. ON recently announced a 5GW partnership, with 3GW currently under construction across multiple hyperscale data center campuses. With patented technology and proprietary software, ON.energy develops projects worldwide that set new benchmarks for resilience.

Role Summary

ON.energy is building the power infrastructure that makes the AI era possible. Our systems are deployed across 2.5 GW of hyper-scale campuses, validated by top U.S. national labs, and certified for grid-safe operation by major utilities. You’ll be the fourth engineer on a growing data team, helping scale our AWS-based data lakehouse and working primarily with industrial data sources and high-volume time-series data. No prior experience in the energy sector is required, we’ll support you in learning the domain.


Key Responsibilities
  • Build and maintain scalable ETL/ELT pipelines for batch and real-time processing, including high-frequency time-series ingestion.
  • Evolve the Data Lakehouse — optimizing storage, performance, cost, and data consistency.
  • Manage orchestration workflows with complex dependencies and error handling.
  • Deliver production-ready, analytics-optimized datasets, engaging directly with end users to understand how they consume data.
  • Implement data governance, security controls, and audit policies across the AWS ecosystem.
  • Build monitoring, alerting, and data quality testing for platform reliability.

Key Requirements
  • Bachelor’s degree in Computer Science, Computer Engineering, or a closely related discipline.
  • 3+ years of hands-on Data Engineering on AWS, with real exposure to modern data lakehouse architectures.
  • English at B2 or above. Daily work with English-speaking teams and written documentation.
  • Production experience with open table formats, preferably Apache Iceberg (table design, partitioning, schema evolution). Coming from Delta Lake or Hudi? We’ll support you in transitioning.
  • Core AWS data services: Glue, Athena, Lambda, and S3.
  • Designing and maintaining ETL/ELT pipelines for batch and streaming workloads.
  • Python (PySpark / Python Shell) and advanced SQL (window functions, CTEs, execution plan tuning).
  • Data modeling for analytical workloads in a lakehouse context — medallion architectures, incremental loads, deduplication, historical backfills.
  • Infrastructure as Code: Terraform or CloudFormation.

Preferred Experience

Step Functions · Kinesis · Lake Formation · DynamoDB · custom ETL with boto3, pyiceberg, pyarrow · CI/CD for Glue or Lambda · time-series and IoT/telemetry data at scale


#LI-AD1

For US-based roles - What you’ll get:

  • Competitive salary + annual performance-based bonus eligibility
  • Medical, dental, and vision insurance
  • 401(k) with company match
  • Paid time off and company holidays 

For Mexico-based roles - What you’ll get:

  • Competitive salary + annual performance bonus eligibility
  • Christmas Bonus (Aguinaldo): 30 days
  • Major medical expenses and life insurance
  • Paid time off and holidays (per local policy)

For all roles:

  • Professional development and growth opportunities
  • Opportunity to grow with a mission-driven team shaping the future of clean energy
  • Equal Opportunity: ON.energy is committed to equal employment opportunity and to maintaining a work environment free of harassment, discrimination, or retaliation.
  • Accommodations: If you need an accommodation during the application process, email [email protected]
  • Benefits vary by role and location and are subject to change.

Agency Notice: ON.energy does not accept unsolicited resumes from staffing agencies, search firms, or third-party recruiters. Resumes submitted without a fully executed Master Services Agreement (MSA) and a written request from an authorized member of our Talent Acquisition team will be considered the property of ON.energy. No placement fees or compensation will be paid for unsolicited candidate submissions.

Skills Required

  • Bachelor's degree in Computer Science, Computer Engineering, or a closely related discipline
  • 3+ years of hands-on data engineering experience on AWS
  • English proficiency at B2 level or above
  • Production experience with open table formats, preferably Apache Iceberg
  • Experience with AWS Glue, Athena, Lambda, and S3
  • Experience designing and maintaining ETL/ELT pipelines for batch and streaming workloads
  • Python experience, including PySpark or Python Shell
  • Advanced SQL proficiency, including window functions, CTEs, and execution plan tuning
  • Data modeling for analytical workloads in a lakehouse context
  • Experience with medallion architectures, incremental loads, deduplication, and historical backfills
  • Infrastructure-as-code experience with Terraform or CloudFormation
  • Experience with Step Functions, Kinesis, Lake Formation, or DynamoDB
  • Experience with boto3, pyiceberg, or pyarrow for custom ETL
  • CI/CD experience for Glue or Lambda
  • Experience with time-series and IoT or telemetry data at scale
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Miami, Florida
165 Employees

What We Do

ON.energy is building the backbone of energy and AI infrastructure, powering grid-safe data centers and mission-critical facilities. The company supplies and operates hyperscale power systems that solve the toughest resilience challenges, delivering custom solutions for AI data centers, mission-critical facilities, and front-of-the-meter assets. Its track record spans industrial, manufacturing, infrastructure, transportation, and grid-scale storage. With patented technology and proprietary software, ON.energy develops projects worldwide that set new benchmarks for resilience.

Similar Jobs

Coderio Logo Coderio

Senior Data Engineer

Software • Design • App development
In-Office or Remote
8 Locations
223 Employees

Nortal Logo Nortal

(1542) Senior Data / Reporting Engineer (Python, SQL, AWS)

Big Data • Blockchain • Software • Business Intelligence • App development • Big Data Analytics • Automation
Remote
11 Locations
2800 Employees

Coderio Logo Coderio

Senior Data Engineer

Software • Design • App development
In-Office or Remote
7 Locations
223 Employees

ResilientCo Logo ResilientCo

Data Engineer

Professional Services • Consulting
Remote
11 Locations
12 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account