DataOps Engineer

Posted 16 Hours Ago
Be an Early Applicant
Englewood Cliffs, NJ, USA
In-Office
110K-125K Annually
Senior level
Information Technology • Professional Services
The Role
Design, build, and operate a production data platform using Apache Iceberg and containerized data engines. Own CI/CD, automation, observability, incident response, security controls, and SLA monitoring for ETL/ELT and streaming pipelines; maintain catalogs and table health, run on-call rotations, and document practices.
Summary Generated by Built In
Company Description

For More Open Positions Visit us at:
http://recruiting.woongjininc.com/

Our Mission

WOONGJIN, Inc. is a rapidly growing team who provides a range of unique, exceptional, and enhanced services to our clients. We have a strong moral code that includes the service of goodness without expectations of reward. We are motivated by the sense of responsibility and servant leadership.

Benefits

  • Medical Insurance
  • Vision Insurance
  • Dental Insurance
  • 401(k)
  • Paid Sick hours

Job Description

We are looking for a mid‑level engineer to build and operate a data platform that uses Apache Iceberg as the lake‑house table format and Docker‑based micro‑services (Spark, Flink, Presto, etc.). you will own the end‑to‑end delivery pipeline, monitoring, security, and incident response, ensuring the platform runs reliably at scale.

 

Key Responsibilities

  • Iceberg operations: support tables, manage schema changes, partitions, snapshot retention, and keep the catalog (Hive Metastore, AWS Glue, Nessie, …) synchronized.
  • Docker image creation & testing: write multi‑stage Dockerfiles for Spark/Flink/Presto, run local test environments with Docker‑Compose, and conduct vulnerability scans (Trivy, Snyk, …).
  • Data pipeline development: build ETL/ELT jobs that ingest raw data and write to Iceberg tables; add simple streaming components using Kafka, Pulsar, or Kinesis when needed.
  • CI/CD automation: configure pipelines (GitHub Actions, GitLab CI, Azure DevOps, …) to lint Dockerfiles, scan images, version Iceberg metadata, and deploy pipelines without downtime.
  • Automation with Ansible/Python: script cluster provisioning, catalog configuration, vacuum/compaction, and other routine housekeeping tasks.
  • Observability: instrument services with OpenTelemetry, Prometheus, Grafana, and Loki; create dashboards showing pipeline latency, resource usage, table health, and error rates; set up basic alerts.
  • SLA monitoring: measure data freshness, job success rates, and query response times against agreed‑upon targets and report deviations.
  • Incident response: join the on‑call rotation, perform first‑line diagnosis and resolution of pipeline failures, Iceberg metadata issues, or container crashes; write concise root‑cause analyses and suggest improvements.
  • Security & compliance support: help enforce image signing, mTLS, IAM roles, and bucket policies; collaborate with the security team to meet GDPR, HIPAA, or ISO 27001 requirements.
  • Knowledge sharing: keep internal documentation up to date and run short tech demos or brown‑bag sessions on Iceberg, Docker best practices, and automation techniques.

     

 

Qualifications

  • Bachelor’s degree in Computer Science, IT, Data Engineering, or a related field (Master’s a plus).
  • 5+ years of hands‑on experience building and operating large‑scale data platforms (lake‑house, data‑warehouse, or big‑data ecosystems).
  • Proven production experience with Apache Iceberg (table creation, partition management, schema evolution, catalog integration).
  • Strong Docker skills: multi‑stage builds, Docker‑Compose testing, routine image security scanning.
  • Experience with at least one major data‑processing engine (Spark, Flink, or Presto/Trino) and its connection to Iceberg tables.
  • Proficiency in Python and/or Ansible for automating infrastructure and platform tasks.
  • Experience building CI/CD pipelines that include Docker linting, vulnerability scanning, and automated deployment of data‑pipeline code.
  • Familiarity with observability tooling (Prometheus, Grafana, OpenTelemetry, Loki) and ability to create useful alerts and dashboards.
  • Ability to respond to incidents, write clear root‑cause analysis reports, and contribute to post‑mortem actions.
  • Willingness to participate in an on‑call rotation as a first‑line responder.
  • Availability to work on‑site in New Jersey for the initial assignment and relocate to Dallas by October 2026.

 

Preferred Qualifications

  • Experience with cloud‑native data services on AWS, Azure, or GCP (EMR, Dataproc, Synapse, etc.).
  • Familiarity with other lake‑house formats such as Delta Lake or Apache Hudi and ability to evaluate trade‑offs against Iceberg.
  • Knowledge of streaming platforms (Kafka, Pulsar, Kinesis) and real‑time processing patterns.
  • Relevant certifications (Databricks Lakehouse Associate, Google Professional Data Engineer, AWS Certified Data Analytics – Specialty, etc.).
  • Background supporting data platforms in regulated industries (pharma, finance, healthcare) and understanding of associated compliance frameworks.

Additional Information

All your information will be kept confidential according to EEO guidelines.

 *** NO C2C ***

Skills Required

  • Bachelor's degree in Computer Science, IT, Data Engineering, or related field
  • 5+ years building and operating large-scale data platforms (lake-house, data-warehouse, or big-data ecosystems)
  • Production experience with Apache Iceberg (table creation, partition management, schema evolution, catalog integration)
  • Strong Docker skills including multi-stage Dockerfiles and Docker-Compose testing
  • Experience with at least one data-processing engine: Spark, Flink, or Presto/Trino and connecting it to Iceberg tables
  • Proficiency in Python and/or Ansible for automation and provisioning
  • Experience building CI/CD pipelines with Docker linting, vulnerability scanning, and zero-downtime deployment
  • Familiarity with observability tooling: OpenTelemetry, Prometheus, Grafana, Loki and creating alerts/dashboards
  • Ability to respond to incidents, produce clear root-cause analyses, and participate in on-call rotation
  • Willingness/ability to work on-site in New Jersey initially and relocate to Dallas by October 2026
  • Knowledge of Iceberg catalogs such as Hive Metastore, AWS Glue, or Nessie
  • Routine image security scanning experience (Trivy, Snyk)
  • Experience with schema changes, snapshot retention, partitioning, and compaction/vacuum operations
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Buena Park, CA
24 Employees
Year Founded: 1976

What We Do

Woongjin Inc. is a Korean based company in the US providing unique IT Consulting and Recruiting/Staffing services.

Similar Jobs

SBT Global, Inc. Logo SBT Global, Inc.

DataOps Engineer

Automotive • eCommerce • Manufacturing
In-Office
Englewood Cliffs, NJ, USA
122K-122K Annually

Zscaler Logo Zscaler

Principal, Product Marketing - ZPA

Cloud • Information Technology • Security • Software • Cybersecurity
Easy Apply
Remote or Hybrid
USA
8697 Employees
200K-285K Annually

ZS Logo ZS

Consultant

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
3 Locations
15000 Employees
160K-177K Annually

ZS Logo ZS

Consultant

Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Hybrid
3 Locations
15000 Employees
120K-137K Annually

Similar Companies Hiring

Scrunch  Thumbnail
Artificial Intelligence • Information Technology • Marketing Tech • Software • SEO
Salt Lake City, Utah
Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account