Data Engineer (Temporal & Apache Kafka required)

Posted 4 Days Ago
McLean, VA, USA
In-Office
90K-154K Annually
Mid level
AdTech • Information Technology • Marketing Tech
The Role
Design and scale event-driven data platforms using Apache Kafka and Temporal. Build batch and streaming ETL/ELT pipelines with Python, SQL, and Spark; develop data schemas, validation, contracts, and quality controls; optimize warehouse and lakehouse models; and implement reliable, observable workflows. Collaborate with software, machine learning, and analytics teams on distributed systems, CI/CD, and production data governance.
Summary Generated by Built In

About Infinitive
Infinitive is a data and AI consultancy that helps clients modernize, monetize, and operationalize their data to generate lasting value. They pride themselves on their deep industry and technology expertise, ensuring that they drive and sustain the adoption of new capabilities. Infinitive is committed to aligning their team with their clients' culture, ensuring a successful partnership by bringing the right mix of talent and skills for high return on investment.

Infinitive has earned recognition as one of the "Best Small Firms to Work For" by Consulting Magazine, receiving this accolade nine times, most recently in 2026. They have also been honored as a “Top Workplace” by the Washington Post, “Best Places to Work” by the Washington Business Journal, and “Best Places to Work” by Virginia Business.
About the Role

We are seeking an experienced Data Engineer to help design, build, and scale our next-generation event-driven data platforms. In this role, you will be instrumental in bridging high-throughput distributed streaming with complex, fault-tolerant workflow orchestration and strict data governance.

You will work extensively with Apache Kafka for real-time event streaming and Temporal (the open-source, durable execution engine originating from Uber/Cadence) to build resilient, distributed stateful workflows and data pipelines. A core focus of this position is establishing robust data schema design and automated validation to ensure strong data contracts across distributed systems. Alongside these technologies, you will design robust batch and streaming ETL/ELT pipelines leveraging Python, Apache Spark, and modern cloud data warehouses/lakehouses.

Key Responsibilities
  • Stream Processing & Messaging: Architect, deploy, and maintain high-volume distributed data streams using Apache Kafka (producers, consumers, Kafka Connect, Schema Registry).

  • Data Schema Design & Validation: Establish and enforce schema design standards, versioning strategies, and automated schema validation (e.g., Avro, Protobuf, JSON Schema) to maintain strict data contracts across microservices, streaming consumers, and lakehouse storage.

  • Resilient Workflow Orchestration: Design and implement durable execution workflows using Temporal to coordinate long-running distributed pipelines, compensate transactions (Saga pattern), and manage cross-system ETL tasks.

  • Pipeline Development: Build end-to-end batch and near-real-time pipelines using Python, SQL, and Apache Spark / PySpark.

  • Data Modeling & Warehousing: Design and optimize analytical data models (dimensional/star schema) in modern cloud data warehouses/lakehouses (e.g., Snowflake, BigQuery, Databricks, Redshift).

  • Reliability & Data Quality: Implement automated testing, continuous schema validation, data drift detection, and observability across streaming and batch workflows.

  • Cross-Functional Collaboration: Partner with software engineers, machine learning engineers, and analysts to define standard schema definitions, data contracts, and production-grade CI/CD release patterns.

Required Experience

  • 4+ years of professional experience in data engineering, backend distributed systems, or software engineering.

  • Hands-on experience with Temporal (or Cadence): Proven understanding of durable workflows, activities, retries, signals, queries, and long-running distributed task orchestration.

  • Deep expertise with Apache Kafka: Practical experience with message partitioning, consumer groups, offset management, and topic design.

  • Strong background in Data Schema Design & Validation:

    • Demonstrated proficiency with schema definition frameworks (Apache Avro, Protocol Buffers/gRPC, or JSON Schema).

    • Practical experience managing schema evolution, compatibility modes (backward/forward/full), and schema registries (e.g., Confluent Schema Registry, AWS Glue Schema Registry).

    • Experience enforcing data validation rules, contract testing, and data quality checks (e.g., Great Expectations, Pandera, Pydantic, dbt tests).

  • Strong programming proficiency in Python (Go or Java is a plus) with clean code, design patterns, and unit/integration testing standards.

  • Distributed computing experience: Hands-on development with Apache Spark (PySpark/Spark SQL) processing large-scale datasets.

  • Advanced SQL & Data Modeling: Strong experience with relational databases, dimensional data modeling, and query performance tuning.


Infinitive is required by law in some jurisdictions to include a reasonable estimate of the compensation range for this role. The determination of this range includes various factors not limited to skill set, level, experience, relevant training, and licensure and certifications. Compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range for this role in the U.S. is $90,000 - $154,00.00.
Infinitive is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other characteristic protected by applicable federal, state, or local law.

Skills Required

  • 4+ years of professional experience in data engineering, backend distributed systems, or software engineering
  • Hands-on experience with Temporal or Cadence, including durable workflows, activities, retries, signals, queries, and long-running orchestration
  • Deep practical experience with Apache Kafka, including partitioning, consumer groups, offset management, and topic design
  • Proficiency with schema definition frameworks such as Apache Avro, Protocol Buffers/gRPC, or JSON Schema
  • Experience managing schema evolution, compatibility modes, and schema registries
  • Experience with data validation, contract testing, and data quality tools such as Great Expectations, Pandera, Pydantic, or dbt tests
  • Strong programming proficiency in Python
  • Go or Java experience
  • Experience developing with Apache Spark, PySpark, or Spark SQL for large-scale data processing
  • Advanced SQL, relational database, dimensional data modeling, and query performance tuning experience
  • Clean code, design patterns, and unit/integration testing experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Ashburn, VA
174 Employees
Year Founded: 2003

What We Do

Infinitive is a transformation and technology consultancy. We enable global brands to deliver kick-ass results through insights, innovation, and efficiency. We possess deep industry and technology expertise to drive and sustain adoption of new capabilities. We match our people and personalities to our clients’ culture while bringing the right mix of talent and skills to enable a high return on investment. Our strong workplace culture has received recognition from Inc. magazine, The Washington Post, Consulting Magazine, Washington Business Journal and other top media outlets and awards programs.

Similar Jobs

Wipfli Logo Wipfli

Tax Manager

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
106K-160K Annually

Wipfli Logo Wipfli

Senior Manager, Accounting Advisory - Tribal Government Industry

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
142K-195K Annually

Wipfli Logo Wipfli

Manager, Financial Reporting - Skilled Nursing Clients

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
97K-145K Annually

Wipfli Logo Wipfli

Manager, Accounting Advisory - Physician Practices

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
2900 Employees
107K-160K Annually

Similar Companies Hiring

NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account