Senior Data Engineer (Databricks)

Posted 5 Days Ago
Be an Early Applicant
Porto, PRT
Hybrid
Senior level
Cloud • Software • Analytics
The Role
Design, build, and operate Databricks ETL pipelines for real-time IoT and platform data. Deliver production Databricks workloads, BI data models, data quality/testing, lineage and governance. Collaborate with stakeholders, use AI coding tools responsibly, and maintain CI/CD, monitoring, and reusable master data assets for analytics and ML.
Summary Generated by Built In

Join us on the R&D Software team as a Data Engineer (Databricks), and help shape the future of safer, more efficient, and more reliable operations across the globe. Start your journey with Anova today!

Where you’ll work: This is a hybrid role based out of our Porto office. In practice, most of your work can be done remotely, with occasional in-office time in Porto for team collaboration — a flexibility our engineers consistently tell us they value.

Job Duties and Responsibilities: 

You will build and run the Databricks pipelines that turn real-time telemetry and platform data into reliable, well-governed data assets — the master data that reporting, analytics and machine learning across Anova all depend on.

Collaborate for success

  • Deliver Databricks ETL projects end to end, from requirements through to pipelines running in production. 
  • Translate business goals into data solutions and help stakeholders make the right choices about data. 
  • Contribute to technical decisions, take a significant share of the implementation, and monitor the pipelines you own once they are live.

Build the One Anova data stream

  • Work with real-time telemetry from industrial IoT sensors deployed across the globe.
  • Build the BI aggregations that bring data from across platforms together into consistent, reusable data assets.
  • Your pipelines are the backbone for our internal natural-language digital assets that lets any employee query Anova's data without writing SQL. The reliability, freshness and clarity of what you publish directly determines whether that experience can be trusted.
  • Publish and maintain data assets as master data for the organization.

Engineer with AI assistance 

  • Use agentic coding tools — Claude Code, Copilot, Cursor and similar — as a normal part of daily delivery.
  • Hold AI-generated code to the same bar as any other code. You are accountable for what you ship.
  • Keep repositories, tests and documentation structured so both people and agents can work in them effectively.

Advocate for quality 

  • Contribute to and continuously adapt best practices and Ways of Working around data engineering, testing and pipeline operations.
  • Maintain clear data lineage and definitions for the assets you own — as AI agents increasingly query this data directly, untraceable or ambiguous data becomes a governance risk, not just a data-quality one.
  • Treat data quality as a feature: tests, expectations and monitoring, so problems surface before stakeholders find them.

Minimum Requirements - 

  • Bachelor's degree in Computer Science, Data Engineering, Data Science, or a related quantitative field or equivalent combination of education and experience 
  • 5+ years of experience in data engineering or a closely related role, with hands-on production experience in Databricks (6–8 years preferred).
  • Significant experience building data workloads in Databricks, with a very good understanding of PySpark and Delta Lake.
  • Strong SQL — window functions, complex joins and query tuning are everyday tools for you.
  • Experience with streaming or incremental ingestion (Structured Streaming, Auto Loader, or equivalent) and the patterns that keep it correct: idempotency, checkpointing and schema evolution.
  • Data modelling for BI and analytics.
  • Good understanding of testing and CI/CD for Databricks workflows, alongside the software engineering and DevOps basics — git, code review, linters, unit tests and CI/CD pipelines are things you use daily. 
  • Data quality practice: testing data as well as code, using pipeline expectations, dbt tests or similar. 
  • Comfortable using agentic coding tools, with a clear view of where they help and where they need supervision. 
  • Proficient in written and spoken English.

Preferred Qualifications -

  • Databricks platform depth beyond the basics: Lakeflow pipelines (formerly Delta Live Tables), Lakeflow Jobs, Unity Catalog for governance and lineage, and infrastructure as code with Declarative Automation Bundles or Terraform.
  • Performance and cost optimization on Databricks: cluster sizing, Photon, liquid clustering, and partitioning. 
  • The wider Azure data ecosystem: Event Hubs or Data Factory.
  • Master data management or data governance practice: clear ownership, stewardship and agreed definitions for shared data assets.
  • Domain experience in industrial, energy or IoT settings.

Skills Required

  • Bachelor's degree in Computer Science, Data Engineering, Data Science, or related field (or equivalent experience)
  • 5+ years of experience in data engineering or closely related role with hands-on production Databricks experience
  • Significant experience building data workloads in Databricks, strong PySpark and Delta Lake expertise
  • Strong SQL skills including window functions, complex joins, and query tuning
  • Experience with streaming or incremental ingestion (Structured Streaming, Auto Loader) and patterns like idempotency, checkpointing, schema evolution
  • Data modelling for BI and analytics
  • Familiarity with testing and CI/CD for Databricks workflows; software engineering and DevOps basics (git, code review, linters, unit tests, CI/CD)
  • Data quality practices: pipeline expectations, dbt tests or similar
  • Comfortable using agentic coding tools (Claude Code, Copilot, Cursor) and supervising AI-generated code
  • Proficient in written and spoken English
  • Experience with Lakeflow/Delta Live Tables, Lakeflow Jobs, Unity Catalog, and infrastructure-as-code (Declarative Automation Bundles or Terraform)
  • Performance and cost optimization on Databricks (cluster sizing, Photon, liquid clustering, partitioning)
  • Experience with Azure data ecosystem (Event Hubs, Data Factory)
  • Master data management or data governance practice (ownership, stewardship, definitions)
  • Domain experience in industrial, energy, or IoT settings
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Saint Louis, MO
288 Employees
Year Founded: 1989

What We Do

We are the leading global provider of IIoT solutions to remotely manage industrial assets. We have over 30 years of experience in the design, installation and maintenance of wireless hardware, software technologies and cloud-based analytics. With hundreds of thousands of devices monitoring cryogenic gases, LPG/propane, LNG, chemicals, oils, lubricants, fuels, and water, Anova is connecting the industrial world, for better.

Similar Jobs

Pfizer Logo Pfizer

Digital Operations Agentic Lead - Senior Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
29 Locations
121990 Employees

Pfizer Logo Pfizer

Product Specialist

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
29 Locations
121990 Employees

Teya Logo Teya

Back-end Engineer

Fintech • Payments • Financial Services
Hybrid
2 Locations
1000 Employees

Teya Logo Teya

Global Head of Shared Services

Fintech • Payments • Financial Services
Hybrid
2 Locations
1000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account