Top Data Engineer Jobs in Warsaw

17 Days AgoSaved
Hybrid
Warsaw, Warszawa, Mazowieckie, POL
Entry level
Entry level
Artificial Intelligence • Healthtech • Professional Services • Analytics • Consulting
Designs and builds production-grade AI applications natively on Snowflake using Cortex Analyst, Cortex Search, Snowflake Agents, semantic models, Python, and SQL. Develops secure, performant natural-language querying, retrieval, text-to-SQL, and agentic workflows over governed data. Owns delivery of a technical workstream, collaborates with data scientists, engineers, clients, and stakeholders, and supports enterprise GenAI and analytics solutions. Snowflake implementation experience is mandatory; regulated-industry experience is advantageous.
Top Skills: Cortex AnalystCortex SearchGenerative AiLarge Language Models (Llms)PythonSemantic ModelsSemantic ViewsSnowflakeSnowflake AgentsSnowflake CortexSQL
9 Days AgoSaved
In-Office or Remote
32 Locations
124K-207K Annually
Senior level
124K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills: Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL
2 Days AgoSaved
Hybrid
3 Locations
Mid level
Mid level
Hardware • Software
Designs, builds, maintains, and optimizes enterprise data platforms, pipelines, warehouses, models, and integration frameworks supporting Power BI, analytics, AI, and operational reporting. The role develops Azure-based ETL/ELT solutions, improves data quality and reliability, establishes governance and standardized metrics, and collaborates with BI, IT, automation, and AI teams. It is a hybrid position in Katowice, Poland.
Top Skills: APIsAutomated DeploymentAzure Data FactoryAzure DatabricksAzure SqlAzure SynapseCRMDataverseDevOpsEltErpETLMicrosoft FabricPower BIRelational DatabasesSemantic Data LayersSQLTabular ModelsVersion Control
3 Days AgoSaved
In-Office
2 Locations
Mid level
Mid level
Cloud • Information Technology • Internet of Things • Professional Services • Software
Design, implement, and maintain scalable streaming and batch data pipelines using Apache Pinot, Iceberg, Flink, and Spark. Improve pipeline performance, reliability, observability, and data quality while participating in software development, testing, incident resolution, and postmortems. Collaborate with engineering, product, and data teams, use AI coding tools, and continuously develop expertise in production data systems.
Top Skills: Apache FlinkApache IcebergApache PinotSparkCi/CdCloud EnvironmentsCodexCursorGitGithub CopilotJavaPythonScala
Reposted 4 Days AgoSaved
In-Office
Warsaw, Warszawa, Mazowieckie, POL
Mid level
Mid level
Biotech • Pharmaceutical
Develop and maintain production applications and data pipelines to support early drug discovery. Build APIs, backend services, and UIs; implement ETL workflows, data validation, and database integrations; collaborate with scientists and senior engineers; participate in code reviews, CI/CD, monitoring, and incident response to ensure reliable data-driven research systems.
Top Skills: AWSAzureDjangoDockerETLFastapiFlaskGCPGitNoSQLNumpyPandasPythonRestful ApisScikit-LearnSQL
4 Days AgoSaved
In-Office or Remote
Warszawa, Mazowieckie, POL
Junior
Junior
Artificial Intelligence • Big Data • Computer Vision • Machine Learning • Consulting • Conversational AI • Generative AI
Design and develop Azure-based data platforms, ETL/ELT processes, data warehouses, data marts, and data lakes. Write advanced SQL and Python, create Power BI visualizations, model data, and work with Databricks and Spark. Deploy data applications using Docker, Kubernetes, and GitHub while contributing to scalable client projects and maintaining data quality, monitoring, governance, and operational readiness.
Top Skills: SparkAzureAzure Data PlatformCi ServersDatabricksDelta Live TablesDockerGitGitKubernetesMicroservicesPower BIPysparkPythonSQLSQL ServerSsasSsis
Reposted 13 Hours AgoSaved
Remote
29 Locations
Senior level
Senior level
Artificial Intelligence • Fintech • Payments • Software
Build the companys first data stack and data products (internal analytics and customer-facing). Work across data engineering, devops, and software tasks; ship early features, improve developer experience, document proposals, and iterate based on feedback.
Top Skills: AWSAxiomCi/CdGitGoNext.JsTypescriptVercel
10 Days AgoSaved
In-Office or Remote
Warsaw, Warszawa, Mazowieckie, POL
Mid level
Mid level
Fintech • Financial Services
Build and maintain a centralized data platform, including SQL Server warehouses, SSIS processes, scalable SQL/Python/Spark pipelines, data integrations, CI/CD pipelines, and Infrastructure as Code. The role also involves improving legacy code, ensuring data quality and reliability, contributing to architecture and engineering standards, conducting code reviews, documenting solutions, and collaborating with team members.
Top Skills: Apache AirflowSparkChange Data Capture (Cdc)Ci/CdDagsterDatabricksDbtEltETLGrpcInfrastructure As CodeKafkaMs Sql ServerPub/SubPysparkPythonRestSQLSql Server AgentSsis
YesterdaySaved
Remote
27 Locations
Mid level
Mid level
Digital Media • Fintech • Gaming • Sports
Designs, develops, tests, optimizes, and maintains scalable batch and near-real-time data pipelines and architectures. Ensures data quality, builds API integrations, improves internal data processes, and supports machine learning, data science, BI, and product initiatives. The role requires experience with data warehouses, relational and NoSQL databases, data modeling, event-driven architectures, SQL, Python, Airflow, Spark, and AWS data services.
Top Skills: Amazon EksAmazon RdsAmazon RedshiftApache AirflowSparkAws AthenaAws Ec2Aws EmrAws LambdaAws S3DockerKubernetesNosql DatabasesPythonRelational DatabasesSQL
2 Days AgoSaved
Remote
5 Locations
Entry level
Entry level
Information Technology • Software • Consulting
Design and deliver cloud-native data platforms, scalable big data pipelines, lakehouses, warehouses, transformation frameworks, BI solutions, and ML-enabled data workflows. Lead client-facing pre-sales engagements, technical workshops, proof-of-concepts, executive conversations, and enterprise data modernization initiatives. Build reusable architectural IP, influence senior stakeholders, and translate ambiguous business challenges into production-ready solutions using AWS and modern data technologies.
Top Skills: Amazon AthenaAmazon BedrockAmazon EmrAmazon KinesisAmazon QuicksightAmazon RedshiftApache AirflowApache HadoopApache HudiApache IcebergApache KafkaAWSAws GlueAws Lake FormationCi/CdCollibraDatabricksDbtDelta LakeGitGlue CatalogHdfsHiveLookerMlflowPower BIPysparkPythonSagemakerScalaSnowflakeSparkSQLTableauUnity CatalogWeights & BiasesYarn
2 Days AgoSaved
Remote
Poland
Senior level
Senior level
Information Technology • Software • Business Intelligence
Develop and maintain scalable AWS data engineering platforms and reusable Python modules. Build object-oriented data processing solutions, debug pipelines across Step Functions, Glue, Athena, and Lambda, and apply modern engineering practices including testing, CI/CD, and version control. Collaborate with agile teams to deliver reliable solutions across countries and business environments.
Top Skills: Amazon AthenaAutomated TestingAWSAws GlueAws LambdaAws Step FunctionsAzure DevopsCi/CdDynamoDBGitPandasPython
2 Days AgoSaved
Remote
Poland
Entry level
Entry level
Information Technology • Software • Business Intelligence
Design and implement Azure-based data processing systems, scalable data pipelines, data warehouses, data lakes, and ETL workflows. Transform and optimize structured and unstructured data, tune queries, improve storage and retrieval performance, and resolve system bottlenecks. Collaborate with data scientists, analysts, and stakeholders to deliver technical solutions while working within Agile methodologies.
Top Skills: AgileAzureAzure DevopsDatabricksJIRAPysparkPythonSQL
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
8 Days AgoSaved
In-Office or Remote
30 Locations
177K-294K Annually
Senior level
177K-294K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Owns the design, development, operation, and governance of data pipelines and integration patterns supporting Medical Affairs AI products. Builds APIs and ETL/ELT solutions connecting enterprise systems to analytics platforms, RAG pipelines, vector databases, and GraphRAG applications. Ensures data quality, privacy, lineage, compliance, and protection of sensitive information. Partners with architecture, engineering, product, and business stakeholders to deliver reusable, production-grade data integrations.
Top Skills: Ai/MlApi GatewaysCi/CdEtl/EltEvent-Driven IntegrationGraph DatabasesGraphQLGraphragLow-Code/No-Code ToolsMiddlewareRagRestSalesforce Life Sciences/Marketing CloudSnowflakeSQLStreaming IntegrationVector DatabasesVeeva Crm
3 Days AgoSaved
Remote
Poland
Mid level
Mid level
Information Technology • Software • Business Intelligence
Designs, builds, and supports GCP data ecosystems, including cloud storage, processing, warehouses, data lakes, pipelines, and integrations. Responsibilities include requirements gathering, data modeling, performance optimization, troubleshooting, monitoring, root cause analysis, and client collaboration. The role also mentors junior data engineers and contributes to ETL/ELT, orchestration, transformation, and database management initiatives.
Top Skills: AlteryxApache AirflowSparkBigQueryCloud FunctionsCloud StorageConfluenceDatabricksDataformDbtGitGoogle Cloud PlatformJIRALookerAzureNoSQLPower BIPythonSQLTableauTalend
Reposted 3 Days AgoSaved
Remote
Poland
27-37 Hourly
Senior level
27-37 Hourly
Senior level
Information Technology • Software • Design
Build and maintain AWS-based data pipelines and lifecycles (Glue, DMS, Redshift, S3). Develop production data transformations and CDC using Scala and Python, automate CI/CD with Bash, optimize Spark jobs (partitioning, Parquet, broadcast joins), and deliver reliable, idempotent pipelines while collaborating remotely with cross-functional teams.
Top Skills: AWSBashCdcCi/CdDmsGlueParquetPythonRedshiftS3ScalaSpark
9 Days AgoSaved
In-Office or Remote
32 Locations
163K-272K Annually
Senior level
163K-272K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Designs, builds, and operates production-grade agentic AI systems and orchestration frameworks. Responsibilities include prompt architecture, tool and API integrations, monitoring, evaluation, cost controls, reliability improvements, and governance documentation. The engineer collaborates with data science, data engineering, and governance teams to ensure reliable, compliant workflows using structured healthcare and pharmaceutical data.
Top Skills: LanggraphLlm ApisMlopsPydantic Ai
Reposted 4 Days AgoSaved
Remote
3 Locations
Mid level
Mid level
Fintech • Payments
Design, build, and maintain AWS-based lakehouse solutions and medallion architecture using Iceberg and Glue. Implement CDC and streaming ingestion with Debezium and Kafka. Develop PySpark batch/distributed jobs on EMR/Glue, orchestrate workflows with Airflow, ensure data quality/governance, and collaborate on infrastructure and IaC practices to deliver reliable business-ready datasets.
Top Skills: AirbyteApache AirflowApache IcebergAws EmrAws Glue CatalogAws Glue JobsAws IamAws S3DagsterDebeziumDelta LakeKafkaMariadbMySQLPostgresPysparkPythonShell ScriptingSQLUnix/Linux
5 Days AgoSaved
Remote
4 Locations
Mid level
Mid level
Information Technology • Consulting
Build unstructured-data ingestion and context-layer pipelines for an investment group. Responsibilities include historical backfills, document parsing, entity resolution, deduplication, incremental API synchronization, embeddings, vector and graph storage, access scoping, lineage, PII protection, orchestration, monitoring, cost control, and operational runbooks.
Top Skills: AirflowAWSAws Step FunctionsComposioCrm ApisEmbedding PipelinesGCPGdprGoogle Workspace ApisGraph DatabasesMcpMicrosoft 365 ApisOpensearchPgvectorPineconePythonRagSlack ApisSQLVector Stores
5 Days AgoSaved
Remote
26 Locations
Mid level
Mid level
Cloud • Information Technology • Professional Services • Software • Consulting
Own the data engineering side of DoiT’s integrations framework by building and maintaining third-party billing and usage API integrations. Ensure data correctness, completeness, normalization, reconciliation, and support for negotiated rates. Manage orchestration, scheduling, backfills, and pipeline operations while using AI throughout development. Collaborate with product, support, and engineering teams to expand vendor coverage and improve cloud and SaaS spend visibility.
Top Skills: AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
8 Days AgoSaved
Remote
26 Locations
Mid level
Mid level
Cloud • Information Technology • Professional Services • Software • Consulting
Own data engineering for a cloud and SaaS billing integrations platform. Build and maintain third-party billing API integrations, operate ingestion pipelines, normalize vendor data, manage orchestration and backfills, and ensure financial-grade correctness through reconciliation and quality checks. Use AI throughout the engineering workflow and collaborate with product, support, and platform teams to expand vendor coverage and improve customer cost visibility.
Top Skills: AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
18 Days AgoSaved
In-Office
Warsaw, Warszawa, Mazowieckie, POL
Mid level
Mid level
Gaming • Mobile • Software
Build and maintain Python ETL pipelines, data warehouses, and data lakes supporting marketing analytics and personalized marketing programs. Integrate mobile marketing data, automate scalable and cost-efficient workflows, and ensure data quality, validation, governance, and performance. Collaborate with marketing and studio teams to deliver data-driven solutions. The role uses Snowflake or Firebolt, Airflow, SQL, distributed computing, and modern AI development tools.
Top Skills: Apache AirflowAWSCi/CdDevOpsFireboltPandasPolarsPythonSnowflakeSQL
9 Days AgoSaved
Remote or Hybrid
Poland
Senior level
Senior level
Information Technology • Software
Own the architecture, ingestion, governance, optimization, retention, and correctness of a large-scale financial data lakehouse. Design bronze, silver, and gold layers; manage CDC streaming through Kafka and Debezium; optimize TB-scale queries; ensure reconciliation, freshness, regulatory retention, GDPR deletion, lineage, access controls, PII masking, and infrastructure cost monitoring.
Top Skills: AirflowAmazon AthenaApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGcsKafkaPythonS3SnowflakeSQLTrino
9 Days AgoSaved
Remote or Hybrid
Poland
Mid level
Mid level
Information Technology • Software
Own the company data lakehouse supporting millions of financial events daily. Design bronze, silver, and gold architectures; build CDC streaming ingestion; optimize large-scale storage and queries; manage retention, archival, and GDPR deletion; ensure reconciliation, freshness, and data quality; and implement governance, access controls, PII masking, encryption, lineage, and auditability. Monitor pipeline health, anomalies, and cloud costs across S3 and Snowflake or Databricks.
Top Skills: Amazon AthenaApache AirflowApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGdprKafkaPythonS3SnowflakeSQLTrino
10 Days AgoSaved
Remote
3 Locations
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Software • Consulting • Data Privacy
Design and develop an AI-powered analytics agent in Devin, integrating GitHub, dbt, Databricks, Snowflake, and data ingestion systems. Adapt dbt workloads for Databricks, resolve platform compatibility issues, build and test agent capabilities, and implement guardrails, validation, observability, security, and reliable AI-assisted workflows. Collaborate with client engineering teams, troubleshoot integrations, document the solution, and prepare knowledge-transfer materials.
Top Skills: APIsAWSAzureCi/CdCliDatabricksDatabricks Asset BundlesDatabricks WorkflowsDbtDevinGCPGitGitMySQLSnowflakeSQLUnity Catalog
YesterdaySaved
In-Office
Warsaw, Warszawa, Mazowieckie, POL
Entry level
Entry level
Software
Perform manual functional, regression, integration, exploratory, and end-to-end testing across web applications, backend systems, REST APIs, databases, and integrations. Investigate defects using browser DevTools, network tools, proxies, logs, API tools, and SQL. Support release readiness, validate fixes, communicate risks, and improve QA processes. Collaborate cross-functionally while adapting testing approaches across products and technical environments.
Top Skills: Charles ProxyChatgptCi/CdClaude CodeCloud EnvironmentsDockerFiddlerGitHttp/HttpsKafkaKubernetesPostmanProxymanRelational DatabasesRest ApisSQLSwagger
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account