Top Data Engineer Jobs in Krakow

5 Days AgoSaved
In-Office or Remote
32 Locations
124K-207K Annually
Senior level
124K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills: Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL
YesterdaySaved
In-Office
2 Locations
Senior level
Senior level
Information Technology
Design and implement scalable data pipelines, ETL and data integration solutions, SQL-based transformations, and analytical data models. Optimize database and ETL performance, translate business requirements into robust data solutions, and collaborate in agile cross-functional teams on cloud, AI, and big-data initiatives. The role supports multiple experience levels, with increasing autonomy and responsibility, and requires travel to client locations across Europe.
Top Skills: AWSAzureDatabricksDb2GCPHbaseHdfsHiveIbm DatastageInformaticaKafkaKuduMicrosoft Sql ServerNifiOraclePythonRSASSas Data Integration StudioSnowflakeSparkSQLTeradata
Reposted YesterdaySaved
In-Office
2 Locations
Junior
Junior
Information Technology
Design and build scalable batch and streaming data pipelines using Databricks, Apache Spark, PySpark, Spark SQL, and Delta Lake. Develop Lakehouse and Medallion architecture data models, integrate cloud and enterprise data sources, optimize large-scale SQL workloads, and implement data quality, testing, CI/CD, DevOps, and DataOps practices. Collaborate with architects, analysts, and data scientists to deliver production-grade data products. The role is hybrid in Poland and requires travel to client locations across Europe.
Top Skills: SparkAuto LoaderAWSAzureCi/CdDatabricksDatabricks WorkflowsDbtDelta LakeDelta Live TablesEltETLFeature StoresGCPGitKafkaLakehouseMedallion ArchitectureMlflowPysparkPythonSpark SqlSQLStructured StreamingTerraformUnity CatalogVector Databases
Reposted 3 Days AgoSaved
In-Office
Kraków, Małopolskie, POL
Mid level
Mid level
Software
Build and manage Python-based data ingestion workflows and pipelines, integrating downstream systems such as ServiceNow. Use Snowflake, SQL, and OpenTelemetry to support data storage, validation, optimization, and observability. Troubleshoot pipeline failures, maintain data integrity, and improve workflow automation. Preferred experience includes orchestration tools such as Airflow, Prefect, or Dagster, ITSM and ITIL practices, and CI/CD pipelines for data products.
Top Skills: Apache AirflowCi/CdDagsterOpentelemetryPrefectPythonServicenowSnowflakeSQL
One Month AgoSaved
Hybrid
4 Locations
Senior level
Senior level
Artificial Intelligence • Automotive • Computer Vision • Information Technology • Internet of Things • Logistics • Software
Design and maintain scalable data pipelines, connectors, and reusable data products to make enterprise data AI-ready. Enable RAG, semantic search, agentic workflows, and production deployment of AI solutions while enforcing data quality, governance, metadata, security, and cross-team collaboration.
Top Skills: AirflowAWSAws BedrockAzure Ai FoundryAzure Data FactoryCopilot StudioDatabricksEmbeddingsJavaKafkaKubernetesM365 CopilotPythonRetrieval-Augmented Generation (Rag)ScalaSemantic SearchSnowflakeSQLVector Databases
5 Days AgoSaved
In-Office or Remote
32 Locations
163K-272K Annually
Senior level
163K-272K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Designs, builds, and operates production-grade agentic AI systems and orchestration frameworks. Responsibilities include prompt architecture, tool and API integrations, monitoring, evaluation, cost controls, reliability improvements, and governance documentation. The engineer collaborates with data science, data engineering, and governance teams to ensure reliable, compliant workflows using structured healthcare and pharmaceutical data.
Top Skills: LanggraphLlm ApisMlopsPydantic Ai
YesterdaySaved
Remote
4 Locations
Mid level
Mid level
Information Technology • Consulting
Build unstructured-data ingestion and context-layer pipelines for an investment group. Responsibilities include historical backfills, document parsing, entity resolution, deduplication, incremental API synchronization, embeddings, vector and graph storage, access scoping, lineage, PII protection, orchestration, monitoring, cost control, and operational runbooks.
Top Skills: AirflowAWSAws Step FunctionsComposioCrm ApisEmbedding PipelinesGCPGdprGoogle Workspace ApisGraph DatabasesMcpMicrosoft 365 ApisOpensearchPgvectorPineconePythonRagSlack ApisSQLVector Stores
YesterdaySaved
Remote
26 Locations
Mid level
Mid level
Cloud • Information Technology • Professional Services • Software • Consulting
Own the data engineering side of DoiT’s integrations framework by building and maintaining third-party billing and usage API integrations. Ensure data correctness, completeness, normalization, reconciliation, and support for negotiated rates. Manage orchestration, scheduling, backfills, and pipeline operations while using AI throughout development. Collaborate with product, support, and engineering teams to expand vendor coverage and improve cloud and SaaS spend visibility.
Top Skills: AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
4 Days AgoSaved
Remote
26 Locations
Mid level
Mid level
Cloud • Information Technology • Professional Services • Software • Consulting
Own data engineering for a cloud and SaaS billing integrations platform. Build and maintain third-party billing API integrations, operate ingestion pipelines, normalize vendor data, manage orchestration and backfills, and ensure financial-grade correctness through reconciliation and quality checks. Use AI throughout the engineering workflow and collaborate with product, support, and platform teams to expand vendor coverage and improve customer cost visibility.
Top Skills: AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
Reposted 5 Days AgoSaved
Remote
29 Locations
Senior level
Senior level
Artificial Intelligence • Fintech • Payments • Software
Build the companys first data stack and data products (internal analytics and customer-facing). Work across data engineering, devops, and software tasks; ship early features, improve developer experience, document proposals, and iterate based on feedback.
Top Skills: AWSAxiomCi/CdGitGoNext.JsTypescriptVercel
5 Days AgoSaved
Remote or Hybrid
Poland
Senior level
Senior level
Information Technology • Software
Own the architecture, ingestion, governance, optimization, retention, and correctness of a large-scale financial data lakehouse. Design bronze, silver, and gold layers; manage CDC streaming through Kafka and Debezium; optimize TB-scale queries; ensure reconciliation, freshness, regulatory retention, GDPR deletion, lineage, access controls, PII masking, and infrastructure cost monitoring.
Top Skills: AirflowAmazon AthenaApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGcsKafkaPythonS3SnowflakeSQLTrino
5 Days AgoSaved
Remote or Hybrid
Poland
Mid level
Mid level
Information Technology • Software
Own the company data lakehouse supporting millions of financial events daily. Design bronze, silver, and gold architectures; build CDC streaming ingestion; optimize large-scale storage and queries; manage retention, archival, and GDPR deletion; ensure reconciliation, freshness, and data quality; and implement governance, access controls, PII masking, encryption, lineage, and auditability. Monitor pipeline health, anomalies, and cloud costs across S3 and Snowflake or Databricks.
Top Skills: Amazon AthenaApache AirflowApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGdprKafkaPythonS3SnowflakeSQLTrino
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
6 Days AgoSaved
Remote
3 Locations
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Software • Consulting • Data Privacy
Design and develop an AI-powered analytics agent in Devin, integrating GitHub, dbt, Databricks, Snowflake, and data ingestion systems. Adapt dbt workloads for Databricks, resolve platform compatibility issues, build and test agent capabilities, and implement guardrails, validation, observability, security, and reliable AI-assisted workflows. Collaborate with client engineering teams, troubleshoot integrations, document the solution, and prepare knowledge-transfer materials.
Top Skills: APIsAWSAzureCi/CdCliDatabricksDatabricks Asset BundlesDatabricks WorkflowsDbtDevinGCPGitGitMySQLSnowflakeSQLUnity Catalog
6 Days AgoSaved
Remote
PL
Senior level
Senior level
Artificial Intelligence • Information Technology • Machine Learning • Software • Virtual Reality • Analytics
Lead the architecture and scaling of high-throughput data services, ingestion pipelines, and distributed data platforms. Build low-latency, memory-efficient infrastructure in Rust; design resilient APIs, schemas, and data contracts; integrate services with cloud warehouses and lakehouses; and optimize distributed processing, storage, caching, and query performance.
Top Skills: Amazon RedshiftAvroAWSAws GlueAzureDatabricksDelta LakeGCPGoogle BigqueryGrpcOpenapiParquetProtobufRustSnowflake
Reposted 22 Days AgoSaved
Easy Apply
Hybrid
Kraków, Małopolskie, POL
Easy Apply
186-26K Hourly
Senior level
186-26K Hourly
Senior level
AdTech • Enterprise Web • Information Technology • Machine Learning • Marketing Tech • Sales
Design, build, and maintain large-scale data processing systems and pipelines (Spark, BigQuery, dbt/Dataform). Improve scalability, performance, and reliability; own projects end-to-end; extend and operate Airflow; work with cloud (GCP/AWS), Kubernetes, and CI/CD practices.
Top Skills: AirflowSparkAWSBigQueryCi/CdDataformDbtDockerGoogle Cloud Platform (Gcp)JavaKubernetesNoSQLPythonRdbmsScalaSQL
19 Days AgoSaved
In-Office or Remote
Kraków, Małopolskie, POL
Junior
Junior
Information Technology
Develop and support modern data pipelines and Lakehouse architectures using Azure Databricks, Spark, PySpark, SQL, and Azure Data Factory. Maintain SQL Server data platforms, ETL solutions, data models, and warehouse structures while integrating multiple data sources. Ensure data quality, scalability, and performance, and collaborate with BI teams supporting Power BI and reporting platforms. Participate in cloud data architecture initiatives under senior guidance.
Top Skills: SparkAzure Data FactoryAzure DatabricksDelta LakePower BIPysparkSpark SqlSQLSQL ServerStructured Streaming
11 Days AgoSaved
Remote
Poland
Senior level
Senior level
Information Technology • Software • Business Intelligence
Develop and maintain scalable AWS data pipelines, modify Lambda functions and SQL processing jobs, design Iceberg-based data models, optimize Athena and Spark workloads, and implement data quality, observability, testing, governance, and documentation practices. The role also requires collaboration with engineering, product, and data stakeholders, along with GitLab-based version control and CI/CD or Infrastructure as Code familiarity.
Top Skills: Amazon AuroraAmazon EventbridgeAmazon S3Amazon SagemakerApache AirflowApache IcebergSparkAWSAws AthenaAws GlueAws LambdaCi/CdDbtGitlabInfrastructure As CodePythonSQLTest-Driven Development
YesterdaySaved
In-Office
2 Locations
Senior level
Senior level
Information Technology
Design and implement scalable data pipelines, ETL processes, data integration solutions, and analytical data models. Work with GCP services, SQL, dbt, Python, and Data Vault 2.0 in agile, cross-functional AI teams. Support scalable AI/ML initiatives, collaborate with technical and business stakeholders, and travel to client locations across Europe or occasionally beyond as project needs require.
Top Skills: BigQueryCi/CdCloud ComposerData Vault 2.0DataprocDbtDockerGitlabGoogle Cloud PlatformGoogle Cloud StorageKubernetesPysparkPythonScalaSparkSQL
Reposted 24 Days AgoSaved
In-Office
Kraków, Małopolskie, POL
Senior level
Senior level
Big Data • Cloud • Database
Design, build, and operate GitOps-driven data and AI platforms and pipelines (ETL/ELT, feature stores, batch/real-time inference). Enable GenAI/RAG (embeddings, vector ingestion), implement IaC and CI/CD, enforce data quality, governance, observability and MLOps, and support Azure/AWS-based production deployments while collaborating with security, product, and operations teams.
Top Skills: Ai SearchAnthropic ClaudeApplication InsightsAWSAzureAzure Container RegistryAzure DevopsAzure Kubernetes ServiceAzure MlAzure Ml Feature StoreAzure Ml Model RegistryAzure MonitorAzure OpenaiAzure StorageDatabricksDatabricks Feature StoreDrataEntra IdGitGoogle GeminiIntuneJSONMl MonitoringNetSuiteOpenaiPythonSalesforceShell ScriptingSQLTerraformYaml
One Month AgoSaved
Hybrid
Kraków, Małopolskie, POL
Senior level
Senior level
Cloud • Fintech • Information Technology • Machine Learning • Software • App development • Generative AI
Design, build, and maintain ELT/ETL pipelines into Snowflake, model data for analytics, deliver Power BI reports, embed data quality and observability, and use AI tools to accelerate development while collaborating across engineering, product, and finance.
Top Skills: Apache AirflowAzure Data FactoryClaudeCursorDaxDbtPower BIPower QueryPythonSnowflakeSQL
One Month AgoSaved
Hybrid
Kraków, Małopolskie, POL
Senior level
Senior level
Cloud • Fintech • Information Technology • Machine Learning • Software • App development • Generative AI
Design, build, and maintain ELT/ETL pipelines into Snowflake, produce governed Power BI analytics, implement data quality/observability, optimize data models for performance and cost, and use AI tools to accelerate development while collaborating across product and finance teams.
Top Skills: Apache AirflowAws GlueAzure Data FactoryAzure Data LakeAzure Event HubsAzure SynapseClaudeCursorDaxDbtKafkaMicrosoft FabricPower BIPower QueryPythonS3SnowflakeSQL
4 Days AgoSaved
In-Office or Remote
Kraków, Małopolskie, POL
Senior level
Senior level
Software
Lead architecture and technical standards for a greenfield, enterprise-scale data platform supporting analytics, operational products, and AI workloads. Design Iceberg-based event flows, canonical data models, batch and near-real-time pipelines, governance, data quality, observability, tenant isolation, and regional data strategies. Evaluate platform technologies, optimize performance and costs, mentor engineers, and collaborate with leadership and platform teams to deliver a production-ready foundation.
Top Skills: Apache FlinkApache IcebergSparkAWSCi/CdInfrastructure As CodeJavaPythonScalaSQLTrino
27 Days AgoSaved
In-Office
Kraków, Małopolskie, POL
18K-27K Annually
Senior level
18K-27K Annually
Senior level
Fintech • Payments • Financial Services
Build and operate production data pipelines covering ingestion, transformation, orchestration, monitoring, and data quality. Develop dbt models, author and debug Airflow DAGs, provision AWS infrastructure with infrastructure-as-code, and maintain CI/CD using GitHub Actions. Support AI-ready data initiatives and Redshift access, while owning requirements through deployment and incident resolution. The role requires strong Python, SQL, AWS, analytical data modeling, autonomy, and stakeholder collaboration.
Top Skills: Amazon EcrAmazon EcsAmazon KinesisAmazon RedshiftAmazon S3Apache AirflowAWSAws CdkAws CloudformationAws GlueAws IamAws LambdaCi/CdDbtGithub ActionsMcpPythonSQLTerraform
Reposted 27 Days AgoSaved
In-Office
Kraków, Małopolskie, POL
Expert/Leader
Expert/Leader
Automotive
Design, implement, and optimize ETL/data pipelines, data quality metrics, APIs, and enrichment for AI/ML platforms; research big-data technologies; run trials, write technical documentation, and collaborate in project teams.
5 Days AgoSaved
In-Office
Kraków, Małopolskie, POL
Mid level
Mid level
Automotive
Develops and maintains Microsoft Lists, Power Automate flows, Excel-based analysis tools, and Power BI semantic models, reports, dashboards, and paginated reports. Responsibilities include automating data processing, designing linked lists, improving data-entry workflows, manipulating multi-source data with Power Query, creating DAX measures, managing relationships, and supporting end-user customization of Power BI visuals. Requires a relevant bachelor’s degree and 3–4 years of experience developing data tools for electronics components.
Top Skills: CsvDaxGoogle BigqueryHTMLExcelMicrosoft ListsPower AutomatePower BIPower QuerySharepoint
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account