Maximum of 25 job preferences reached.
Top Data Engineer Jobs in Krakow
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills:
Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL
Information Technology
Design and implement scalable data pipelines, ETL and data integration solutions, SQL-based transformations, and analytical data models. Optimize database and ETL performance, translate business requirements into robust data solutions, and collaborate in agile cross-functional teams on cloud, AI, and big-data initiatives. The role supports multiple experience levels, with increasing autonomy and responsibility, and requires travel to client locations across Europe.
Top Skills:
AWSAzureDatabricksDb2GCPHbaseHdfsHiveIbm DatastageInformaticaKafkaKuduMicrosoft Sql ServerNifiOraclePythonRSASSas Data Integration StudioSnowflakeSparkSQLTeradata
Information Technology
Design and build scalable batch and streaming data pipelines using Databricks, Apache Spark, PySpark, Spark SQL, and Delta Lake. Develop Lakehouse and Medallion architecture data models, integrate cloud and enterprise data sources, optimize large-scale SQL workloads, and implement data quality, testing, CI/CD, DevOps, and DataOps practices. Collaborate with architects, analysts, and data scientists to deliver production-grade data products. The role is hybrid in Poland and requires travel to client locations across Europe.
Top Skills:
SparkAuto LoaderAWSAzureCi/CdDatabricksDatabricks WorkflowsDbtDelta LakeDelta Live TablesEltETLFeature StoresGCPGitKafkaLakehouseMedallion ArchitectureMlflowPysparkPythonSpark SqlSQLStructured StreamingTerraformUnity CatalogVector Databases
Software
Build and manage Python-based data ingestion workflows and pipelines, integrating downstream systems such as ServiceNow. Use Snowflake, SQL, and OpenTelemetry to support data storage, validation, optimization, and observability. Troubleshoot pipeline failures, maintain data integrity, and improve workflow automation. Preferred experience includes orchestration tools such as Airflow, Prefect, or Dagster, ITSM and ITIL practices, and CI/CD pipelines for data products.
Top Skills:
Apache AirflowCi/CdDagsterOpentelemetryPrefectPythonServicenowSnowflakeSQL
Artificial Intelligence • Automotive • Computer Vision • Information Technology • Internet of Things • Logistics • Software
Design and maintain scalable data pipelines, connectors, and reusable data products to make enterprise data AI-ready. Enable RAG, semantic search, agentic workflows, and production deployment of AI solutions while enforcing data quality, governance, metadata, security, and cross-team collaboration.
Top Skills:
AirflowAWSAws BedrockAzure Ai FoundryAzure Data FactoryCopilot StudioDatabricksEmbeddingsJavaKafkaKubernetesM365 CopilotPythonRetrieval-Augmented Generation (Rag)ScalaSemantic SearchSnowflakeSQLVector Databases
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Designs, builds, and operates production-grade agentic AI systems and orchestration frameworks. Responsibilities include prompt architecture, tool and API integrations, monitoring, evaluation, cost controls, reliability improvements, and governance documentation. The engineer collaborates with data science, data engineering, and governance teams to ensure reliable, compliant workflows using structured healthcare and pharmaceutical data.
Top Skills:
LanggraphLlm ApisMlopsPydantic Ai
Information Technology • Consulting
Build unstructured-data ingestion and context-layer pipelines for an investment group. Responsibilities include historical backfills, document parsing, entity resolution, deduplication, incremental API synchronization, embeddings, vector and graph storage, access scoping, lineage, PII protection, orchestration, monitoring, cost control, and operational runbooks.
Top Skills:
AirflowAWSAws Step FunctionsComposioCrm ApisEmbedding PipelinesGCPGdprGoogle Workspace ApisGraph DatabasesMcpMicrosoft 365 ApisOpensearchPgvectorPineconePythonRagSlack ApisSQLVector Stores
Cloud • Information Technology • Professional Services • Software • Consulting
Own the data engineering side of DoiT’s integrations framework by building and maintaining third-party billing and usage API integrations. Ensure data correctness, completeness, normalization, reconciliation, and support for negotiated rates. Manage orchestration, scheduling, backfills, and pipeline operations while using AI throughout development. Collaborate with product, support, and engineering teams to expand vendor coverage and improve cloud and SaaS spend visibility.
Top Skills:
AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
Cloud • Information Technology • Professional Services • Software • Consulting
Own data engineering for a cloud and SaaS billing integrations platform. Build and maintain third-party billing API integrations, operate ingestion pipelines, normalize vendor data, manage orchestration and backfills, and ensure financial-grade correctness through reconciliation and quality checks. Use AI throughout the engineering workflow and collaborate with product, support, and platform teams to expand vendor coverage and improve customer cost visibility.
Top Skills:
AirflowAWSBigQueryClickhouseDagsterDbtGoGCPAzurePrefectPythonRedshiftRest ApisSnowflakeSQL
Artificial Intelligence • Fintech • Payments • Software
Build the companys first data stack and data products (internal analytics and customer-facing). Work across data engineering, devops, and software tasks; ship early features, improve developer experience, document proposals, and iterate based on feedback.
Top Skills:
AWSAxiomCi/CdGitGoNext.JsTypescriptVercel
Information Technology • Software
Own the architecture, ingestion, governance, optimization, retention, and correctness of a large-scale financial data lakehouse. Design bronze, silver, and gold layers; manage CDC streaming through Kafka and Debezium; optimize TB-scale queries; ensure reconciliation, freshness, regulatory retention, GDPR deletion, lineage, access controls, PII masking, and infrastructure cost monitoring.
Top Skills:
AirflowAmazon AthenaApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGcsKafkaPythonS3SnowflakeSQLTrino
Information Technology • Software
Own the company data lakehouse supporting millions of financial events daily. Design bronze, silver, and gold architectures; build CDC streaming ingestion; optimize large-scale storage and queries; manage retention, archival, and GDPR deletion; ensure reconciliation, freshness, and data quality; and implement governance, access controls, PII masking, encryption, lineage, and auditability. Monitor pipeline health, anomalies, and cloud costs across S3 and Snowflake or Databricks.
Top Skills:
Amazon AthenaApache AirflowApache HudiApache IcebergBigQueryDagsterDatabricksDebeziumDelta LakeGdprKafkaPythonS3SnowflakeSQLTrino
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Artificial Intelligence • Cloud • Information Technology • Software • Consulting • Data Privacy
Design and develop an AI-powered analytics agent in Devin, integrating GitHub, dbt, Databricks, Snowflake, and data ingestion systems. Adapt dbt workloads for Databricks, resolve platform compatibility issues, build and test agent capabilities, and implement guardrails, validation, observability, security, and reliable AI-assisted workflows. Collaborate with client engineering teams, troubleshoot integrations, document the solution, and prepare knowledge-transfer materials.
Top Skills:
APIsAWSAzureCi/CdCliDatabricksDatabricks Asset BundlesDatabricks WorkflowsDbtDevinGCPGitGitMySQLSnowflakeSQLUnity Catalog
Artificial Intelligence • Information Technology • Machine Learning • Software • Virtual Reality • Analytics
Lead the architecture and scaling of high-throughput data services, ingestion pipelines, and distributed data platforms. Build low-latency, memory-efficient infrastructure in Rust; design resilient APIs, schemas, and data contracts; integrate services with cloud warehouses and lakehouses; and optimize distributed processing, storage, caching, and query performance.
Top Skills:
Amazon RedshiftAvroAWSAws GlueAzureDatabricksDelta LakeGCPGoogle BigqueryGrpcOpenapiParquetProtobufRustSnowflake
AdTech • Enterprise Web • Information Technology • Machine Learning • Marketing Tech • Sales
Design, build, and maintain large-scale data processing systems and pipelines (Spark, BigQuery, dbt/Dataform). Improve scalability, performance, and reliability; own projects end-to-end; extend and operate Airflow; work with cloud (GCP/AWS), Kubernetes, and CI/CD practices.
Top Skills:
AirflowSparkAWSBigQueryCi/CdDataformDbtDockerGoogle Cloud Platform (Gcp)JavaKubernetesNoSQLPythonRdbmsScalaSQL
Information Technology
Develop and support modern data pipelines and Lakehouse architectures using Azure Databricks, Spark, PySpark, SQL, and Azure Data Factory. Maintain SQL Server data platforms, ETL solutions, data models, and warehouse structures while integrating multiple data sources. Ensure data quality, scalability, and performance, and collaborate with BI teams supporting Power BI and reporting platforms. Participate in cloud data architecture initiatives under senior guidance.
Top Skills:
SparkAzure Data FactoryAzure DatabricksDelta LakePower BIPysparkSpark SqlSQLSQL ServerStructured Streaming
Information Technology • Software • Business Intelligence
Develop and maintain scalable AWS data pipelines, modify Lambda functions and SQL processing jobs, design Iceberg-based data models, optimize Athena and Spark workloads, and implement data quality, observability, testing, governance, and documentation practices. The role also requires collaboration with engineering, product, and data stakeholders, along with GitLab-based version control and CI/CD or Infrastructure as Code familiarity.
Top Skills:
Amazon AuroraAmazon EventbridgeAmazon S3Amazon SagemakerApache AirflowApache IcebergSparkAWSAws AthenaAws GlueAws LambdaCi/CdDbtGitlabInfrastructure As CodePythonSQLTest-Driven Development
Information Technology
Design and implement scalable data pipelines, ETL processes, data integration solutions, and analytical data models. Work with GCP services, SQL, dbt, Python, and Data Vault 2.0 in agile, cross-functional AI teams. Support scalable AI/ML initiatives, collaborate with technical and business stakeholders, and travel to client locations across Europe or occasionally beyond as project needs require.
Top Skills:
BigQueryCi/CdCloud ComposerData Vault 2.0DataprocDbtDockerGitlabGoogle Cloud PlatformGoogle Cloud StorageKubernetesPysparkPythonScalaSparkSQL
Big Data • Cloud • Database
Design, build, and operate GitOps-driven data and AI platforms and pipelines (ETL/ELT, feature stores, batch/real-time inference). Enable GenAI/RAG (embeddings, vector ingestion), implement IaC and CI/CD, enforce data quality, governance, observability and MLOps, and support Azure/AWS-based production deployments while collaborating with security, product, and operations teams.
Top Skills:
Ai SearchAnthropic ClaudeApplication InsightsAWSAzureAzure Container RegistryAzure DevopsAzure Kubernetes ServiceAzure MlAzure Ml Feature StoreAzure Ml Model RegistryAzure MonitorAzure OpenaiAzure StorageDatabricksDatabricks Feature StoreDrataEntra IdGitGoogle GeminiIntuneJSONMl MonitoringNetSuiteOpenaiPythonSalesforceShell ScriptingSQLTerraformYaml
Cloud • Fintech • Information Technology • Machine Learning • Software • App development • Generative AI
Design, build, and maintain ELT/ETL pipelines into Snowflake, model data for analytics, deliver Power BI reports, embed data quality and observability, and use AI tools to accelerate development while collaborating across engineering, product, and finance.
Top Skills:
Apache AirflowAzure Data FactoryClaudeCursorDaxDbtPower BIPower QueryPythonSnowflakeSQL
Cloud • Fintech • Information Technology • Machine Learning • Software • App development • Generative AI
Design, build, and maintain ELT/ETL pipelines into Snowflake, produce governed Power BI analytics, implement data quality/observability, optimize data models for performance and cost, and use AI tools to accelerate development while collaborating across product and finance teams.
Top Skills:
Apache AirflowAws GlueAzure Data FactoryAzure Data LakeAzure Event HubsAzure SynapseClaudeCursorDaxDbtKafkaMicrosoft FabricPower BIPower QueryPythonS3SnowflakeSQL
Software
Lead architecture and technical standards for a greenfield, enterprise-scale data platform supporting analytics, operational products, and AI workloads. Design Iceberg-based event flows, canonical data models, batch and near-real-time pipelines, governance, data quality, observability, tenant isolation, and regional data strategies. Evaluate platform technologies, optimize performance and costs, mentor engineers, and collaborate with leadership and platform teams to deliver a production-ready foundation.
Top Skills:
Apache FlinkApache IcebergSparkAWSCi/CdInfrastructure As CodeJavaPythonScalaSQLTrino
Fintech • Payments • Financial Services
Build and operate production data pipelines covering ingestion, transformation, orchestration, monitoring, and data quality. Develop dbt models, author and debug Airflow DAGs, provision AWS infrastructure with infrastructure-as-code, and maintain CI/CD using GitHub Actions. Support AI-ready data initiatives and Redshift access, while owning requirements through deployment and incident resolution. The role requires strong Python, SQL, AWS, analytical data modeling, autonomy, and stakeholder collaboration.
Top Skills:
Amazon EcrAmazon EcsAmazon KinesisAmazon RedshiftAmazon S3Apache AirflowAWSAws CdkAws CloudformationAws GlueAws IamAws LambdaCi/CdDbtGithub ActionsMcpPythonSQLTerraform
Automotive
Design, implement, and optimize ETL/data pipelines, data quality metrics, APIs, and enrichment for AI/ML platforms; research big-data technologies; run trials, write technical documentation, and collaborate in project teams.
Automotive
Develops and maintains Microsoft Lists, Power Automate flows, Excel-based analysis tools, and Power BI semantic models, reports, dashboards, and paginated reports. Responsibilities include automating data processing, designing linked lists, improving data-entry workflows, manipulating multi-source data with Power Query, creating DAX measures, managing relationships, and supporting end-user customization of Power BI visuals. Requires a relevant bachelor’s degree and 3–4 years of experience developing data tools for electronics components.
Top Skills:
CsvDaxGoogle BigqueryHTMLExcelMicrosoft ListsPower AutomatePower BIPower QuerySharepoint
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in Krakow Hiring Data Engineers
See AllPopular Job Searches
Tech Jobs & Startup Jobs in Poland
Software Engineer Jobs in Poland
Data Science Jobs in Poland
Machine Learning Jobs in Poland
Artificial Intelligence Jobs in Poland
Product Manager Jobs in Poland
Front End Developer Jobs in Poland
QA Engineer Jobs in Poland
Cybersecurity Jobs in Poland
IT Jobs in Poland
Data Engineer Jobs in Poland
Data Analyst Jobs in Poland
UX Designer Jobs in Poland
Graphic Designer Jobs in Poland
Finance Jobs in Poland
HR Jobs in Poland
Marketing Jobs in Poland
Operations Manager Jobs in Poland
Project Manager Jobs in Poland
Sales Jobs in Poland
Tech Jobs & Startup Jobs in Krakow
Software Engineer Jobs in Krakow
Data Science Jobs in Krakow
Machine Learning Jobs in Krakow
Artificial Intelligence Jobs in Krakow
Product Manager Jobs in Krakow
Front End Developer Jobs in Krakow
QA Engineer Jobs in Krakow
Cybersecurity Jobs in Krakow
IT Jobs in Krakow
Data Engineer Jobs in Krakow
Data Analyst Jobs in Krakow
UX Designer Jobs in Krakow
Graphic Designer Jobs in Krakow
Finance Jobs in Krakow
HR Jobs in Krakow
Marketing Jobs in Krakow
Operations Manager Jobs in Krakow
Project Manager Jobs in Krakow
Sales Jobs in Krakow
Tech Jobs & Startup Jobs in Warsaw
Software Engineer Jobs in Warsaw
Data Science Jobs in Warsaw
Machine Learning Jobs in Warsaw
Artificial Intelligence Jobs in Warsaw
Product Manager Jobs in Warsaw
Front End Developer Jobs in Warsaw
QA Engineer Jobs in Warsaw
Cybersecurity Jobs in Warsaw
IT Jobs in Warsaw
Data Engineer Jobs in Warsaw
Data Analyst Jobs in Warsaw
UX Designer Jobs in Warsaw
Graphic Designer Jobs in Warsaw
Finance Jobs in Warsaw
HR Jobs in Warsaw
Marketing Jobs in Warsaw
Operations Manager Jobs in Warsaw
Project Manager Jobs in Warsaw
Sales Jobs in Warsaw
Remote Jobs in Poland
Remote Jobs in Warsaw
Remote Jobs in Krakow
All Filters
Total selected ()
No Results
No Results

























