Maximum of 25 job preferences reached.
Top Data Engineer Jobs in Cleveland, OH
Artificial Intelligence • Information Technology • Software
Design, build, and migrate large-scale batch data pipelines into Databricks. Develop and optimize production-grade Python/Spark/SQL solutions, support parallel legacy and cloud systems during migration, lead technical design discussions, use AI tools to improve workflows, ensure pipeline reliability through testing/monitoring/tuning, and participate in a rotating on-call schedule.
Top Skills:
Anthropic ClaudeSparkAzure Data FactoryAzure DevopsCi/CdCosmos DbDatabricksDelta LakeEvent HubsGitGithub CopilotJenkinsNoSQLPythonSQLSynapse AnalyticsTeradata
HR Tech • Other • Professional Services
Own design and implementation of a greenfield analytical data platform: ingesting PostgreSQL and Kafka streams, establishing governance, data lineage, quality, monitoring, and compliance-ready pipelines. Evaluate AWS-native tooling, Snowflake, and Databricks; produce RFCs and documentation; mentor junior engineers and ensure analytical systems remain downstream of transactional sources.
Top Skills:
AWSAws MskBlockchainDatabricksKafkaPostgresSnowflakeTerraform
Fitness • Healthtech • Retail • Pharmaceutical
The Senior Data Engineer will design and maintain scalable cloud-native data pipelines, develop APIs, implement advanced analytics, manage AI workflows, and enhance data operations collaboratively.
Top Skills:
AirflowAWSAws GlueAzureAzure Data FactoryBigQueryCi/CdCloud FunctionsDataflowDevOpsFastapiFlaskGCPPub/SubPythonSQLVertex Ai
Artificial Intelligence • Healthtech • Machine Learning
Lead design and scaling of backend data infrastructure for Beacon's Datastore: build event-driven pipelines, data models, and APIs; draft RFCs; integrate services; improve operational robustness, documentation, and tooling to support clinical and scientific workflows.
Top Skills:
Amazon KinesisHelmJavaScriptJuliaKafkaKubernetesNode.jsPostgresPulsarPythonRRabbitMQSQLTerraformTypescript
Reposted 5 Days AgoSaved
Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
The Senior Data Engineer will design and implement data processing pipelines using Java and Spark, optimize performance, and mentor junior engineers while ensuring compliance with data governance standards.
Top Skills:
SparkAWSCi/CdDockerGitJavaKubernetesRestful ApisSpring Boot
Fintech • Information Technology • Payments • Financial Services
Founding data engineer to design and build Payabli's data platform: architect lakehouse/warehouse, build batch and streaming pipelines, model canonical datasets, ensure data quality/observability, enforce access/masking/lineage for regulated financial data, and enable analytics/ML feature pipelines while establishing team standards and CI/CD.
Top Skills:
AirflowAWSAzureCdcDagsterDatabricksDbtDelta LakeEltFeature StoreFivetranFlinkGCPGreat ExpectationsKafkaKinesisMlflowMonte CarloOpenlineagePci DssPrefectPythonSnowflakeSparkSQLUnity Catalog
Insurance • Software
Build and maintain scalable ETL/data pipelines and internal frameworks to support ML and product teams. Drive architecture, observability, data quality, and engineering best practices across AI/ML workflows and prompt-engineering pipelines.
Top Skills:
AirflowAWSCi/CdDagsterDbtGCPPrefectPythonSQLTerraform
Edtech
Lead design, build, and operate scalable Databricks/Delta Lake data pipelines and Kimball dimensional models using dbt. Manage governance with Unity Catalog, optimize performance and cost, operationalize ML with MLflow, and coordinate/quality-check offshore vendor engineering. Mentor engineers, uphold standards, and deliver trusted data products that power analytics, reporting, and MLOps workflows.
Top Skills:
Azure Event HubsDatabricksDatabricks Dashboards (Genie)DbtDelta LakeDelta LakehouseDockerKafkaLakeflow ConnectMlflowMlopsPower BIPysparkPythonSalesforceSparkSQLStructured StreamingTableauUnity Catalog
Healthtech
Design, build, and maintain dbt data models and scalable ELT pipelines; optimize performance-aware SQL on Snowflake; ensure observability, testing, and CI/CD; partner with product and engineering teams to meet analytic needs, mentor engineers, document business logic, and troubleshoot ingestion, transformation, and warehouse issues.
Top Skills:
Aws LambdaAws S3Aws SnsAws SqsCi/CdDbtEltFhirFivetranGitHevoOrchestration ToolsPythonSnowflakeSQL
Healthtech • Professional Services • Software
Design, build, and maintain scalable data pipelines, data models, ETL/ELT, and BI solutions. Collaborate with data scientists and stakeholders to improve data quality, accessibility, and PHI security. Develop and document workflows, optimize processing/storage, support data warehouse and visualization, and adopt AI tools to drive data-informed outcomes.
Top Skills:
Ai ToolsBigQueryCdcCloud SqlDatastreamDbtGCPHvrSnowflakeSQLSQL ServerSsisTableau
Information Technology • Consulting
Design, build, deploy, and maintain large-scale analytics solutions and ETL pipelines. Implement real-time streaming and batch processing, optimize SQL and data flows, ensure data quality, contribute to dev-ops and QA, cross-train team members, and participate in community and best-practice activities.
Top Skills:
AWSAzureCassandraGCPHadoopHbaseHiveImpalaJavaKafkaMapreducePigPower BIPythonScalaSparkSQLStormTableau
Information Technology • Consulting
Design, build, deploy and maintain large-scale analytics and ETL solutions. Manage data ingestion, real-time streaming and batch processing across data stores, tune SQL and data flows, ensure quality, performance and maintainability, participate in QA/devops, cross-train team members, and contribute to community and best practices.
Top Skills:
AWSAzureCassandraGCPHadoopHbaseHiveImpalaJavaKafkaMapreducePigPower BIPythonScalaSparkSQLStormTableau
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Healthtech
Design, build, and optimize scalable cloud data pipelines and infrastructure. Maintain and refactor SQL and ETL processes, automate data workflows, ensure data quality/security, and troubleshoot complex issues. Provide technical leadership and mentor junior engineers while collaborating with stakeholders to translate requirements.
Top Skills:
APIsAWSAzureData ModelingData WarehousingDistributed SystemsEltETLGCPJavaKotlinPythonScalaSQLVersion Control
Information Technology • Database • Consulting
Design, build, and operate scalable data transformation pipelines and dimensional models using dbt, SQL, and AWS. Implement infrastructure-as-code (AWS CDK), CI/CD, and query performance optimizations while ensuring data quality, governance, and collaboration with cross-functional and offshore teams.
Top Skills:
Amazon MwaaAmazon RedshiftAws CdkAws CodecommitAws CodepipelineAws GlueAws LambdaAws Step FunctionsCi/CdDbtJinjaPysparkPythonSQLVersion Control
Information Technology • Database • Consulting
Lead Data Engineer builds and maintains scalable data pipelines and lakehouse/warehouse platforms (Databricks) to support analytics, BI, reporting, and AI. Responsibilities include data modeling, ETL/ELT, medallion architecture, orchestration and monitoring (Airflow/Cron), performance tuning, CDC/incremental loads, data quality, and leading/mentoring a team while collaborating with stakeholders and maintaining documentation.
Top Skills:
Apache AirflowAWSAzureCronData WarehousingDatabricksEltETLGCPHadoopHbaseHiveLakehouseMedallion ArchitecturePigPower BIPysparkPythonSparkSQLTableau
Healthtech
Own and build scalable ELT pipelines and event-driven ingestion in AWS, using Dagster for orchestration. Implement IaC (Terraform), data quality, lineage, monitoring, and troubleshoot cross-system issues while partnering with analytics and product teams to ensure downstream analytical correctness.
Top Skills:
APIsAws LambdaAws S3Aws SnsAws SqsCdcCi/CdDagsterGoJavaScriptPythonSnowflakeSQLTerraformTypescriptVersion ControlWebhooks
Agency • Information Technology
Design, build, and maintain scalable ETL/ELT pipelines into Snowflake, enforce data quality, optimize data models and query performance for Power BI reporting, produce documentation, and troubleshoot data issues while collaborating with stakeholders.
Top Skills:
AWSAzureEltETLGCPOraclePower BISnowflakeSQL
Agency • Information Technology
Design, implement, and maintain PySpark-based data reconciliation solutions for financial systems. Build matching algorithms, integrate with rules engines, process large distributed datasets, identify and resolve discrepancies, and collaborate with analysts and architects to improve data quality and governance.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisPyspark
Agency • Information Technology
Design, build, and optimize PySpark applications and ETL pipelines to process large-scale datasets from SQL/NoSQL sources, data lakes, and streaming platforms. Ensure data quality, error handling, performance tuning, and collaborate with analysts, scientists, and architects to deliver scalable data solutions.
Top Skills:
Apache AirflowSparkData LakeETLLuigiNoSQLPysparkPythonSQLStreaming Platforms
Agency • Information Technology
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisNoSQLPysparkPythonSQL
Agency • Information Technology
Design, develop, test, and deploy high-performance Spark/Scala data processing applications and ETL pipelines using the Cloudera Hadoop ecosystem. Optimize Spark and platform performance, ensure data integrity and security, collaborate with data scientists and analysts, troubleshoot issues, and implement version control and CI/CD for Spark applications.
Top Skills:
SparkCdhCloudera HadoopFlumeGitGitlabHbaseHdfsHiveImpalaJenkinsKafkaNifiNoSQLOoziePostgresScalaSpark SqlSQLSqoop
Agency • Information Technology
Design and develop Big Data applications using Java, Spark, and MapReduce on Cloudera/Hadoop ecosystems. Work with Hive, Impala, YARN, Kafka, and large datasets; write tests (JUnit), perform data analysis, and script in Unix/Python. Lead and manage global technology teams and follow industry best practices.
Top Skills:
SparkClouderaHadoopHiveImpalaJavaJunitKafkaMapreducePythonUnix ShellYarn
Reposted 7 Days AgoSaved
Agency • Information Technology
Design and implement data federation and lakehouse architectures, build scalable ETL/ELT pipelines with Python and Spark, optimize performance across federated queries, manage Delta/Iceberg/Hudi tables, and enforce governance, security, and access controls for analytics and AI teams.
Top Skills:
AdlsApache IcebergSparkAws EmrAzure DatabricksData FederationData LakehouseDbtDelta LakeDremioGCPGlueHudiKubernetesPulumiPysparkPythonS3SQLStarburstTerraformTrino (Presto)
Agency • Information Technology
Design and implement big data solutions (Spark, Hive, Java, CDP). Analyze and consolidate disparate data sources, produce functional specifications, review data models, gather stakeholder requirements, validate implementations, support production deployments, investigate data quality and data lineage, and collaborate with technology leads to ensure data completeness and accuracy.
Top Skills:
CdpData LineageData TracingDatabasesExcelHiveJavaPowerPointSparkSQLVisioWord
Agency • Information Technology
Design, build, test, and maintain high-performance Python data applications and backend services. Lead technical design, review code, mentor junior engineers, integrate with databases and cloud services, optimize performance, and ensure security, scalability, and reliability.
Top Skills:
DjangoFastapiFlaskGitLinux/UnixMicroservicesMongoDBMySQLNoSQLPostgresPythonRestful ApisSQL
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in Cleveland, OH Hiring Data Engineers
See AllPopular Cleveland, OH Job Searches
Tech Jobs & Startup Jobs in Cleveland
Remote Jobs in Cleveland
Content Jobs in Cleveland
Customer Success Jobs in Cleveland
IT Jobs in Cleveland
Cyber Security Jobs in Cleveland
Tech Support Jobs in Cleveland
Data & Analytics Jobs in Cleveland
Analysis Reporting Jobs in Cleveland
Analytics Jobs in Cleveland
Business Intelligence Jobs in Cleveland
Data Engineer Jobs in Cleveland
Data Science Jobs in Cleveland
Machine Learning Jobs in Cleveland
Data Management Jobs in Cleveland
UX Designer Jobs in Cleveland
Software Engineer Jobs in Cleveland
Android Developer Jobs in Cleveland
C# Jobs in Cleveland
C++ Jobs in Cleveland
DevOps Jobs in Cleveland
Front End Developer Jobs in Cleveland
Golang Jobs in Cleveland
Hardware Engineer Jobs in Cleveland
iOS Developer Jobs in Cleveland
Java Developer Jobs in Cleveland
Javascript Jobs in Cleveland
Linux Jobs in Cleveland
Engineering Manager Jobs in Cleveland
.NET Developer Jobs in Cleveland
PHP Developer Jobs in Cleveland
Python Jobs in Cleveland
QA Jobs in Cleveland
Ruby Jobs in Cleveland
Salesforce Developer Jobs in Cleveland
Scala Jobs in Cleveland
Finance Jobs in Cleveland
HR Jobs in Cleveland
Internships in Cleveland
Legal Jobs in Cleveland
Marketing Jobs in Cleveland
Operations Jobs in Cleveland
Office Manager Jobs in Cleveland
Operations Manager Jobs in Cleveland
Product Manager Jobs in Cleveland
Project Manager Jobs in Cleveland
Sales Jobs in Cleveland
Account Executive (AE) Jobs in Cleveland
Account Manager (AM) Jobs in Cleveland
Sales Leadership Jobs in Cleveland
Sales Development Representative Jobs in Cleveland
Sales Engineer Jobs in Cleveland
Sales Operations Jobs in Cleveland
All Filters
Total selected ()
No Results
No Results





















