Maximum of 25 job preferences reached.
Top Data Engineer Jobs in Houston, TX
Information Technology • Database • Consulting
Lead Data Engineer builds and maintains scalable data pipelines and lakehouse/warehouse platforms (Databricks) to support analytics, BI, reporting, and AI. Responsibilities include data modeling, ETL/ELT, medallion architecture, orchestration and monitoring (Airflow/Cron), performance tuning, CDC/incremental loads, data quality, and leading/mentoring a team while collaborating with stakeholders and maintaining documentation.
Top Skills:
Apache AirflowAWSAzureCronData WarehousingDatabricksEltETLGCPHadoopHbaseHiveLakehouseMedallion ArchitecturePigPower BIPysparkPythonSparkSQLTableau
Healthtech
Design and build cloud-based data applications and pipelines to analyze clinical data, enrich and provision datasets, and support clinical/operational processes. Mentor junior engineers, drive R&D for repeatable templates, collaborate with Product/Platform/Architecture, and prioritize source control, documentation, and simple solutions using modern data, ETL, and LLM tooling.
Top Skills:
AdfAirflowBigQueryDbtEhrEmbeddingsEmrFastapiFhirFivetranFlaskFlinkGlueHadoopHl7InformaticaKafkaLangchainLinuxLlamaindexLlmNifiNoSQLOlapPythonRagRedshiftSnowflakeSparkSQLSynapseVector Dbs
Cloud • Information Technology • Other • Productivity • Software
Architect, build, and maintain Snowflake-based ELT pipelines and production data models. Implement observability, data quality, and governance; optimize cost and freshness. Partner with Analytics, Product, and Engineering, use dbt/CI-CD and AI-assisted tooling, and participate in reviews and architecture planning to support analytics, BI, and ML.
Top Skills:
Apache AirflowClaude CodeCursorDbtFivetranGitGithub ActionsGithub CopilotLlmsPrefectPythonSnowflakeSQL
Agency • Information Technology
Design and develop Big Data applications using Java, Spark, and MapReduce on Cloudera/Hadoop ecosystems. Work with Hive, Impala, YARN, Kafka, and large datasets; write tests (JUnit), perform data analysis, and script in Unix/Python. Lead and manage global technology teams and follow industry best practices.
Top Skills:
SparkClouderaHadoopHiveImpalaJavaJunitKafkaMapreducePythonUnix ShellYarn
Reposted 23 Hours AgoSaved
Agency • Information Technology
Design and implement data federation and lakehouse architectures, build scalable ETL/ELT pipelines with Python and Spark, optimize performance across federated queries, manage Delta/Iceberg/Hudi tables, and enforce governance, security, and access controls for analytics and AI teams.
Top Skills:
AdlsApache IcebergSparkAws EmrAzure DatabricksData FederationData LakehouseDbtDelta LakeDremioGCPGlueHudiKubernetesPulumiPysparkPythonS3SQLStarburstTerraformTrino (Presto)
Agency • Information Technology
Design and implement big data solutions (Spark, Hive, Java, CDP). Analyze and consolidate disparate data sources, produce functional specifications, review data models, gather stakeholder requirements, validate implementations, support production deployments, investigate data quality and data lineage, and collaborate with technology leads to ensure data completeness and accuracy.
Top Skills:
CdpData LineageData TracingDatabasesExcelHiveJavaPowerPointSparkSQLVisioWord
Agency • Information Technology
Design, build, test, and maintain high-performance Python data applications and backend services. Lead technical design, review code, mentor junior engineers, integrate with databases and cloud services, optimize performance, and ensure security, scalability, and reliability.
Top Skills:
DjangoFastapiFlaskGitLinux/UnixMicroservicesMongoDBMySQLNoSQLPostgresPythonRestful ApisSQL
Reposted 23 Hours AgoSaved
Agency • Information Technology
Design, develop, maintain, and optimize enterprise Tableau dashboards and BI solutions. Integrate and cleanse data from Oracle/Sybase/Big Data, tune SQL, automate Tableau extracts, manage Tableau Server on Linux, support testing, UAT, deployments, and troubleshoot data integrity and performance issues.
Top Skills:
Big DataLinuxOracleSQLSybaseTableau ApisTableau DesktopTableau PrepTableau Server
Agency • Information Technology
Design, build, and maintain scalable ETL/ELT pipelines into Snowflake, enforce data quality, optimize data models and query performance for Power BI reporting, produce documentation, and troubleshoot data issues while collaborating with stakeholders.
Top Skills:
AWSAzureEltETLGCPOraclePower BISnowflakeSQL
Agency • Information Technology
Design, implement, and maintain data quality processes using cloud platforms (Azure/AWS/GCP). Collaborate with stakeholders to define standards, perform data profiling and cleansing, document issues and remediation, generate quality metrics, and support cross-functional teams to improve data architectures and integration.
Top Skills:
AWSAzureGCPJavaPythonSQL
Agency • Information Technology
Design, implement, and maintain PySpark-based data reconciliation solutions for financial systems. Build matching algorithms, integrate with rules engines, process large distributed datasets, identify and resolve discrepancies, and collaborate with analysts and architects to improve data quality and governance.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisPyspark
AdTech • Agency
Design, build, and operate large-scale data pipelines and ETL processes to normalize varied partner datasets. Lead technical decisions, mentor engineers, implement testing and monitoring, manage data warehouses/lakes, and collaborate cross-functionally to deliver robust data infrastructure that supports ML and product teams.
Top Skills:
SparkAWSData LakeData WarehouseDockerGithub ActionsTerraform
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Agency • Information Technology
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisNoSQLPysparkPythonSQL
Agency • Information Technology
Design, develop, test, and deploy high-performance Spark/Scala data processing applications and ETL pipelines using the Cloudera Hadoop ecosystem. Optimize Spark and platform performance, ensure data integrity and security, collaborate with data scientists and analysts, troubleshoot issues, and implement version control and CI/CD for Spark applications.
Top Skills:
SparkCdhCloudera HadoopFlumeGitGitlabHbaseHdfsHiveImpalaJenkinsKafkaNifiNoSQLOoziePostgresScalaSpark SqlSQLSqoop
Agency • Information Technology
Design, build, and optimize PySpark applications and ETL pipelines to process large-scale datasets from SQL/NoSQL sources, data lakes, and streaming platforms. Ensure data quality, error handling, performance tuning, and collaborate with analysts, scientists, and architects to deliver scalable data solutions.
Top Skills:
Apache AirflowSparkData LakeETLLuigiNoSQLPysparkPythonSQLStreaming Platforms
Agency • Information Technology
Design, develop, and maintain PySpark applications and ETL pipelines to process, transform, and integrate large-scale datasets from SQL, NoSQL, data lakes, and streaming sources. Optimize Spark job performance, implement robust error handling, and collaborate with data analysts, scientists, and architects using orchestration tools like Airflow or Luigi.
Top Skills:
Apache AirflowSparkData LakeLuigiNoSQLPysparkPythonSQL
Industrial • Automation
Lead design and implementation of scalable cloud data lakes, warehouses, and marts. Build ELT/ETL pipelines, apply Data Vault 2.0 and dimensional modeling, implement data quality, lineage, observability, and governance. Optimize Snowflake/Databricks/Synapse platforms, champion CI/CD and DataOps, mentor engineers, and enable AI/ML-ready data products through collaboration with architecture and product teams.
Top Skills:
Azure SynapseBitbucketCi/CdData LakehouseData Vault 2.0DatabricksDbtDbt CloudDbt CoreEltETLGitInfrastructure As CodeJIRANoSQLPythonSnowflakeSQL
Consulting
Build and maintain data pipelines, applications, workflows, and governance within Palantir Foundry. Apply industrial engineering and operations research methods to analyze depot, MRO, and supply chain processes, integrate ERP/MRO data into models, run capacity/throughput analyses, deliver decision‑support briefings, and support adoption of advanced analytics and digital engineering. Coordinate with stakeholders and support business development activities.
Top Skills:
AipErpFoundry Code RepositoriesMachineryMroPalantir FoundryPipeline BuilderPysparkSlateWorkshop
Biotech
Design, develop, and maintain Customer Data Platform solutions and data pipelines to create unified customer profiles, manage identity resolution, enable audience segmentation and activation, ensure data quality and governance, and support integrations with marketing, CRM, and cloud data systems.
Top Skills:
Adobe Experience PlatformAPIsAWSAzureEtl/EltGCPIdentity ResolutionMaster Data ManagementMparticleSalesforce Data CloudSegmentSQLTealiumTreasure Data
Automotive
Own the 2-3 year architecture and reliability of high-volume catalog, pricing, and inventory ingestion. Lead migrations from legacy batch to modern systems, define data SLAs, build observability and orchestration, mentor senior engineers, and resolve cross-team, high-severity data problems.
Top Skills:
Ai Coding ToolsAws EksBigQueryDatabricksEc2FlinkKafkaKinesisKubernetesPysparkPythonRdsRedpandaSnowflakeSparkSqs
Mobile • Other • Software • Analytics
Design, build, and maintain scalable data systems and ETLs on GCP; connect production data to business systems; generate analytics and ML-driven insights for GTM teams; work cross-functionally and introduce modern tools (Airflow, DBT, Presto, Hightouch).
Top Skills:
AirflowAirflowData LakeData WarehouseDbtETLGCPGoogle AnalyticsHightouchIntercomMarketoMetabasePrestoPythonSalesforce
Marketing Tech • Consulting
Design and implement BI solutions: build data models, automations, ETL/ELT workflows, dashboards, and ad-hoc analyses. Monitor and analyze RF/video performance using Dataminer/Qligent or similar tools. Collaborate with cross-functional teams, document requirements, and deliver actionable insights to optimize hybrid terrestrial and satellite video delivery.
Top Skills:
AzureDataminerEltETLPower BIQligentRf MonitoringSQLTableauVideo MonitoringXML
HR Tech • Logistics • Software
Design, build, and maintain scalable batch and streaming data pipelines, improve data modeling and governance, contribute to enterprise data architecture, and collaborate with cross-functional teams to ensure reliable, high-quality data for analytics and operational use.
Top Skills:
Ai-Enabled ToolsAws EventbridgeAws GlueAws LambdaAws RedshiftAws S3CdcData WarehouseEtl/EltEvent-DrivenGraph Data ModelsLlm/Ml PipelinesMedallion ArchitecturePythonRole-Based Access Control (Rbac)SQLStreaming
Industrial • Automation
Lead design and delivery of scalable, secure cloud data pipelines and products. Translate requirements into architecture, apply advanced data modeling (Data Vault 2.0/dimensional), embed data quality, observability, security, and DataOps/CI-CD practices, optimize performance and cost, mentor engineers, and enable AI/ML-ready datasets within an agile DevSecOps model.
Top Skills:
AdlsAirflowAzureAzure Data FactoryCi/CdData Vault 2.0DatabricksDataopsDevsecopsEpicInfrastructure As CodeSnowflakeSQLSynapseTidal
Industrial • Automation
Design, build, and operate scalable cloud data pipelines and platforms (Azure preferred). Implement DevSecOps, CI/CD, automated testing, monitoring, performance tuning, and production readiness. Mentor engineers, drive operational excellence, and translate requirements into robust technical designs. Apply healthcare domain knowledge and collaborate in agile teams to deliver secure, reliable data products.
Top Skills:
AdlsAirflowAzureBitbucketData FactoryData Vault 2.0DatabricksDbtEpicGitSnowflakeSQLSynapseTidal
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in Houston, TX Hiring Data Engineers
See AllPopular Houston, TX Job Searches
Tech Jobs & Startup Jobs in Houston
Remote Jobs in Houston
Content Jobs in Houston
Customer Success Jobs in Houston
IT Jobs in Houston
Cyber Security Jobs in Houston
Tech Support Jobs in Houston
Data & Analytics Jobs in Houston
Analysis Reporting Jobs in Houston
Analytics Jobs in Houston
Business Intelligence Jobs in Houston
Data Engineer Jobs in Houston
Data Science Jobs in Houston
Machine Learning Jobs in Houston
Data Management Jobs in Houston
UX Designer Jobs in Houston
Software Engineer Jobs in Houston
Android Developer Jobs in Houston
C# Jobs in Houston
C++ Jobs in Houston
DevOps Jobs in Houston
Front End Developer Jobs in Houston
Golang Jobs in Houston
Hardware Engineer Jobs in Houston
iOS Developer Jobs in Houston
Java Developer Jobs in Houston
Javascript Jobs in Houston
Linux Jobs in Houston
Engineering Manager Jobs in Houston
.NET Developer Jobs in Houston
PHP Developer Jobs in Houston
Python Jobs in Houston
QA Jobs in Houston
Ruby Jobs in Houston
Salesforce Developer Jobs in Houston
Scala Jobs in Houston
Finance Jobs in Houston
HR Jobs in Houston
Internships in Houston
Legal Jobs in Houston
Marketing Jobs in Houston
Operations Jobs in Houston
Office Manager Jobs in Houston
Operations Manager Jobs in Houston
Product Manager Jobs in Houston
Project Manager Jobs in Houston
Sales Jobs in Houston
Account Executive (AE) Jobs in Houston
Account Manager (AM) Jobs in Houston
Sales Leadership Jobs in Houston
Sales Development Representative Jobs in Houston
Sales Engineer Jobs in Houston
Sales Operations Jobs in Houston
All Filters
Total selected ()
No Results
No Results











.png)











