Maximum of 25 job preferences reached.
Top Data Engineer Jobs in San Diego, CA
Healthtech
Design and build cloud-based data applications and pipelines to analyze clinical data, enrich and provision datasets, and support clinical/operational processes. Mentor junior engineers, drive R&D for repeatable templates, collaborate with Product/Platform/Architecture, and prioritize source control, documentation, and simple solutions using modern data, ETL, and LLM tooling.
Top Skills:
AdfAirflowBigQueryDbtEhrEmbeddingsEmrFastapiFhirFivetranFlaskFlinkGlueHadoopHl7InformaticaKafkaLangchainLinuxLlamaindexLlmNifiNoSQLOlapPythonRagRedshiftSnowflakeSparkSQLSynapseVector Dbs
Cloud • Information Technology • Other • Productivity • Software
Architect, build, and maintain Snowflake-based ELT pipelines and production data models. Implement observability, data quality, and governance; optimize cost and freshness. Partner with Analytics, Product, and Engineering, use dbt/CI-CD and AI-assisted tooling, and participate in reviews and architecture planning to support analytics, BI, and ML.
Top Skills:
Apache AirflowClaude CodeCursorDbtFivetranGitGithub ActionsGithub CopilotLlmsPrefectPythonSnowflakeSQL
Reposted 10 Days AgoSaved
Agency • Information Technology
Design and implement data federation and lakehouse architectures, build scalable ETL/ELT pipelines with Python and Spark, optimize performance across federated queries, manage Delta/Iceberg/Hudi tables, and enforce governance, security, and access controls for analytics and AI teams.
Top Skills:
AdlsApache IcebergSparkAws EmrAzure DatabricksData FederationData LakehouseDbtDelta LakeDremioGCPGlueHudiKubernetesPulumiPysparkPythonS3SQLStarburstTerraformTrino (Presto)
Agency • Information Technology
Design and develop Big Data applications using Java, Spark, and MapReduce on Cloudera/Hadoop ecosystems. Work with Hive, Impala, YARN, Kafka, and large datasets; write tests (JUnit), perform data analysis, and script in Unix/Python. Lead and manage global technology teams and follow industry best practices.
Top Skills:
SparkClouderaHadoopHiveImpalaJavaJunitKafkaMapreducePythonUnix ShellYarn
Agency • Information Technology
Design and implement big data solutions (Spark, Hive, Java, CDP). Analyze and consolidate disparate data sources, produce functional specifications, review data models, gather stakeholder requirements, validate implementations, support production deployments, investigate data quality and data lineage, and collaborate with technology leads to ensure data completeness and accuracy.
Top Skills:
CdpData LineageData TracingDatabasesExcelHiveJavaPowerPointSparkSQLVisioWord
Agency • Information Technology
Design, build, test, and maintain high-performance Python data applications and backend services. Lead technical design, review code, mentor junior engineers, integrate with databases and cloud services, optimize performance, and ensure security, scalability, and reliability.
Top Skills:
DjangoFastapiFlaskGitLinux/UnixMicroservicesMongoDBMySQLNoSQLPostgresPythonRestful ApisSQL
Reposted 10 Days AgoSaved
Agency • Information Technology
Design, develop, maintain, and optimize enterprise Tableau dashboards and BI solutions. Integrate and cleanse data from Oracle/Sybase/Big Data, tune SQL, automate Tableau extracts, manage Tableau Server on Linux, support testing, UAT, deployments, and troubleshoot data integrity and performance issues.
Top Skills:
Big DataLinuxOracleSQLSybaseTableau ApisTableau DesktopTableau PrepTableau Server
Agency • Information Technology
Design, build, and maintain scalable ETL/ELT pipelines into Snowflake, enforce data quality, optimize data models and query performance for Power BI reporting, produce documentation, and troubleshoot data issues while collaborating with stakeholders.
Top Skills:
AWSAzureEltETLGCPOraclePower BISnowflakeSQL
Agency • Information Technology
Design, implement, and maintain PySpark-based data reconciliation solutions for financial systems. Build matching algorithms, integrate with rules engines, process large distributed datasets, identify and resolve discrepancies, and collaborate with analysts and architects to improve data quality and governance.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisPyspark
Agency • Information Technology
Design, implement, and maintain data quality processes using cloud platforms (Azure/AWS/GCP). Collaborate with stakeholders to define standards, perform data profiling and cleansing, document issues and remediation, generate quality metrics, and support cross-functional teams to improve data architectures and integration.
Top Skills:
AWSAzureGCPJavaPythonSQL
Agency • Information Technology
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisNoSQLPysparkPythonSQL
Agency • Information Technology
Design, build, and optimize PySpark applications and ETL pipelines to process large-scale datasets from SQL/NoSQL sources, data lakes, and streaming platforms. Ensure data quality, error handling, performance tuning, and collaborate with analysts, scientists, and architects to deliver scalable data solutions.
Top Skills:
Apache AirflowSparkData LakeETLLuigiNoSQLPysparkPythonSQLStreaming Platforms
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Agency • Information Technology
Design, develop, test, and deploy high-performance Spark/Scala data processing applications and ETL pipelines using the Cloudera Hadoop ecosystem. Optimize Spark and platform performance, ensure data integrity and security, collaborate with data scientists and analysts, troubleshoot issues, and implement version control and CI/CD for Spark applications.
Top Skills:
SparkCdhCloudera HadoopFlumeGitGitlabHbaseHdfsHiveImpalaJenkinsKafkaNifiNoSQLOoziePostgresScalaSpark SqlSQLSqoop
Agency • Information Technology
Design, develop, and maintain PySpark applications and ETL pipelines to process, transform, and integrate large-scale datasets from SQL, NoSQL, data lakes, and streaming sources. Optimize Spark job performance, implement robust error handling, and collaborate with data analysts, scientists, and architects using orchestration tools like Airflow or Luigi.
Top Skills:
Apache AirflowSparkData LakeLuigiNoSQLPysparkPythonSQL
Financial Services
Design and build a modern, scalable data platform (warehouse, lakehouse, or hybrid) to support analytics, data science, and near-real-time processing. Lead platform architecture, CDC-based ingestion, ETL/ELT pipelines, data quality/observability, and governance. Partner cross-functionally, provide technical leadership and mentorship, and modernize legacy data processes for low-latency enterprise reporting and ML workloads.
Top Skills:
Azure Data Lake StorageCdcData WarehouseDatabricksDimensional ModelingEltETLLakehouseMicrosoft FabricMl PipelinesPythonSnowflakeSnowflake SchemaSQLStar Schema
Automotive
Own the 2-3 year architecture and reliability of high-volume catalog, pricing, and inventory ingestion. Lead migrations from legacy batch to modern systems, define data SLAs, build observability and orchestration, mentor senior engineers, and resolve cross-team, high-severity data problems.
Top Skills:
Ai Coding ToolsAws EksBigQueryDatabricksEc2FlinkKafkaKinesisKubernetesPysparkPythonRdsRedpandaSnowflakeSparkSqs
Mobile • Other • Software • Analytics
Design, build, and maintain scalable data systems and ETLs on GCP; connect production data to business systems; generate analytics and ML-driven insights for GTM teams; work cross-functionally and introduce modern tools (Airflow, DBT, Presto, Hightouch).
Top Skills:
AirflowAirflowData LakeData WarehouseDbtETLGCPGoogle AnalyticsHightouchIntercomMarketoMetabasePrestoPythonSalesforce
HR Tech • Logistics • Software
Design, build, and maintain scalable batch and streaming data pipelines, improve data modeling and governance, contribute to enterprise data architecture, and collaborate with cross-functional teams to ensure reliable, high-quality data for analytics and operational use.
Top Skills:
Ai-Enabled ToolsAws EventbridgeAws GlueAws LambdaAws RedshiftAws S3CdcData WarehouseEtl/EltEvent-DrivenGraph Data ModelsLlm/Ml PipelinesMedallion ArchitecturePythonRole-Based Access Control (Rbac)SQLStreaming
eCommerce • Hardware • Healthtech • Software
Design, build, and own end-to-end ingestion pipelines and dbt transformation layers on Databricks. Lead GCP->AWS/Databricks migration, validate parity, and decommission legacy systems. Implement CDC and batch patterns, ensure data quality and observability, document assets in Unity Catalog, and build reverse ETL integrations to operational systems. Partner with business stakeholders and write architectural decision records to define engineering patterns and acceptance criteria.
Top Skills:
AirflowAutoloaderAWSBigQueryCdcCloud FunctionsCloud RunDatabricksDbtDelta LakeDelta Live TablesGCPGreat ExpectationsLakeflowMulesoftNetSuitePysparkPythonSalesforceStripeUnity Catalog
Artificial Intelligence • Machine Learning • Software • Analytics
Work directly with customers to design, build, and maintain data integrations and pipelines, troubleshoot ingestion/transformation/delivery issues, optimize large-scale data processing, collaborate with data scientists, and guide technical onboarding and long-term success.
Top Skills:
Automated TestingAws AthenaAws BatchAws Ec2Aws LambdaAws S3Ci/CdDuckdbGitPostgresPythonServerlessSQL
Real Estate • Travel • PropTech
Build and maintain production data pipelines and foundations for People Analytics and AI-driven tools. Integrate HR systems (Workday, Greenhouse), design data models, ensure governance for sensitive employee data, support data science and reporting, and deliver dashboards and Streamlit apps. Optimize queries, enable LLM consumption with quality and latency controls, and collaborate cross-functionally to transition AI prototypes to production.
Top Skills:
AirflowAirtableAmazon S3Aws (Ssh)GitHiveLlm/Agentic Ai FrameworksPostgresPrestoPythonSftpSQLStreamlitTrinoUbuntuWeb Apis
Healthtech
Maintain and build data products, pipelines, and automation to support care management and clinical programs. Integrate multiple data sources into single sources of truth, perform root cause analysis, and serve as a subject-matter expert for partner teams to drive clinical, operational, and financial insights.
Top Skills:
AccessDatabricksExcelPowerPointPysparkPythonSASSQLWord
Reposted 12 Days AgoSaved
Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
The Senior Data Engineer will design scalable data processing frameworks using Java and Apache Spark, manage data pipelines, and mentor junior engineers.
Top Skills:
SparkAWSDockerIcebergJavaKubernetesParquetRestful ApisSpring Boot
Fintech • Real Estate • Software
As a Data Engineering Architect, lead the transformation of data engineering into a modern platform, oversee architecture and technical leadership, design scalable systems, ensure data quality and governance, and mentor teams for a data-driven culture.
Top Skills:
AWSAzureDbtFlinkGCPJavaLookerPower BIPythonScalaSparkSQLTableau
Security • Software
Design and execute SQL-based ETL processes to migrate legacy public safety data into Mark43. Collaborate with customers, contractors, and internal teams to map legacy schemas, handle edge cases, ensure data fidelity, and improve migration tooling while managing timelines and stakeholder relationships.
Top Skills:
ETLRelational DatabasesSQL
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in San Diego, CA Hiring Data Engineers
See AllPopular San Diego, CA Job Searches
Tech Jobs & Startup Jobs in San Diego, CA
Remote Jobs in San Diego, CA
Content Jobs in San Diego, CA
Customer Success Jobs in San Diego, CA
IT Jobs in San Diego, CA
Cyber Security Jobs in San Diego, CA
Tech Support Jobs in San Diego, CA
Data & Analytics Jobs in San Diego, CA
Analysis Reporting Jobs in San Diego, CA
Analytics Jobs in San Diego, CA
Business Intelligence Jobs in San Diego, CA
Data Engineer Jobs in San Diego, CA
Data Science Jobs in San Diego, CA
Machine Learning Jobs in San Diego, CA
Data Management Jobs in San Diego, CA
UX Designer Jobs in San Diego, CA
Software Engineer Jobs in San Diego, CA
Android Developer Jobs in San Diego, CA
C# Jobs in San Diego, CA
C++ Jobs in San Diego, CA
DevOps Jobs in San Diego, CA
Front End Developer Jobs in San Diego, CA
Golang Jobs in San Diego, CA
Hardware Engineer Jobs in San Diego, CA
iOS Developer Jobs in San Diego, CA
Java Developer Jobs in San Diego, CA
Javascript Jobs in San Diego, CA
Linux Jobs in San Diego, CA
Engineering Manager Jobs in San Diego, CA
.NET Developer Jobs in San Diego, CA
PHP Developer Jobs in San Diego, CA
Python Jobs in San Diego, CA
QA Jobs in San Diego, CA
Ruby Jobs in San Diego, CA
Salesforce Developer Jobs in San Diego, CA
Scala Jobs in San Diego, CA
Finance Jobs in San Diego, CA
HR Jobs in San Diego, CA
Internships in San Diego, CA
Legal Jobs in San Diego, CA
Marketing Jobs in San Diego, CA
Operations Jobs in San Diego, CA
Office Manager Jobs in San Diego, CA
Operations Manager Jobs in San Diego, CA
Product Manager Jobs in San Diego, CA
Project Manager Jobs in San Diego, CA
Sales Jobs in San Diego, CA
Account Executive (AE) Jobs in San Diego, CA
Account Manager (AM) Jobs in San Diego, CA
Sales Leadership Jobs in San Diego, CA
Sales Development Representative Jobs in San Diego, CA
Sales Engineer Jobs in San Diego, CA
Sales Operations Jobs in San Diego, CA
All Filters
Total selected ()
No Results
No Results






.png)


_1.png)















