Maximum of 25 job preferences reached.
Top Data Engineer Jobs in Phoenix, AZ
Artificial Intelligence • Enterprise Web • Software • Design • Generative AI
Designs and operates batch, streaming, and real-time data platforms and pipelines using Spark, Kafka, Iceberg, Airflow, and cloud infrastructure. Owns data lake evolution, data quality, observability, reliability, governance, event instrumentation, schema management, and privacy controls for PII, retention, deletion, and residency. Leads complex initiatives end-to-end, troubleshoots production issues, mentors engineers, and develops AI agent harnesses that enforce data engineering standards and quality gates.
Top Skills:
AirflowAmazon EmrApache IcebergChange Data Capture (Cdc)Ci/CdDruidEksInfrastructure As CodeKafkaKubernetesMwaaSparkSpark Structured StreamingSQL
Information Technology
Builds and maintains scalable data pipelines, ETL infrastructure, data warehouses, and analytics tools. The role assembles complex datasets, automates processes, optimizes data delivery architecture, and supports cross-functional data initiatives. Collaborates with executive, product, data, and design stakeholders to resolve technical issues and improve analytics systems. Requires proficiency in SAS, Python, Java, MATLAB, SQL, NoSQL, Hadoop, MapReduce, Hive, Presto, UNIX, and Linux, with JavaScript, R, and basic machine learning preferred.
Top Skills:
ETLHadoopHiveJavaJavaScriptLinuxMapreduceMatlabNoSQLPrestoPythonRSASSQLUnix
Information Technology • Legal Tech • Consulting
Lead the architecture, development, and optimization of Snowflake and enterprise data platforms. Design scalable ETL/ELT pipelines, dimensional data models, and medallion architectures while establishing engineering standards for testing, deployment, monitoring, security, and data quality. Partner with stakeholders on technical roadmaps, mentor engineers and analysts, conduct design and code reviews, modernize legacy pipelines, and deliver reliable, reusable data solutions.
Top Skills:
Azure Data FactoryDbtDimensional Data ModelingEltETLMedallion ArchitecturePythonSigmaSnowflakeSQLStar Schema
Artificial Intelligence • Software
Build and operate large-scale data pipelines, web-scraping systems, and data-collection tools for AI datasets. Maintain analytical databases, optimize queries, troubleshoot pipeline and data-quality issues, manage containerized workloads across Kubernetes and Linux infrastructure, and support CI/CD workflows. The role also involves documenting technical processes, researching improvements, and developing scalable APIs in a fully remote environment.
Top Skills:
ArgocdBigQueryCeleryClickhouseDatabendDockerGithub ActionsHelmKafkaKubernetesLinuxPythonRabbitMQ
Cloud • Information Technology
Design and modify data solutions using Databricks and Palantir Foundry; build data models and reports from ERP and legacy-system datasets; define and execute tests; and translate stakeholder reporting needs into technical specifications. The role requires SQL, Python, PySpark, database architecture, cloud architecture, software engineering practices, independent execution, documentation, and collaboration. It supports an Army client in a hybrid Arlington, Virginia environment and requires an active Secret clearance.
Top Skills:
AWSAzureData PipelinesDatabricksEtl PipelinesNon-Relational DatabasesPalantir FoundryPysparkPythonRelational DatabasesSap ErpSQLVersion ControlWorkflow Automation
Insurance • Legal Tech • Social Impact
As the first in-house Data Engineer, you will own the data infrastructure, build and maintain data pipelines, and collaborate with teams to ensure quality data access and insights.
Top Skills:
AirflowBigQueryDagsterDbtGoPythonSQL
Healthtech
Design and build PerfectServe’s unified cloud data platform, including ingestion pipelines, integrations, data models, metrics, and consumer-facing data layers. Establish architecture, engineering, quality, and testing standards; enable safe extraction from production systems; partner with data users; and mentor data engineers. The role requires substantial cloud data platform experience, deep SQL, warehouse expertise, ELT, change-data-capture, AWS, modeling, orchestration, and data quality skills.
Top Skills:
AirflowAmazon S3AWSBigQueryBusiness IntelligenceChange Data CaptureDagsterData Quality TestingDatabricksDbtDimensional ModelingEltMedallion ArchitecturePrefectSnowflakeSQL
Information Technology
Leads the design, development, and maintenance of scalable AWS data platforms, pipelines, ETL/ELT workflows, data models, and orchestration systems. Uses Python, PySpark, SQL, dbt, AWS services, modern lakehouse technologies, and CI/CD practices to deliver secure federal data solutions. Responsibilities include platform modernization, performance optimization, AI-enabled data integration, compliance support, technical mentoring, architecture discussions, and cross-functional Agile collaboration.
Top Skills:
Amazon AthenaAmazon BedrockAmazon CloudwatchAmazon EmrAmazon EventbridgeAmazon MwaaAmazon RdsAmazon RedshiftAmazon S3Amazon S3 VectorsAmazon SnsAmazon SqsApache AirflowApache HiveApache IcebergSparkAvroAws CloudformationAws DmsAws GlueAws LambdaAws Step FunctionsCi/CdDbtFedrampGitGraphdbHadoopHarnessIbm DatastageJavaNist Sp 800-53Nosql DatabasesOpensearchOpensearch Vector IndexesOrcParquetPostgresPysparkPythonRagRundeckShell ScriptingSQLTrinoVector Search
Renewable Energy
Staff Data Engineer responsible for defining data architecture strategy, building fault-tolerant batch and real-time streaming infrastructure, and managing scalable lakehouse and database systems. The role processes telemetry from millions of connected devices, drives performance tuning and data quality, supports on-call operations, evaluates modern technologies, and mentors engineers. It requires strong expertise in Python, SQL, cloud data platforms, streaming systems, infrastructure as code, and enterprise-scale architecture.
Top Skills:
Apache FlinkApache IcebergApache KafkaSparkAWSAws CdkAws GlueAws KinesisAws LambdaAws RedshiftAws S3CdktfCi/CdDelta LakeGCPGcp Pub/SubPostgresql AuroraPrefectPythonRaySQLTerraform
Artificial Intelligence • Software
Build and manage scalable ETL/ELT pipelines using Microsoft Fabric and Azure Databricks, implement medallion architecture, support data migrations and legacy SSRS decommissioning, develop Power BI data models, and maintain governance, security, and CI/CD practices. The role also supports migration events and modernizes data infrastructure for advanced analytics and reporting.
Top Skills:
SparkAzureAzure DatabricksAzure DevopsMicrosoft FabricMicrosoft PurviewOnelakePower BIPythonSQLSQL ServerSsrsUnity Catalog
Fintech • Financial Services
Designs and maintains scalable cloud data architectures, pipelines, ETL processes, data warehouses, databases, and data models. Ensures data quality, accuracy, security, availability, and consistency across sources. Develops SQL queries for analytics and reporting, and documents pipelines, schemas, architectures, and models. The role requires expertise with Python, Java, Scala, R, SQL, relational and NoSQL databases, cloud platforms, and Snowflake.
Top Skills:
AWSETLGCPJavaAzureNosql DatabasesPythonRRdbmsScalaSnowflakeSQL
Healthtech • Software
Designs and delivers large-scale, governed data pipelines and infrastructure across the full data lifecycle. Responsibilities include coding in Python and SQL, optimizing Airflow, dbt, AWS, Kafka, and Iceberg/Parquet workflows, implementing data quality and observability, contributing to architecture and tool evaluations, mentoring engineers, collaborating with stakeholders, and participating in on-call support.
Top Skills:
AirflowApache IcebergAPIsAthenaAWSCi/CdDatabricksDbtEmrHipaaKafkaParquetPythonS3SnowflakeSQL
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Security • Cybersecurity
Designs and manages scalable data pipelines, contracts, transformation layers, and analytics models. The role focuses on data governance, observability, lineage, configuration validation, and collaboration with product engineering teams. It requires processing large datasets, integrating cloud and on-premises industrial technologies, and implementing streaming and message-queuing solutions, ideally in cybersecurity or industrial environments.
Top Skills:
DockerGoJvmKubernetesNode.jsPython
Healthtech • Social Impact • Software
Designs and maintains scalable data architecture, warehouses, ETL/ELT pipelines, and data platforms. Responsibilities include ensuring data quality, integrity, governance, accessibility, reliability, and performance; supporting customer onboarding and analytics tools; contributing to architecture and engineering best practices; troubleshooting complex data issues; and participating in on-call rotations. The role collaborates with engineering, product, solutions, business intelligence, and other cross-functional teams in healthcare and social care data environments.
Top Skills:
Amazon EcsAmazon RdsAmazon RedshiftAmazon S3Apache AirflowAWSCi/CdDbtDockerEltETLGithub ActionsJenkinsKafkaKubernetesLookerPostgresPythonSigmaSnowflakeSQLTableauTerraformThoughtspot
Machine Learning • Software
Own the analytics data layer by translating stakeholder needs into certified dimensional models, governed metrics, semantic layers, dashboards, and reports. Build with dbt, SQL, Databricks, Power BI, and Airflow while ensuring data quality, observability, reconciliation, documentation, and dependable orchestration. Partner closely with finance, operations, and client-facing teams, automate recurring workflows, and use AI-assisted engineering practices in a fully remote environment.
Top Skills:
AirflowBigQueryCubeDatabricksDatabricks Unity CatalogDaxDbtExcelGitLlmsLookerLookmlMetricflowPower BIPythonRedshiftSharepointSnowflakeSpark SqlSQLTableau
Cloud • Information Technology
Build and operate cloud data capabilities, including ingestion pipelines, transformations, data models, curated data products, and data-quality controls. Use Snowflake, dbt, SQL, and Python to deliver governed data for analytics, applications, automation, and AI workflows. Implement testing, CI/CD, documentation, lineage, access controls, monitoring, incident resolution, and performance optimization. Collaborate with engineering, analytics, governance, security, and business teams in an Agile, product-oriented environment with active AI-assisted development.
Top Skills:
APIsAzureCi/CdDbtGraphQLNetSuitePythonRestSalesforceSnowflakeSQL
Consulting
Designs and operates an AWS-native lakehouse supporting analytics, reporting, visualization, and AI/ML. Builds batch and streaming ingestion pipelines using Python and PySpark, Apache Iceberg data structures, governance and metadata services, quality controls, security, observability, and CI/CD automation. Optimizes performance, scalability, reliability, and cost while collaborating with cross-functional technical teams and documenting architecture, operations, and secure configurations.
Top Skills:
Amazon AthenaAmazon EmrAmazon KinesisAmazon MskAmazon QuicksightAmazon RedshiftAmazon S3Apache IcebergApache ParquetAws CdkAws CloudformationAws CodepipelineAws DmsAws GlueAws KmsAws Lake FormationAws LambdaAws Step FunctionsChatgptCursorDatabricksDelta LakeDockerGitGithub ActionsGithub CopilotIamJenkinsKiroMwaaPower BIPysparkPythonSQLTableauTerraform
Edtech
Lead design, build, and operate scalable Databricks/Delta Lake data pipelines and Kimball dimensional models using dbt. Manage governance with Unity Catalog, optimize performance and cost, operationalize ML with MLflow, and coordinate/quality-check offshore vendor engineering. Mentor engineers, uphold standards, and deliver trusted data products that power analytics, reporting, and MLOps workflows.
Top Skills:
Azure Event HubsDatabricksDatabricks Dashboards (Genie)DbtDelta LakeDelta LakehouseDockerKafkaLakeflow ConnectMlflowMlopsPower BIPysparkPythonSalesforceSparkSQLStructured StreamingTableauUnity Catalog
Real Estate • Travel • PropTech
Provides hands-on technical leadership for foundational data engineering and analytics engineering. Owns multi-year data architecture, production pipelines, dimensional models, and trusted metrics spanning cloud costs, traffic, infrastructure, and AI usage. Designs new datasets integrating cost, utilization, performance, and reliability signals; establishes standards, coaches engineers, aligns stakeholders, and drives adoption across Infrastructure, Finance, Data Science, and Product Engineering.
Top Skills:
AirflowMinervaPythonScalaSQL
Biotech • Agriculture
Designs, develops, and maintains Databricks-based ETL pipelines, data integration solutions, data warehouses, and reporting datasets. Builds batch and streaming architectures using Python, Spark, SQL, and Delta Lake; optimizes performance and costs; migrates legacy SSIS and SQL Server processes; maintains SSRS reports; implements data quality controls; and collaborates with business and technical teams supporting analytics and AI/ML workloads.
Top Skills:
Ai-Assisted Development ToolsSparkAzure Data FactoryData WarehousingDatabricksDelta LakeETLAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
Agriculture
Designs, develops, and maintains Databricks-based ETL pipelines, data lake architectures, datasets, and data warehouses supporting reporting, analytics, and AI/ML workloads. Troubleshoots and optimizes Databricks and SSIS processes, maintains SSRS reports, implements data quality controls, migrates legacy systems, and collaborates with business and technical teams. The role also promotes AI-assisted development practices and documents technical procedures.
Top Skills:
SparkAzure Data FactoryDatabricksDelta LakeAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
Reposted One Month AgoSaved
Easy Apply
Easy Apply
AdTech • Artificial Intelligence • Marketing Tech • Software • Analytics
Build, deploy, and operate production-grade data pipelines and data products for healthcare audiences. Design transformations, data models, and governed views using Python, SQL, Airflow, S3, Snowflake, and EMR. Implement data-quality, monitoring, and privacy-by-design controls for PHI/PII. Partner with product, analytics, and platform teams to onboard sources, support audience discovery, segmentation, activation, measurement, and troubleshoot production issues.
Top Skills:
Amazon EmrAmazon S3Apache AirflowAthenaHivePythonSnowflakeSQL
Cloud • Information Technology • Cybersecurity
Designs, analyzes, documents, integrates, and migrates complex environmental datasets for AirHub business lines. Responsibilities include analyzing legacy databases, creating logical and physical data models, documenting mappings and business rules, developing migration specifications and API integrations, supporting data governance and testing, and producing technical documentation. The role also requires close collaboration with business stakeholders in an Agile delivery environment.
Top Skills:
Api IntegrationsEtl/EltPostgresSQL
Automotive • Insurance • Machine Learning • Mobile • Software
Lead financial data engineering initiatives involving commissions, incentive programs, and backend financial systems. Build scalable, accurate, and reliable data pipelines, platform tooling, partner integrations, analytics infrastructure, and data science infrastructure. Shape technical strategy and roadmaps, collaborate with Finance, Product, Data Science, and Engineering, mentor engineers, contribute code, lead incident response, and establish engineering standards.
Top Skills:
Amazon RedshiftApache AirflowApache IcebergSparkAWSDbtPythonSQLTerraform
Healthtech
Develops scalable data models, Databricks pipelines, backend applications, ETL processes, and data ingestion systems. Partners with engineering, product, and business stakeholders on technical roadmaps; monitors and optimizes data systems for performance, reliability, and scalability. Provides technical leadership, mentors junior engineers, conducts code reviews, and builds secure solutions for sensitive healthcare data.
Top Skills:
BigQueryCi/CdDatabricksDockerJavaKafkaKubernetesNode.jsPythonRedshiftSnowflakeSQL
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in Phoenix, AZ Hiring Data Engineers
See AllPopular Phoenix, AZ Job Searches
Tech Jobs & Startup Jobs in Phoenix
Remote Jobs in Phoenix
Content Jobs in Phoenix
Customer Success Jobs in Phoenix
IT Jobs in Phoenix
Cyber Security Jobs in Phoenix
Tech Support Jobs in Phoenix
Data & Analytics Jobs in Phoenix
Analysis Reporting Jobs in Phoenix
Analytics Jobs in Phoenix
Business Intelligence Jobs in Phoenix
Data Engineer Jobs in Phoenix
Data Science Jobs in Phoenix
Machine Learning Jobs in Phoenix
Data Management Jobs in Phoenix
UX Designer Jobs in Phoenix
Software Engineer Jobs in Phoenix
Android Developer Jobs in Phoenix
C# Jobs in Phoenix
C++ Jobs in Phoenix
DevOps Jobs in Phoenix
Front End Developer Jobs in Phoenix
Golang Jobs in Phoenix
Hardware Engineer Jobs in Phoenix
iOS Developer Jobs in Phoenix
Java Developer Jobs in Phoenix
Javascript Jobs in Phoenix
Linux Jobs in Phoenix
Engineering Manager Jobs in Phoenix
.NET Developer Jobs in Phoenix
PHP Developer Jobs in Phoenix
Python Jobs in Phoenix
QA Jobs in Phoenix
Ruby Jobs in Phoenix
Salesforce Developer Jobs in Phoenix
Scala Jobs in Phoenix
Finance Jobs in Phoenix
HR Jobs in Phoenix
Internships in Phoenix
Legal Jobs in Phoenix
Marketing Jobs in Phoenix
Operations Jobs in Phoenix
Office Manager Jobs in Phoenix
Operations Manager Jobs in Phoenix
Product Manager Jobs in Phoenix
Project Manager Jobs in Phoenix
Sales Jobs in Phoenix
Account Executive (AE) Jobs in Phoenix
Account Manager (AM) Jobs in Phoenix
Sales Leadership Jobs in Phoenix
Sales Development Representative Jobs in Phoenix
Sales Engineer Jobs in Phoenix
Sales Operations Jobs in Phoenix
All Filters
Total selected ()
No Results
No Results









%20(1).png)




.png)

_1.png)














