Top Data Engineer Jobs in Pittsburgh, PA

Reposted 9 Days AgoSaved
Hybrid
Pittsburgh, PA
124K-280K Annually
Senior level
124K-280K Annually
Senior level
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Lead design and implementation of enterprise data architecture and cloud-based data solutions. Manage large data engineering projects, develop data models and pipelines, ensure governance and security compliance, advise clients strategically, and build high-performing inclusive teams.
Top Skills: AWSAzureDatabricksGCPSnowflake
YesterdaySaved
Remote
Pittsburgh, PA
121K-164K Annually
Junior
121K-164K Annually
Junior
Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Build and operate production data pipelines and dimensional models using Spark, SparkSQL, and cloud lakehouse technologies. Own pipelines from requirements through deployment, monitoring, and iteration; improve data quality, lineage, reliability, and cost efficiency. Partner with data scientists, analysts, product managers, and engineers to support datamarts, KPIs, reporting, and analysis. Participate in business-hours on-call rotations and improve runbooks and alerting.
Top Skills: AirflowC++DatabricksJavaKafkaKinesisMonte CarloPythonScalaSparkSparksqlSQLStructured Streaming
14 Days AgoSaved
Remote or Hybrid
Pittsburgh, PA
125K-159K Annually
Mid level
125K-159K Annually
Mid level
AdTech • Automotive • Big Data • Consumer Web
Administer and enhance Edmunds’ Databricks data platform and AWS infrastructure. Build and maintain ETL pipelines, infrastructure-as-code tooling using Terraform or CDK, and operational dashboards, alerts, and reports. Collaborate with business, engineering, analytics, security, and infrastructure teams to support data platform users and AI solutions. Evaluate new technologies, troubleshoot platform issues, and improve operational, cost, and security visibility.
Top Skills: SparkAWSAws CdkDatabricksInfrastructure As Code (Iac)PythonScalaSQLTerraform
19 Days AgoSaved
Remote or Hybrid
Pittsburgh, PA
160K-260K Annually
Expert/Leader
160K-260K Annually
Expert/Leader
Artificial Intelligence • Cloud • Payments • Software • Business Intelligence • Generative AI • Automation
Define and govern enterprise-scale data architecture across batch, streaming, warehouse, lakehouse, transactional, and AI use cases. Establish standards for data quality, lineage, access, cataloging, governance, observability, and SLAs. Architect AI-enabled workflows, resolve complex architecture issues, influence roadmaps, and mentor engineers through hands-on technical leadership. The role requires 15+ years of software, data engineering, or architecture experience and expertise in large-scale data platforms and modeling.
Top Skills: AIBatch ProcessingBigQueryData CatalogsData WarehousesDbtFeature StoresGCPLakehousesOlapOltpStreaming ArchitecturesVector Stores
Reposted 7 Days AgoSaved
In-Office
Pittsburgh, PA
103K-125K Annually
Senior level
103K-125K Annually
Senior level
Healthtech • Software
Designs and maintains operational data infrastructure for smart manufacturing, including pipelines, data warehouses, ETL/ELT processes, time-series data systems, integrations, governance, and trusted datasets. Uses SQL and Python to support reporting, analytics, automation, dashboards, and operational decision-making. Partners across manufacturing and enterprise systems, investigates complex data challenges, evaluates emerging technologies, and supports advanced analytics and digital transformation initiatives.
Top Skills: AgileAzure Data FactoryData MartsData WarehousesDimensional ModelingEltErpETLGrafanaIotKepwareLabviewLeanMachine LearningMesOpc-UaPower BIPredictive ModelingPythonScadaScrumSemantic LayersSix SigmaSQLSQL ServerSsrsTableauTime-Series Databases
Reposted YesterdaySaved
Remote
Pittsburgh, PA
800K-1M Annually
Mid level
800K-1M Annually
Mid level
Information Technology • Professional Services • Consulting
Build and maintain large-scale data processing pipelines using Apache Spark (PySpark/Scala) and Python. Design ETL workflows, implement distributed data processing, use version control (Git), and work with cloud platforms (GCP/AWS/Azure).
Top Skills: SparkAWSAzureETLGCPGitPysparkPythonScala
YesterdaySaved
Remote
Pittsburgh, PA
175K-240K Annually
Senior level
175K-240K Annually
Senior level
Software
Own and evolve WorkOS’s internal data platform, including ingestion, orchestration, Snowflake, dbt models, data governance, reverse ETL, semantic views, and AI-enabled data workflows. Ensure pipeline reliability, freshness, security, and scalability while partnering with Product, Finance, RevOps, GTM, Engineering, and Security. Build monitoring, runbooks, CI/CD gates, access controls, masking policies, and infrastructure automation across AWS and Kubernetes.
Top Skills: AirflowAWSCi/CdDagsterDbtIamKubernetesPostgresPrefectPythonRbacSalesforceSlackSnowflakeSQLTerraform
Reposted YesterdaySaved
Remote
Pittsburgh, PA
Mid level
Mid level
Fintech • Software • Analytics • Financial Services
Design, build, and maintain scalable data pipelines, data lakes, and databases; ingest and map customer financial datasets; monitor pipeline reliability; translate business requirements into data flows and analytical insights to ensure high-quality, usable data.
Top Skills: APIsBigQueryBigtableData LakesDatabricksDbtFivetranGithub ActionsMariadbMongoDBMySQLNoSQLPostgresRedshiftSnowflakeSQL
YesterdaySaved
Remote
Pittsburgh, PA
77K-129K Annually
Mid level
77K-129K Annually
Mid level
Consulting
Build and maintain data pipelines, governance documentation, and data architectures supporting Department of State systems. Collect, clean, transform, integrate, validate, and analyze data from multiple sources using Palantir Foundry and relational databases. Collaborate with nontechnical stakeholders to gather requirements, identify data-quality issues, develop reporting tools and visualizations, and recommend improvements supporting data-driven decisions.
Top Skills: ArchibusData PipelinesKahuaPalantir FoundryPower BIPythonRelational DatabasesSQLTableau
YesterdaySaved
Remote
Pittsburgh, PA
Senior level
Senior level
Artificial Intelligence • Information Technology • Software • Consulting
Design, build, and optimize complex ETL pipelines, data workflows, monitoring, and alerting systems supporting legal and eDiscovery infrastructure. Lead end-to-end data engineering initiatives, improve query and pipeline performance, document deployment strategies, and ensure privacy, security, and compliance. Collaborate with attorneys, data scientists, engineers, and other stakeholders in regulated environments while contributing to internal consulting and thought leadership initiatives.
Top Skills: Ai/MlApache AirflowETLMySQLPostgresPrestoPythonSQL
YesterdaySaved
Remote
Pittsburgh, PA
Mid level
Mid level
Artificial Intelligence • Information Technology • Software • Consulting
Build and optimize scalable ETL pipelines, monitoring and alerting systems, and reliable data solutions for legal and eDiscovery infrastructure. Responsibilities include query and workflow optimization, cross-functional requirements gathering, system documentation, deployment planning, and compliance with data privacy and security standards. The role also contributes to internal consulting, thought leadership, and practice development.
Top Skills: Apache AirflowETLMySQLPostgresPrestoPythonSQL
YesterdaySaved
In-Office or Remote
Pittsburgh, PA
40K-73K Hourly
Junior
40K-73K Hourly
Junior
Healthtech
Designs and builds cloud-based, data-centric applications supporting clinical and operational healthcare processes. Develops data pipelines, transformations, enrichment processes, provisioning layers, and user interfaces. Collaborates with Product, Platform, and Architecture teams using modern software development practices, source control, documentation, and regular delivery methods.
Top Skills: Big DataCloud ComputingData PipelinesData ScienceData TransformationsSource Control
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
YesterdaySaved
Remote
Pittsburgh, PA
125K-135K Annually
Senior level
125K-135K Annually
Senior level
Consulting
Designs, builds, tests, deploys, and maintains secure ETL/ELT pipelines for public health data. Integrates Power Platform, Azure services, APIs, SQL Server, CDC/NIOSH systems, and approved file-transfer platforms. Develops batch, event-driven, and near-real-time data workflows supporting intake, validation, reporting, analytics, notifications, and results delivery while meeting FISMA Moderate, federal privacy, Agile, and DevSecOps requirements.
Top Skills: AgileAPIsDataverseDevsecopsEltETLAzureMicrosoft Power PlatformPower AppsPower AutomatePower PagesPythonSQLSQL Server
YesterdaySaved
In-Office or Remote
Pittsburgh, PA
Entry level
Entry level
Biotech • Agriculture
Designs and maintains scalable batch and streaming data pipelines and datasets in Databricks using Python, Spark, and SQL. Supports SQL Server, SSIS, SSRS, Azure, data warehousing, reporting, analytics, and AI/ML workloads. Responsibilities include optimizing performance and reliability, implementing data quality controls, integrating enterprise data sources, modernizing platforms, and partnering with business and IT teams to deliver scalable data solutions.
Top Skills: SparkDatabricksDelta LakeAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
29 Days AgoSaved
Easy Apply
Remote or Hybrid
Pittsburgh, PA
Easy Apply
Senior level
Senior level
Fintech • News + Entertainment • Software • Database • Financial Services
Lead the architecture and development of scalable AWS-based data ingestion, transformation, and orchestration pipelines. Build reliable data infrastructure using Python, SQL, Airflow, Lambda, ECS, SQS, and Terraform. Establish data modeling, quality, lineage, monitoring, testing, and observability practices while partnering with analysts, scientists, and backend engineers. Mentor senior engineers, guide technical decisions, and provide hands-on leadership for complex data platform initiatives.
Top Skills: Amazon EcsAmazon KinesisAmazon RedshiftAmazon S3Amazon SqsApache AirflowApache FlinkAWSAws GlueAws LambdaBeautifulsoupCi/CdDatabricksDockerGreat ExpectationsKafkaMonte CarloMwaaPythonScrapySnowflakeSQLTerraform
YesterdaySaved
In-Office or Remote
Pittsburgh, PA
Entry level
Entry level
Agriculture
Designs, builds, and maintains scalable batch and streaming data pipelines in Databricks using Python, Spark, SQL, and Delta Lake. Supports SQL Server, SSIS, and SSRS environments while improving performance, reliability, data quality, and scalability. Integrates enterprise data sources, applies warehousing best practices, and partners with business and IT teams to deliver reporting, analytics, and AI/ML data solutions.
Top Skills: SparkDatabricksDelta LakeAzureMicrosoft FabricPythonSQLSQL ServerSsisSsrsT-Sql
7 Days AgoSaved
Easy Apply
Remote or Hybrid
Pittsburgh, PA
Easy Apply
232K-348K Annually
Senior level
232K-348K Annually
Senior level
Artificial Intelligence • Cloud • Software
Lead the architecture and development of Vercel’s next-generation data platform, supporting batch and real-time integrations, analytics, data warehousing, and AI/ML workloads. Design scalable systems using Kafka, ClickHouse, Tinybird, and Snowflake; establish data governance and security standards; guide architectural decisions and roadmaps; collaborate with engineering, product, security, compliance, and leadership teams; write production code; and mentor engineers.
Top Skills: AWSAzureBig Data FrameworksClickhouseConfluent PlatformData GovernanceData WarehousingETLGCPKafkaKafka StreamsSnowflakeTinybird
2 Days AgoSaved
Remote
Pittsburgh, PA
Mid level
Mid level
Travel
Build and maintain scalable ETL/ELT pipelines, integrate travel and customer data from internal and external systems, develop cloud data warehouse models, implement data quality monitoring, support real-time streaming, and ensure data privacy and regulatory compliance. Collaborate with analysts, data scientists, and product teams to deliver reliable datasets for personalization, forecasting, pricing, route optimization, and customer experience initiatives.
Top Skills: AirflowAWSAzureBigQueryCcpaDbtGCPGdprIdmcJavaKafkaPythonRedshiftScalaSnowflakeSQL
Reposted 2 Days AgoSaved
Remote
Pittsburgh, PA
Mid level
Mid level
Healthtech
Design, build, and maintain scalable data pipelines, data models, and analytics solutions using Microsoft Fabric and Azure. Develop Lakehouse and Warehouse architectures, Fabric Notebooks with PySpark and Spark SQL, ETL/ELT workflows, and Power BI semantic models. Optimize performance, ensure data quality, security, governance, and regulatory compliance, and collaborate with analysts, data scientists, and business stakeholders to support business intelligence and advanced analytics.
Top Skills: Azure Data LakeAzure Data ServicesAzure Synapse AnalyticsData FactoryDataflows Gen2Delta LakeDirect LakeFabric LakehouseFabric PipelinesFabric WarehouseMicrosoft FabricOnelakePower BIPysparkPythonScalaSpark SqlSQL
2 Days AgoSaved
Remote
Pittsburgh, PA
Senior level
Senior level
Artificial Intelligence • Information Technology • Software • Consulting
Design and scale batch and streaming data pipelines, ETL/ELT workflows, data models, warehouses, quality frameworks, dashboards, and AI/ML feature pipelines. Automate reporting and analytics while collaborating with product, engineering, and operations teams. The role also includes internal consulting practice participation, thought leadership, client case studies, and contribution to digital transformation initiatives.
Top Skills: Adobe AnalyticsAirflowBigQueryCi/CdGenerative AiGitGoogle AnalyticsHiveInfrastructure As CodeKafkaLlmsPythonSnowflakeSparkSpark StreamingSQLTableau
2 Days AgoSaved
Remote
Pittsburgh, PA
Entry level
Entry level
Software
Build and maintain Databricks data pipelines across bronze, silver, and gold layers. Ingest APIs, logs, billing exports, and reference data; implement attribution logic, governance, data quality monitoring, and cost optimization. Manage Unity Catalog permissions, lineage, refresh schedules, incremental processing, and CI/CD workflows while supporting multi-cloud storage and high-volume caller-identity data.
Top Skills: SparkAsset BundlesAuto LoaderAWSAzureCi/CdDatabricksDatabricks WorkflowsDelta LakeGCPGitPysparkPythonRest ApisSQLUnity Catalog
Reposted 2 Days AgoSaved
Remote
Pittsburgh, PA
Internship
Internship
Consumer Web • Digital Media • eCommerce • News + Entertainment
Build and maintain ETL/ELT pipelines, ingest and transform structured and unstructured data, integrate APIs and external sources, improve data quality and monitoring, support dashboards and ML pipelines, optimize warehouse and cloud performance, and contribute to data governance and documentation.
Top Skills: APIsBigQueryEltETLGCPPostgresPythonSnowflakeSQL
8 Days AgoSaved
Remote or Hybrid
Pittsburgh, PA
165K-235K Annually
Senior level
165K-235K Annually
Senior level
Big Data • Cloud • Productivity • Software • Database • Analytics • Automation
Build and maintain Databricks-based data platforms, including ingestion, transformation, storage, governance, data modeling, and serving pipelines. Establish medallion architecture standards, canonical data models, quality controls, lineage, schema evolution, and reliable batch or incremental processing. Improve pipeline observability, scalability, idempotency, and recoverability while moving curated data to systems such as ClickHouse. Collaborate across application and analytics teams to create durable, governed production datasets.
Top Skills: Amazon AuroraAmazon RdsApache AirflowSparkBigQueryCdcClickhouseCloud Object StorageDatabricksDelta LakeIamOpenmetadataPostgresSnowflakeUnity Catalog
Reposted 2 Days AgoSaved
Remote
Pittsburgh, PA
Senior level
Senior level
Information Technology • Professional Services • Consulting
Design, build, and maintain Keboola-based data pipelines integrating multiple business systems to deliver accurate, scalable data for supply chain, logistics, and transportation analytics; collaborate with analysts and engineers; support documentation, governance, and pipeline automation.
Top Skills: APIsAzure DevopsAzure Devops BoardsCi/CdGitJSONKeboolaPythonSnowflakeSQL
13 Days AgoSaved
In-Office
Pittsburgh, PA
Mid level
Mid level
Edtech
Designs, develops, tests, maintains, and optimizes data pipelines and software systems using SQL, Python, and Apache Airflow. Supports technical projects, documents systems, troubleshoots application performance, evaluates user requirements, and develops technical specifications. Collaborates with epidemiologists, statisticians, and researchers, while mentoring staff and translating academic research needs into software solutions.
Top Skills: Apache AirflowDjangoMySQLPostgresPythonSQL
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account