Principal Data Engineer

Posted 5 Days Ago
2 Locations
In-Office
Expert/Leader
Software
The Role
Leads enterprise data platform architecture and modernization, including migration from Azure SQL Server to Databricks, Lakehouse design, Kafka streaming, Spark optimization, governance, reliability, and CI/CD. Enables AI and agentic data capabilities through governed Delta tables, feature stores, vector stores, RAG pipelines, and Databricks tools. Defines engineering standards, security, disaster recovery, and observability practices while mentoring engineers and partnering cross-functionally on scalable data solutions.
Summary Generated by Built In

Overview

The Principal Data Engineer is a senior technical leader within AvidXchange's Data Engineering organization responsible for architecting, building, and scaling modern data platforms. In this role, you will drive the migration of legacy Azure SQL Server workloads to Databricks, design real-time streaming pipelines with Apache Kafka, enable AI and agentic capabilities such as Databricks Genie, and define the long-term data architecture strategy. You will partner closely with Software Engineering, Product, Architecture, DevOps, and Analytics teams to deliver secure, high-performing, and reliable data solutions at enterprise scale.


What You'll Do

Data Platform Architecture & Modernization

·       Lead the design and implementation of scalable, cloud-native data architectures on Databricks (Delta Lake, Unity Catalog, Lakehouse patterns).

·       Own and execute the migration strategy from legacy Azure SQL Server to Databricks, including schema translation, ETL/ELT re-platforming, data validation, and cutover planning.

·       Define data modeling standards (medallion architecture, star/snowflake schemas) and ensure consistency across all pipelines and domains.

·       Evaluate and recommend tools, frameworks, and platforms to support long-term data strategy and organizational goals.

·       Collaborate with Solution and Enterprise Architects to review and approve new data architecture designs.

Streaming & Real-Time Data Engineering

·       Architect and implement Kafka-based streaming pipelines for real-time data ingestion, transformation, and delivery.

·       Design event-driven architectures and streaming topologies using Kafka Streams, ksqlDB, or Spark Structured Streaming on Databricks.

·       Establish patterns for schema management (Confluent Schema Registry), consumer group strategy, offset management, and dead-letter queuing.

·       Ensure streaming pipelines meet SLA requirements for latency, throughput, and fault tolerance.

Optimization, Quality & Standards

·       Debug and optimize Spark jobs, Delta Lake tables, and SQL workloads for performance, cost efficiency, and maintainability.

·       Lead code reviews focused on senior engineers to enforce standards, best practices, and technical quality.

·       Manage pipeline quality, data models, and CI/CD delivery workflows; guide teams on continuous improvement.

·       Promote strong data management practices — data quality, lineage, observability, and governance.

·       Identify opportunities to improve service delivery methods, processes, and resource utilization.

Leadership, Mentorship & Strategy

·       Mentor data engineers at all levels, with particular emphasis on developing senior talent.

·       Establish and evolve data engineering standards, best practices, and management of technical debt.

·       Stay current on data platform trends (Databricks releases, Kafka ecosystem, open table formats) and contribute to long-term architectural vision.

·       Develop plans for data security, disaster recovery, backup, business continuity, and archiving across the Lakehouse.

AI, Agentic Capabilities & Intelligent Data Products

·       Design and enable AI and agentic capabilities on the Databricks platform, including Databricks Genie for natural language data exploration and self-service analytics.

·       Architect data foundations — clean, governed, well-documented Delta tables — that power Genie spaces, AI/BI dashboards, and LLM-driven data agents.

·       Collaborate with ML and AI teams to build and maintain feature stores, vector stores, and retrieval-augmented generation (RAG) pipelines on Databricks.

·       Evaluate and integrate emerging agentic frameworks (LangChain, Mosaic AI Agent Framework) to automate data workflows and enable intelligent data products.

·       Define governance and observability standards for AI-driven data pipelines, ensuring reliability, auditability, and responsible AI practices.

Cross-Functional Collaboration

·       Partner with project managers and business leaders on initiatives involving enterprise data.

·       Collaborate across teams to influence and strengthen data engineering practices organization-wide.

·       Work with Analytics, ML, and product engineers to design and deliver end-to-end data solutions that meet business needs.

What We're Looking For

Required

·       Bachelor's degree in Computer Science, Engineering, or a related field with 10+ years of data engineering experience in a high-availability, business-critical environment.

·       Hands-on Databricks expertise: Delta Lake, Unity Catalog, Databricks Workflows, Databricks SQL, and cluster/job optimization.

·       Proven experience migrating from legacy Azure SQL Server (or other relational RDBMS) to a Databricks Lakehouse — schema translation, data validation, and cutover strategies.

·       Strong proficiency in Apache Kafka for streaming pipelines — producers/consumers, topic design, partitioning strategy, and Kafka Connect.

·       Expert-level PySpark and/or Scala Spark skills, including performance tuning, broadcasting, partitioning, and caching.

·       Deep understanding of data architecture patterns: medallion (bronze/silver/gold), Lambda/Kappa, event sourcing, and streaming-first designs.

·       Hands-on experience with Azure cloud services (ADLS Gen2, Azure Event Hubs, ADF, Azure Key Vault, Azure Monitor).

·       Strong knowledge of infrastructure components — networking, cloud storage (Delta/Parquet), and cloud cost optimization.

Preferred

·       Databricks Certified Data Engineer Associate or Professional certification.

·       Experience with ksqlDB, Kafka Streams, or Spark Structured Streaming for stateful stream processing.

·       Confluent Platform experience: Schema Registry, Kafka Connect connectors, RBAC, and cluster management.

·       Advanced Azure SQL Server expertise (2018+/Azure SQL MI) — stored procedures, indexing strategies, query plan analysis — valuable for migration contexts.

·       Experience with dbt (data build tool) for SQL-based transformation layers on Databricks.

·       Proficiency with Git and CI/CD pipelines for data engineering (Azure DevOps, GitHub Actions).

·       Experience working in Agile environments (Scrum/Kanban).

·       Familiarity with data governance frameworks, data cataloging (Unity Catalog, Microsoft Purview), and data quality tooling (Great Expectations, Monte Carlo).

·       Familiarity with secure coding practices, including OWASP Top 10 and secrets management.

·       Experience implementing DataOps practices: automated testing, data observability, pipeline CI/CD, data contracts, and SLA monitoring across the Lakehouse.

·       Hands-on MLOps experience: model versioning (MLflow), experiment tracking, model serving, and integrating ML pipelines with production data workflows on Databricks.

·       Familiarity with Databricks Mosaic AI (formerly MLflow + Model Serving) for end-to-end MLOps lifecycle management.

·       Experience with real-time ML feature stores or Lakehouse-based ML pipelines is a plus.


About AvidXchange

AvidXchange is a leading provider of accounts payable (“AP”) automation software and payment solutions for middle-market businesses and their suppliers. By trade, we are a technology company, but if you ask anyone who works here, they’ll tell you our people are at the core of who we are. At AvidXchange, mindset is everything. We are Connected as People, Growth Minded, and Customer Obsessed. These three mindsets represent our culture – who we are, who we’ve always been, and they guide us to improve every day. Since our founding in 2000 in Charlotte, NC, we’ve created a company of over 1,500 teammates working across the U.S., or remotely. AvidXchange is proud to be Certified™ as a Great Place to Work®. The prestigious recognition is based on anonymous data from our teammates and makes official what our teammates have known for years – that AvidXchange is a Great Place to Work®. 

Who you are: 

  • A go-getter with an entrepreneurial mindset – that means you are not afraid of taking risks, winning big or facing the unknown. 
  • Someone who understands that business is people centric. Connecting with others as humans first allows you to develop mutually beneficial working relationships. 
  • Focused on making a difference for our customers. AvidXchange exists to help solve complex problems for our customers so we can all realize our potential. 

What you’ll get:  

AvidXchange teammates (we call them AvidXers) get the perks and prestige of a growing tech company paired with the flexibility of a founder-led startup. We help our AvidXers develop as professionals and as human beings, providing work/life balance, development programs, and competitive benefits. At AvidXchange, we are building more than a tech company – we are building an experience. We remain committed to a culture where you can fully be 'you’ – connected with others, chasing big goals, and making a meaningful impact. If you want to help us grow while realizing your potential and creating stories you’ll tell for years, you’ve come to the right place.

AvidXers enjoy:  

  • 18 days PTO* 
  • 11 Holidays (8 company recognized & 3 floating holidays) 
  • 16 hours per year of paid Volunteer Time Off (VTO) 
  • Competitive Healthcare 
    • High Deductible Heath Plan Option that has $0 monthly premium for teammate-only coverage 
    • 100% AvidXchange paid Dental Base Plan Coverage
    • 100% AvidXchange paid Life Insurance 
    • 100% AvidXchange paid Long-Term Disability 
    • 100% AvidXchange paid Short-Term Disability  
    • Employee Assistance Program (EAP) - Provides counseling services, legal and financial consultations and health advocacy for Teammates and their eligible dependents
    • Onsite Health Clinic with Atrium Health - available to Teammates and their eligible dependents
  • 401(k) Match: 100% match on the first 3% of your salary, plus 50% match on the next 2%
  • Parental Leave: 8 weeks 100% paid by AvidXchange** 
  • Discounts on Pet, Home, and Auto insurance 
  • WeeCare Childcare Service: helps teammates find affordable daycare, childcare, and tutors 40% less expensive than traditional daycare centers 
  • Perks at Work: free discount program that provides teammates the opportunity to save on items from electronics, movie tickets, car buying, vacations, and more 
  • Onsite gym fitness center, yoga studio, and basketball court
  • Tuition Reimbursement up to the federal maximum of $5,250***
  • Hybrid Workplace Flexibility
  • Free parking

*Fully granted from beginning of year, pro-rated if hired mid-year 

**Must be full-time for at least 3 months

***Must be full-time for at least one year 

Equal Employment Opportunity

AvidXchange is an equal opportunity employer. AvidXchange is committed to equal employment opportunity in accordance with applicable federal, state, and local laws. AvidXchange will not discriminate against applicants for employment on any legally recognized basis. This includes, but is not limited to veteran status, race, color, religion, sex, sexual orientation, gender identity, gender expression, national origin, age and physical or mental disability. 

Skills Required

  • Bachelor’s degree in Computer Science, Engineering, or a related field
  • 10+ years of data engineering experience in a high-availability, business-critical environment
  • Hands-on expertise with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, Databricks SQL, and cluster/job optimization
  • Experience migrating legacy Azure SQL Server or another relational RDBMS to a Databricks Lakehouse
  • Strong Apache Kafka experience, including producers, consumers, topic design, partitioning, and Kafka Connect
  • Expert PySpark and/or Scala Spark skills, including performance tuning, broadcasting, partitioning, and caching
  • Deep understanding of medallion, Lambda, Kappa, event sourcing, and streaming-first data architecture patterns
  • Hands-on experience with Azure cloud services including ADLS Gen2, Azure Event Hubs, ADF, Azure Key Vault, and Azure Monitor
  • Knowledge of networking, cloud storage, Delta/Parquet, and cloud cost optimization
  • Databricks Certified Data Engineer Associate or Professional certification
  • Experience with ksqlDB, Kafka Streams, or Spark Structured Streaming
  • Confluent Platform experience with Schema Registry, Kafka Connect connectors, RBAC, and cluster management
  • Advanced Azure SQL Server expertise, including stored procedures, indexing, and query plan analysis
  • Experience with dbt for SQL-based transformation layers on Databricks
  • Proficiency with Git and CI/CD pipelines, including Azure DevOps or GitHub Actions
  • Experience working in Agile environments such as Scrum or Kanban
  • Familiarity with data governance, data cataloging, and data quality tools
  • Familiarity with secure coding practices, OWASP Top 10, and secrets management
  • Experience implementing DataOps practices, automated testing, observability, data contracts, and SLA monitoring
  • Hands-on MLOps experience with MLflow, model serving, and ML pipeline integration
  • Familiarity with Databricks Mosaic AI for MLOps lifecycle management
  • Experience with real-time ML feature stores or Lakehouse-based ML pipelines

AvidXchange Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about AvidXchange and has not been reviewed or approved by AvidXchange.

  • Healthcare Strength — Healthcare coverage is positioned as comprehensive, spanning medical, dental, vision, disability, life insurance, mental health support, and FSAs. On-site fitness and wellness resources at the Charlotte headquarters further strengthen the perceived health-and-wellbeing offering.
  • Retirement Support — Retirement support includes a 401(k) with a company match and immediate vesting, which can improve perceived value early in tenure. This feature supports longer-term financial planning without delayed eligibility.
  • Leave & Time Off Breadth — Time-off options include PTO, floating holidays, and paid volunteer time, providing multiple avenues for planned and purpose-driven time away from work. These programs can improve overall perceived total rewards even when base-pay sentiment varies.

AvidXchange Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Charlotte, NC
1,300 Employees
Year Founded: 2000

What We Do

AvidXchange is the accounts payable automation industry leader for mid-market businesses serving more than 5,550 customers & 400,000 suppliers nationwide.

Similar Jobs

Zscaler Logo Zscaler

Sales Engineer

Cloud • Information Technology • Security • Software • Cybersecurity
Easy Apply
Remote or Hybrid
USA
8697 Employees
176K-251K Annually

Applied Systems Logo Applied Systems

Data Engineer

Artificial Intelligence • Cloud • Payments • Software • Business Intelligence • Generative AI • Automation
Remote or Hybrid
United States
3116 Employees
160K-260K Annually

Citizens Logo Citizens

Data Engineer

Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
In-Office or Remote
2 Locations
17000 Employees

Ahold Delhaize USA Logo Ahold Delhaize USA

Data Engineer

AdTech • eCommerce • Food • Marketing Tech • Retail
Hybrid
Salisbury, NC, USA
10000 Employees
163K-245K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account