Principal Data Engineer

Posted Yesterday
Easy Apply
Hiring Remotely in United States
Remote
200K-220K Annually
Senior level
Information Technology • Software • Analytics
The Role
The Principal Data Engineer will lead the Data Resolution team, focusing on complex data challenges using Spark, system architecture, and mentorship, to optimize graph data pipelines.
Summary Generated by Built In
About Sayari: 

Sayari is a risk intelligence provider that equips the public and private sectors with immediate visibility into complex commercial relationships by delivering the largest commercially available collection of corporate and trade data from over 250 jurisdictions worldwide. Sayari's solutions enable risk resilience, mission-critical investigations, and better economic decisions. 

Headquartered in Washington, D.C., its solutions are trusted by Fortune 500 companies, financial institutions, and government agencies, and are used globally by thousands of users in over 35 countries. Funded by world-class investors, with a strategic $228 million investment by TPG Inc. (NASDAQ: TPG) in 2024, Sayari has been recognized by the Inc. 5000 and the Deloitte Technology Fast 500 as one of the fastest growing private companies in the United States and was featured as one of Inc.’s “Best Workplaces” for 2025.

POSITION DESCRIPTION

We are looking for a Principal Data Engineer to join our Data Resolution team and serve as a technical anchor for our most complex data challenges. In this role, you will be a "player-coach," spending the majority of your time (70%) hands-on with Spark and graph data logic while dedicating the remainder of your time to system architecture, design planning, and technical mentorship. You will be instrumental in evolving our graph build pipelines, optimizing our cloud footprint, and overseeing the long-term planning and execution of major data pipeline re-architectures. This is a high-impact role where your work directly powers the data products used by global systems defenders.


JOB RESPONSIBILITIES
  • Design and implement complex Spark data logic, focusing on performance optimization, data volume tuning, and robust execution.
  • Own the architectural design of graph build pipelines, ensuring they are scalable, automated, and highly resilient.
  • Plan and oversee the strategic re-architecture of data pipelines to meet evolving business needs and scale.
  • Optimize infrastructure-as-code and schema designs to reduce cloud costs and improve pipeline latency.
  • Act as a technical consultant for the team, fostering a collaborative and engineer-led approach to design decisions.
  • Support the development of the engineering team through code reviews, design docs, and architectural best practices.
  • Ensure the accuracy of mission-critical data outputs.
SKILLS & EXPERIENCE

Required Skills & Experience

  • 8+ years of experience in the big data space, with a proven track record of implementing large-scale features and leading process redesigns.
  • Expert-level mastery of Apache Spark for large-scale data processing.
  • Strong experience with orchestration tools (Airflow) and cloud computing environments.
  • Hands-on experience architecting and managing data flows into databases such as Elasticsearch, Memgraph, and Cassandra.
  • Demonstrated ability in system architecture, including Infrastructure as Code (IaC) and schema design.
  • A "builder" mindset with experience evolving and improving existing architectures to meet new scale requirements.

Preferred Skills & Experience

  • Experience working specifically with graph data or graph databases.
  • Prior experience with entity resolution or identity resolution systems.
  • Experience evaluating and selecting modern analytical database architectures.

The target base salary for this position is $200,000-$220,000 plus company bonus and equity. Final offer amounts are determined by multiple factors including location, local market variances, candidate experience and expertise, internal peer equity, and may vary from the amounts listed above.


Benefits: 
  • 100% fully paid medical, vision, and dental for employees and their dependents
  • Generous time off; we observe all US federal holidays, close our office for a winter break (12/24-12/31), in addition to granting 18 PTO days and 10 sick days
  • Outstanding compensation package; competitive commissions for revenue roles and quarterly bonuses for non-revenue positions
  • A strong commitment to diversity, equity, and inclusion
  • Eligibility to participate in additional benefits such as 401k match up to 5%, 100% paid life insurance (up to $100,000 coverage),, and parental leave
  • A collaborative and positive culture - your team will be as smart and driven as you
  • Limitless growth and learning opportunities
 
Sayari is an equal opportunity employer and strongly encourages diverse candidates to apply. We believe diversity and inclusion mean our team members should reflect the diversity of the United States. No employee or applicant will face discrimination or harassment based on race, color, ethnicity, religion, age, gender, gender identity or expression, sexual orientation, disability status, veteran status, genetics, or political affiliation. We strongly encourage applicants of all backgrounds to apply.
Pay Range
$200,000$220,000 USD

Top Skills

Airflow
Spark
Cassandra
Elasticsearch
Infrastructure As Code (Iac)
Memgraph
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Washington, DC
84 Employees
Year Founded: 2015

What We Do

The world’s largest provider of companies, their key people, and their most important relationships. From financial intelligence to anti-counterfeiting, and from free trade zones to war zones, Sayari powers cross-border and cross-lingual insight into customers, counterparties, and competitors. Thousands of analysts and investigators in over 30 countries rely on our products to safely conduct cross-border trade, research front-page news stories, confidently enter new markets, and prevent financial crimes such as corruption and money laundering.

Similar Jobs

CrowdStrike Logo CrowdStrike

Data Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
USA
10000 Employees
170K-260K Annually

NVIDIA Logo NVIDIA

Principal Software Engineer

Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
In-Office or Remote
Santa Clara, CA, USA
21960 Employees
272K-489K Annually
In-Office or Remote
El Segundo, CA, USA
215K-270K Annually

Autodesk Logo Autodesk

Machine Learning Engineer

Big Data • Cloud • Digital Media • Machine Learning • Mobile • Software • Industrial
In-Office or Remote
8 Locations
13285 Employees

Similar Companies Hiring

Scotch Thumbnail
Software • Retail • Payments • Fintech • eCommerce • Artificial Intelligence • Analytics
US
25 Employees
Milestone Systems Thumbnail
Software • Security • Other • Big Data Analytics • Artificial Intelligence • Analytics
Lake Oswego, OR
1500 Employees
Fairly Even Thumbnail
Software • Sales • Robotics • Other • Hospitality • Hardware
New York, NY

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account