Data Engineer

Posted 16 Days Ago
Arlington, VA, USA
In-Office
Mid level
Mobile • Software
The Role
Design, build, and maintain Databricks-based lakehouse data pipelines and products. Develop ETL/ELT (batch and streaming) using Python, PySpark, Spark, SQL, and Delta Lake; manage notebooks, jobs, clusters, and CI/CD; implement data quality, monitoring, lineage, governance, and security to support analytics and AI/ML for mission-critical defense workloads.
Summary Generated by Built In

540 is seeking a Data Engineer to support a mission-critical technology modernization effort for the Department of War. You will design, build, and maintain Databricks-based data pipelines and lakehouse capabilities that enable secure data integration, analytics, AI/ML, and operational workloads at enterprise scale.

Working with software engineers, AI/ML engineers, cybersecurity teams, and mission stakeholders, you will build scalable and reliable solutions using Databricks, Python, Apache Spark, and Delta Lake. The ideal candidate enjoys solving complex engineering challenges and developing trusted data products that support national defense missions.

Location: Arlington, VA
Citizenship & Clearance Requirement: Per client requirements, candidates must be U.S. Citizens with an active DoW Secret (or higher) clearance
Education Requirement: Bachelor’s degree in Computer Science, Engineering, or a related technical field preferred; equivalent combinations of education and relevant experience will be considered
540 Internal Thrive Level: Data Engineer II or III

WHY 540?

540 is a forward-thinking company that the government turns to in order to #getshitdone. We don’t just talk about innovation – we deliver it. We break down barriers, build impactful technology, and solve mission-critical problems.

HOW YOU’LL DRIVE IMPACT

  • Design, develop, and maintain Databricks-based data pipelines, data products, and lakehouse capabilities
  • Build automated ETL/ELT pipelines that ingest, transform, and deliver mission-critical data
  • Develop production-grade data-processing solutions using Python, SQL, PySpark, Apache Spark, and Delta Lake
  • Design and maintain data models, schemas, tables, and medallion architecture patterns supporting analytical, operational, and AI/ML workloads
  • Build and operate batch and streaming data pipelines supporting mission requirements
  • Develop and manage Databricks notebooks, jobs, workflows, clusters, and compute resources
  • Implement data-quality checks, automated testing, monitoring, lineage, and metadata-management capabilities
  • Support data discovery, governance, and access controls using Unity Catalog or similar technologies
  • Optimize Spark workloads and Databricks resources for performance, scalability, reliability, and cost efficiency
  • Collaborate with engineers, analysts, and data scientists to deliver reusable data products and mission capabilities
  • Support Databricks deployments using CI/CD, infrastructure as code, and source control
  • Partner with cybersecurity teams to implement data-protection, access-control, auditing, and governance requirements
  • Troubleshoot issues affecting Databricks workloads, data pipelines, storage systems, and production data services
  • Document data models, pipeline designs, engineering processes, and operational procedures

REQUIRED SKILLS & EXPERIENCE

  • 4+ years of relevant data engineering or software engineering experience
  • Hands-on experience developing and operating production data pipelines using Databricks
  • Proficiency with Python, SQL, PySpark, Apache Spark, and Delta Lake
  • Experience building automated ETL/ELT pipelines for large-scale datasets
  • Experience designing and maintaining data models, schemas, tables, and lakehouse architectures
  • Experience managing Databricks notebooks, jobs, workflows, and compute resources
  • Experience implementing data quality, automated testing, monitoring, lineage, or metadata-management capabilities
  • Experience working with Databricks and cloud-based data services in AWS, Azure, or Google Cloud
  • Experience working with structured, semi-structured, and unstructured data
  • Understanding of lakehouse architecture, data governance, security, privacy, and access-control principles
  • Ability to troubleshoot data pipelines, Spark workloads, infrastructure, and applications

NICE TO HAVE

  • Databricks certification or equivalent demonstrated platform expertise
  • Experience supporting DoW, federal, Advana, or other enterprise data environments
  • Experience with CI/CD, infrastructure as code, automated testing, and source control
  • Experience using Unity Catalog for data governance, lineage, and access control
  • Experience developing streaming pipelines with Spark Structured Streaming, Kafka, Kinesis, or Pulsar
  • Experience with orchestration tools such as Airflow, Dagster, or Argo Workflows
  • Experience with Docker, Kubernetes, or other containerization and orchestration technologies
  • Experience building cloud-native data platforms in secure, regulated, classified, or mission-critical environments
  • Currently holds, or is willing to obtain within 30 days of employment, an approved certification such as Cloud+, GSEC, Security+, or SSCP

BENEFITS & PERKS

  • Flexible PTO + all Federal holidays off
  • Health, dental and vision insurance plans
  • Flexible Spending Account (FSA)
  • 401k with employer match
  • Company-sponsored life insurance, short- and long-term disability 
  • Professional development (training, certifications, conferences)
  • Paid cloud developer accounts
  • Referral bonuses
  • HQ office perks (parking / metro reimbursement, nitro coffee & lunches) 
  • Annual social events (540 Week, hackathon, charity golf tournament, etc.)
  • Access to 540’s Washington Capitals & Nationals tickets

EQUAL EMPLOYMENT OPPORTUNITY (EEO)

540's policy is to provide equal employment opportunity to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

Skills Required

  • 4+ years of relevant data engineering or software engineering experience
  • U.S. Citizen with an active Department of War (DoW) Secret or higher clearance
  • Hands-on experience developing and operating production data pipelines using Databricks
  • Proficiency with Python
  • Proficiency with SQL
  • Proficiency with PySpark and Apache Spark
  • Experience with Delta Lake and lakehouse architectures
  • Experience building automated ETL/ELT pipelines for large-scale datasets (batch and streaming)
  • Experience managing Databricks notebooks, jobs, workflows, clusters, and compute resources
  • Experience implementing data quality, automated testing, monitoring, lineage, or metadata-management capabilities
  • Experience with cloud-based data services in AWS, Azure, or Google Cloud
  • Experience working with structured, semi-structured, and unstructured data
  • Understanding of data governance, security, privacy, and access-control principles
  • Ability to troubleshoot data pipelines, Spark workloads, infrastructure, and production data services
  • Bachelor's degree in Computer Science, Engineering, or related technical field (or equivalent experience)
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Arlington, VA
64 Employees
Year Founded: 2013

What We Do

540 is a technology consulting firm who helps our government and business clients innovate like start-ups. Through hard work, perseverance, and deep commitment we remove barriers to innovation which enables us to build impactful tech for the government. Contact us for: - Tech Strategy - Software Development - API Design / Dev - DevOps Email us: [email protected]

Similar Jobs

Samsara Logo Samsara

Data Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
United States
4000 Employees
118K-179K Annually

PwC Logo PwC

Data Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
34 Locations
370000 Employees
77K-202K Annually

PwC Logo PwC

Data Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
35 Locations
370000 Employees
99K-232K Annually

PwC Logo PwC

Data Engineer

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
65 Locations
370000 Employees
99K-232K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account