Position Overview
SteerBridge seeks a highly skilled and motivated individual to join our team as a Senior Data Engineer to align data solutions to business requirements by planning and managing data infrastructure and strategy for our Modern Disability Claims AI/ML program. Our team is dedicated to harnessing the power of AI/ML to increase claims processing throughput and reduce adjudication wait times, ultimately improving outcomes for veterans.
Key Responsibilties
Perform data engineering activities across existing systems of record and multiple databases.
Enhance and optimize data entry, management, and extraction processes to improve data usability within proprietary systems.
Conduct data quality checks to identify inconsistencies, errors, and opportunities for improvement.
Analyze data and present findings to support business and operational needs.
Maintain accurate documentation of data processes, workflows, and methodologies.
Collaborate with team members and stakeholders to address data needs and support continuous improvement.
Eligibility requirements: U.S. citizenship is required for this position under applicable federal contract requirements. Candidate must also be able to obtain and maintain a Public Trust clearance; an active Secret or Top Secret security clearance also satisfies this requirement.
Bachelor’s degree or higher in Systems Engineering, Computer Science, Data Science, or a related field.
6+ years of data engineering or related experience designing, developing, and supporting production data platforms and pipelines.
Strong proficiency with Python, SQL, Pandas, PySpark, NumPy, and Git, including experience developing reusable, testable, and performance-optimized data processing and automation solutions.
Experience designing conceptual, logical, and physical data models and working with relational, NoSQL, data warehouse, data lake, and lakehouse architectures.
Hands-on experience developing and orchestrating scalable batch and/or real-time data pipelines using technologies such as Apache Spark, Kafka, Airflow, NiFi, AWS Glue, Azure Data Factory, or GCP Dataflow.
Experience developing cloud-based data solutions in AWS, Azure, and/or GCP, including cloud storage, managed databases, compute services, and modern data warehousing platforms such as Redshift, Snowflake, or BigQuery.
Experience with data quality, governance, security, lineage, metadata management, performance optimization, CI/CD, and software engineering best practices for large-scale data environments.
Preference for candidates located in the Vienna, VA area who are able to work onsite at the SteerBridge Vienna office at least three days per week. Hybrid arrangements may be available at the supervisor’s discretion.
Experience with distributed computing and modern data lake/lakehouse technologies such as Hadoop, Spark, Hive, Presto/Trino, Delta Lake, Apache Iceberg, or Apache Hudi.
Experience with DevOps/DataOps practices, including Infrastructure as Code (IaC), Docker, Kubernetes, Git-based workflows, automated testing, and CI/CD for data infrastructure and pipelines.
Experience optimizing large-scale data environments for query performance, pipeline efficiency, scalability, and cloud resource utilization and cost.
Experience developing resilient, automated data pipelines with monitoring, alerting, retry logic, failure recovery, and other self-healing capabilities.
Familiarity with AI/ML data pipeline technologies such as TensorFlow, PyTorch, Scikit-learn, MLflow, Kubeflow, or feature stores.
Experience mentoring junior engineers, conducting technical reviews, and clearly documenting technical architectures using tools such as Lucidchart, PlantUML, or Draw.io.
Benefits
- Health insurance
- Dental insurance
- Vision insurance
- Life Insurance
- 401(k) Retirement Plan with matching
- Paid Time Off
- Paid Federal Holidays
Skills Required
- U.S. citizenship
- Bachelor's degree or higher in Systems Engineering, Computer Science, or a related field
- Must hold or be able to obtain a Public Trust clearance; active Secret or Top Secret clearance also qualifies
- At least 6 years of experience
- Experience with data pipelines, advanced analytics platforms, and Python
- Experience scripting, developing tooling, and automating large-scale computing environments
- Extensive experience with Python, Pandas, PySpark, NumPy, SciPy, SQL, and Git
- Minor experience with TensorFlow, PyTorch, and Scikit-learn
- Advanced conceptual, logical, and physical data modeling
- Experience with relational, NoSQL, graph, time-series, and document databases
- Experience with modern data warehousing platforms such as Redshift, Snowflake, or BigQuery
- Experience designing OLTP and OLAP systems, metadata-driven pipelines, schema evolution, and data versioning
- Hands-on experience with Kafka, Airflow, Spark, Flink, or NiFi
- Experience building batch and real-time cloud pipelines using AWS Glue, GCP Dataflow, or Azure Data Factory
- Deep expertise in AWS, GCP, or Azure data ecosystems
- Experience with cloud data lakes, data warehouses, data mesh, cloud storage, managed databases, and distributed compute
- Experience with Hadoop, Spark, Hive, Presto, or Trino and lakehouse architectures
- Advanced SQL and NoSQL query tuning, indexing, sharding, partitioning, replication, backup, and disaster recovery
- Experience implementing data privacy, compliance, cataloging, lineage, metadata management, access controls, and data quality testing
- Strong Python and SQL proficiency, object-oriented programming, design patterns, algorithmic complexity, and performance optimization
- Experience with version control, CI/CD, testing, documentation, type hinting, linting, debugging, and profiling
- Experience supporting feature engineering, model deployment, ML orchestration, experiment tracking, feature stores, and real-time inference
- Experience mentoring engineers and leading cross-functional data initiatives
- Experience with infrastructure as code, Docker, Kubernetes, automated deployment, and DataOps
- Preferred local to Vienna, Virginia, and able to work onsite at least 3 days per week
What We Do
SteerBridge is a technology company that provides professional services and solutions to the U.S. Government and Corporate counterparts through a wide array of capabilities and myopic focus on innovative solutions. We leverage decades of federal acquisition and private sector experience to deliver best in class commercial solutions while maximizing Veteran talent to enhance efficiency and surpass expectations. Current & Past Services include: Veteran Relations | Strategic Communications | General Technology Services | Software Engineering | Program Management | Business Process Management | Cybersecurity Professional Services | AI/ML capabilities | Data Management & Analytics.









