Senior Data Scientist - Machine Learning

Posted Yesterday
Be an Early Applicant
Hiring Remotely in United States
Remote
123K-167K Annually
Senior level
Aerospace • Information Technology • Professional Services • Security • Software
The Role
Design, train and validate supervised ML models on multi-billion claim records to score fraud risk; engineer features in-warehouse; deploy and monitor models in production; provide human-readable rationale and claim-level evidence; collaborate with subject-matter experts and stakeholders to ensure actionable, explainable outputs.
Summary Generated by Built In

Type of Requisition:

Regular

Clearance Level Must Currently Possess:

None

Clearance Level Must Be Able to Obtain:

None

Public Trust/Other Required:

None

Job Family:

Data Science and Data Engineering

Job Qualifications:

Skills:

Amazon Web Services (AWS), Healthcare Claims, Predictive Modeling, Python (Programming Language), Supervised Learning

Certifications:

None

Experience:

5 + years of related experience

US Citizenship Required:

No

Job Description:

As the Senior Data Scientist for Machine Learning supporting the Healthcare Fraud Prevention Partnership (HFPP), you will be the first dedicated machine learning practitioner at the Trusted Third Party (TTP), an established Fraud, Waste and Abuse (FWA) analytics program. You will develop predictive models against a multi-billion record claims warehouse assembled from dozens of public and private healthcare payers, and you will establish how machine learning models move from development into production on this program.

The data, the subject matter experts and the payer partnerships are already in place; the modeling capability is yours to build. This is a senior individual contributor position without direct reports, and it is the only role on the team focused primarily on machine learning, meaning the Senior Data Scientist will be establishing practice rather than joining one.

***Work visa sponsorship will not be provided for this position. This is a remote role. Candidates must reside in the United States.

MEANINGFUL WORK AND PERSONAL IMPACT:

  • Designing, training and validating supervised models that score providers and billing patterns for FWA risk, using investigative case-level data, payer feedback on referred leads, and public exclusion and enforcement data as labels, including the entity resolution to link enforcement records to providers in claims.

  • Designing validation for the actual conditions: labels lagging billing behavior by years, coverage limited to leads previously referred, extreme class imbalance, and schemes that shift faster than confirmation arrives.

  • Engineering features against billions of claim records within the warehouse rather than extracting data to local memory, using Python and SQL, alongside data engineers and Business Intelligence Developers.

  • Delivering output that supports action. Investigators need the specific claims, the pattern and the basis for the finding, so each model carries a human-readable rationale and claim-level evidence alongside the score, adjusted for case mix and specialty and ranked so that precision at the top of the review queue is the operative measure.

  • Deploying models into production and keeping them healthy, including scheduled execution, versioning and drift monitoring, and establishing the modeling and deployment practices the Data Science team adopts going forward.

  • Collaborating with FWA Subject Matter Experts to separate genuine anomalies from patterns explained by coverage policy or claim edits, and communicating methodology and limitations to HFPP Partners and stakeholders so that output is adopted and acted upon.

WHAT YOU'LL NEED TO SUCCEED:

  • Master's in a quantitative field (statistics, computer science, engineering, applied mathematics or related), or a Bachelor's with equivalent hands-on experience.

  • 5+ years building, validating and delivering supervised machine learning models on real-world data, including work in which labels were incomplete, delayed or biased.

  • Experience deploying models into production and maintaining them: scheduling execution, versioning and drift monitoring, with data engineers.

  • Python and SQL, including feature engineering within the data warehouse at very large scale rather than extracting to a local environment.

  • 2+ years working with healthcare claims data (Medicare, Medicaid or commercial) and coding systems (e.g., ICD-10, CPT, HCPCS, DRG).

  • Experience with validation design for imbalanced, temporally shifting problems: out-of-time evaluation, leakage detection, calibration and precision-focused metrics over ranked output.

  • Ability to explain model output to a non-technical investigator, defend methodology to technical audiences, and present analytic outcomes to clients and stakeholders.

DESIRED QUALIFICATIONS:

  • Graph or network analytics, entity resolution and record linkage for identifying collusive relationships across payers.

  • Positive-unlabeled, semi-supervised or active learning against a capacity-constrained review queue.

  • Modeling in a regulated or adverse-action setting where explainability and fairness were requirements.

  • Anomaly detection, peer-group construction and case-mix methods (e.g., HCC); AWS and/or Snowflake, including Snowpark or model lifecycle tooling.

  • Healthcare FWA or program integrity datamining in multi-payer databases; payer coverage policy (LCDs, NCDs) and claim edits (e.g., NCCI).


SECURITY CLEARANCE LEVEL:

  • Must be able to obtain/maintain Public Trust.

GDIT IS YOUR PLACE:

The GDIT HFPP TTP is the only data warehouse of its type anywhere, bringing many billions of claims from dozens of public and private payers together solely for fraud, waste and abuse analytics. The cross-payer visibility it provides exists nowhere else.

A combination uncommon in machine learning roles: mature data and established subject matter expertise in place, with the modeling and production practices yours to define.

OWN YOUR OPPORTUNITY:

Explore a career in data science and engineering at GDIT and you'll find endless opportunities to grow alongside colleagues who share your determination for solving complex data challenges.

The likely salary range for this position is $123,250 - $166,750. This is not, however, a guarantee of compensation or salary. Rather, salary will be set based on experience, geographic location and possibly contractual requirements and could fall outside of this range.

Scheduled Weekly Hours:

40

Travel Required:

Less than 10%

Telecommuting Options:

Remote

Work Location:

Any Location / Remote

Additional Work Locations:

Total Rewards at GDIT:

Our benefits package for all US-based employees includes a variety of medical plan options, some with Health Savings Accounts, dental plan options, a vision plan, and a 401(k) plan offering the ability to contribute both pre and post-tax dollars up to the IRS annual limits and receive a company match. To encourage work/life balance, GDIT offers employees full flex work weeks where possible and a variety of paid time off plans, including vacation, sick and personal time, holidays, paid parental, military, bereavement and jury duty leave. GDIT typically provides new employees with 15 days of paid leave per calendar year to be used for vacations, personal business, and illness and an additional 10 paid holidays per year. Paid leave and paid holidays are prorated based on the employee’s date of hire. The GDIT Paid Family Leave program provides a total of up to 160 hours of paid leave in a rolling 12 month period for eligible employees. To ensure our employees are able to protect their income, other offerings such as short and long-term disability benefits, life, accidental death and dismemberment, personal accident, critical illness and business travel and accident insurance are provided or available. We regularly review our Total Rewards package to ensure our offerings are competitive and reflect what our employees have told us they value most.

 



Our Identity Verification Process:

As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during virtual interviews. We reserve the right to take your picture to verify your identity and prevent fraud. By proceeding, you authorize the collection, processing, and use of your biometric data for identity verification and security purposes.

About Our Work:

We are GDIT. A global technology and professional services company that delivers technology solutions and mission services to every major agency across the U.S. government, defense and intelligence community. Our 26,000 experts extract the power of technology to create immediate value and deliver solutions at the edge of innovation. We operate across 50+ countries worldwide, offering leading mission-ready capabilities in AI, cloud, cyber and software development.

Join our Talent Community to stay up to date on our career opportunities and events at

gdit.com/tc.

Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans

Skills Required

  • Master's in a quantitative field or Bachelor's with equivalent hands-on experience
  • 5+ years building, validating and delivering supervised machine learning models on real-world data
  • 2+ years working with healthcare claims data (Medicare, Medicaid or commercial) and coding systems (ICD-10, CPT, HCPCS, DRG)
  • Experience deploying models into production and maintaining them: scheduling execution, versioning and drift monitoring
  • Python and SQL, including feature engineering within the data warehouse at very large scale
  • Experience designing validation for imbalanced, temporally shifting problems: out-of-time evaluation, leakage detection, calibration and precision-focused metrics
  • Ability to explain model output to non-technical investigators and present methodology to stakeholders
  • Must be able to obtain/maintain Public Trust clearance
  • Work visa sponsorship will not be provided; candidates must reside in the United States
  • Graph or network analytics, entity resolution and record linkage for collusive relationship detection
  • Positive-unlabeled, semi-supervised or active learning methods for capacity-constrained review queues
  • Experience with AWS and/or Snowflake, Snowpark or model lifecycle tooling
  • Modeling in regulated or adverse-action settings where explainability and fairness were requirements
  • Anomaly detection, peer-group construction and case-mix methods (e.g., HCC)

General Dynamics Information Technology Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about General Dynamics Information Technology and has not been reviewed or approved by General Dynamics Information Technology.

  • Affordable Benefits Pay and benefits are described as good or okay in multiple places, and the overall package is often portrayed as acceptable even when base pay is not viewed as top-tier.
  • Healthcare Strength Medical, dental, and vision plan options are presented as comprehensive, with additional protections like disability and life insurance contributing to a well-rounded health and protection offering.
  • Retirement Support A 401(k) plan with company match is consistently highlighted as part of the total rewards package, supporting longer-term financial planning.

General Dynamics Information Technology Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Falls Church, VA
21,625 Employees

What We Do

We are GDIT. The people supporting some of the most complex government, defense, and intelligence projects across the country. We deliver. Bringing the expertise needed to understand and advance critical missions. We transform. Shifting the ways clients invest in, integrate, and innovate technology solutions. We ensure today is safe and tomorrow is smarter. We are there. On the ground, beside our clients, in the lab, and everywhere in between. Offering the technology transformations, strategy, and mission services needed to get the job done.

Similar Jobs

CrowdStrike Logo CrowdStrike

Senior Data Scientist

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
USA
11000 Employees
140K-215K Annually

Photon Logo Photon

Senior Data Scientist

Agency • Information Technology
Remote
United States
5017 Employees
60K-210K Annually

Johnson & Johnson Logo Johnson & Johnson

Senior Data Scientist

Healthtech • Biotech • Pharmaceutical • Manufacturing
In-Office or Remote
7 Locations
143612 Employees

Extend Logo Extend

Data Scientist

eCommerce • Insurance • Software
Remote
US
181 Employees
135K-165K Annually

Similar Companies Hiring

Outpost Space Thumbnail
Aerospace • Defense
US
24 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account