Senior Data Scientist

Posted Yesterday
Hiring Remotely in USA
Remote
150K-165K Annually
Senior level
Fintech
The Role
Build and deploy predictive models that drive partner action, analyze conversational and unstructured data using NLP techniques, and monitor model performance in production. Own AI data tooling, evaluation frameworks, in-warehouse model configuration, and reusable analytics foundations. The role also establishes modeling standards, evaluates fairness and limitations, collaborates with engineering and product teams, and prioritizes high-impact work across a lean data organization.
Summary Generated by Built In
Position:  Senior Data ScientistPosition Type: Full Time (Primarily Remote)Salary: $150,000-$165,000 DOEJoin our team at Mainstay, a division of Lemnis! Lemnis is a public charity dedicated to harnessing transformative change to expand learning for all. We are excited to announce a new opportunity for a Senior Data Scientist to join our innovative, forward-thinking, and growing organization. At Lemnis, we are committed to fostering a collaborative and inclusive work environment where every team member can thrive. If you are passionate about expanding learning for all, eager to make a meaningful impact, and ready to take on new challenges, we would love to hear from you. Apply now and be a part of our journey!Position Summary At Lemnis, we believe that the world is changing in exciting ways. It's up to us to create more equitable, flexible, and learner-centered systems that empower young people to rise to the challenges and opportunities of the future. Do you have questions? At Mainstay, our users have millions, so it’s imperative that our systems are stable, robust, and scalable. As a Senior Data Scientist at Mainstay, you will help implement features that help our partners and end users.About Mainstay At Mainstay, we believe one conversation can spark a brighter future. Our Engagement Platform makes it easy for colleges and businesses to start and measure conversations that drive action at scale. From our rigorous research methods to our Behavioral Intelligence framework, everything we do is designed to help people take the next step toward achieving their goals.This is your chance to drive impact at Mainstay and for our users by building and enhancing the systems that power our product. We’ll provide the environment for you to master your skills and find both personal and professional growth. We are invested in promoting from within and providing the support and mentorship aimed at your long-term success.The engineering team is a group of talented full-stack developers, data engineers, analytics engineers, and product experts. Engineers collaborate closely with their counterparts in the product team and external stakeholders to build new functionality and enhance existing products to delight our partners and end users. The team is excited to continue tackling new challenges and leverage cutting-edge technologies to solve impactful problems.About the RoleWe're hiring a Senior Data Scientist to find the predictive signals in our data and turn them into insights our partners can act on.Your early work is less about squeezing out marginal accuracy and more about finding signals. You'll build student-level predictors that sync into partner systems, analyze conversational data at scale to make our AI smarter, improve access to reliable data and insights, and leverage internal tools to scale your expertise.As the senior data scientist on a lean team, you’ll have real say in what we build. We have more promising directions than capacity, so part of the job is deciding which signals are worth pursuing, which analyses will generalize, and which requests to decline. You’ll set the methodology bar for modeling and evaluation work here, and you’ll be the person others come to when they aren’t sure whether to trust a number.You'll work alongside our Senior Data Engineer and Senior Analytics Engineer. They own infrastructure, pipelines, and the shared semantic layer. You own predictive modeling, unstructured data analysis, AI evaluation, and the analytics surfaces that make data accessible and trustworthy. If you're looking to train models full-time, this isn't that role.What You'll Do
Find and ship predictive signals
  • Investigate signals across student engagement, outcomes, and partner health, prioritizing by business impact over technical interest.
  • Build predictive models that reach partners through the systems they already use, where output drives real action.
  • Set thresholds against the alert volume teams can actually act on, and be explicit about the cost of a false positive when predictions reach a partner.
  • Evaluate performance across student populations, not just in aggregate, and document known limitations alongside the model.
  • Monitor deployed models for drift, and retire models that stop earning their place.
  • Know when not to build a model. Some questions are better answered with an analysis, a definition change, or a conversation.
  • Own scheduling and monitoring for your models, using orchestration the whole team can maintain.
 
Analyze our conversational and unstructured data
  • Apply embeddings, clustering, and classification to conversational data, support and service records, and other unstructured sources to surface themes, gaps, and emerging concerns.
  • Turn what you find into changes that improve the product and the partner experience.
  • Build the text analysis foundations that make our unstructured data retrievable and useful to AI tooling.
 
Own our AI data tooling and evaluation
  • Own our in-warehouse AI configuration, including verified queries, prompts, and agent tooling, and the semantic views that expose your model output. You'll partner with an Analytics Engineer where this depends on the shared semantic layer.
  • Build and maintain the AI evaluation framework for data team work: rubrics precise enough that two reviewers agree, a defensible sampling approach, and reporting that shows whether a model or prompt change actually improved anything.
  • Set the evaluation standard for data team work, and partner with Product and Engineering to share evaluation methodology more broadly.
 
Make your work reusable
  • Equip internal teams with the data and analysis they need for our most strategic partners, favoring work that generalizes over one-off requests.
  • Grow into building the tooling and training that lets those teams answer questions without the data team.
  • Document reasoning, assumptions, and tradeoffs in our internal data knowledge base as part of finishing the work.
  • Work in dbt/code alongside our engineers, contributing models and requesting changes to shared definitions.
 What We're Looking For

Required experience and skills

  • 5+ years building predictive models that someone actually used, including a few years where you owned the problem rather than being handed it, with the judgment that goes with it: calibration, threshold-setting, and knowing when a feature is leaking the answer.

  • Practical NLP experience: text classification, clustering, embeddings, or similar applied work. Calling an LLM API is useful but isn't the same thing.

  • Strong SQL as a primary tool, not a way to get data into a notebook.

  • Working Python for modeling and analysis.

  • Solid applied statistics, with the judgment to know which method fits the question and when the data can't support a conclusion.

  • Modern cloud warehouse experience (Snowflake, BigQuery, Databricks, or similar), especially with in-warehouse AI or agent tooling.

  • Excellent written and verbal communication. You can hand a finding to a non-technical colleague and have them act on it, including knowing what would change your conclusion.

  • Care about how predictions get used. Our scores influence how students get supported, so we want someone who checks whether a model works as well for part-time students as for everyone else, and says so when it doesn't.

  • Comfort with ambiguity and honesty about uncertainty. We'd rather hear "the data can't answer this" than a confident answer that falls apart later.

  • Comfortable working in version control with code review, so your analysis and models are reproducible by someone else.

  • Experience productionizing model output into an operational workflow.

  • Experience scheduling and monitoring recurring jobs in production, and the judgement to reach for tooling the team can maintain rather than a specialized stack that only you know.

  • Track record of choosing what to work on. You’ve turned an ambiguous business goal into a scoped project, made the prioritization case, and been accountable for whether it mattered.


Nice to have

  • Transformation tooling (dbt or similar), dimensional modeling, or analytics engineering exposure.

  • Familiarity with AI evals or prompt evaluation.

  • A modern BI tool (Sigma, Looker, Hex, Tableau, or similar) for making findings usable by others.

  • Experience working closely with analytics or data engineers, where your models depended on someone else's tables.

  • Linguistics or computational linguistics background for intent classification and evaluation work.

  • EdTech, higher education, or student success background.


This probably isn’t the right opportunity for you if

  • You need a dedicated ML platform to be effective. We deliberately don't run one: models run inside our cloud warehouse, features come from our transformation layer, predictions land in tables that downstream systems read.

  • Your experience is primarily research or model development without anyone using the output.

  • You want to specialize. The role ranges across modeling, text analysis, evaluation, and enablement, and none of them will be someone else’s job.

  • You’d rather have your work reviewed rather than review others. As the senior person in this discipline here, you’ll be setting the standard, not inheriting one.

  • You’d introduce a new tool for every problem. We optimize for delivering business value using long-term maintainable solutions, which sometimes means the second-best tool.


Key Performance Metrics

  • Validated predictive signals shipped and in active use by partners or internal teams.

  • Model quality in production terms: calibration and precision at an actionable alert volume.

  • Adoption of AI data tooling, including share of questions answered without data team involvement.

  • Evaluation coverage: share of data team AI surfaces with an active eval, and whether model or prompt changes ship with evidence.

  • Stakeholder feedback from Partner Success, Product, and Leadership.

  • Quality of prioritization: whether the work you chose turned out to matter, and whether you surfaced tradeoffs early.

Skills Required

  • 5+ years building predictive models that are used in practice, including ownership of modeling problems
  • Practical NLP experience, including text classification, clustering, embeddings, or similar applied work
  • Strong SQL skills
  • Working Python experience for modeling and analysis
  • Applied statistics expertise and sound methodological judgment
  • Experience with a modern cloud warehouse such as Snowflake, BigQuery, or Databricks
  • Excellent written and verbal communication skills
  • Experience evaluating model performance across populations and communicating limitations
  • Comfort working with ambiguity and uncertainty
  • Experience with version control and code review
  • Experience productionizing model output into operational workflows
  • Experience scheduling and monitoring recurring production jobs
  • Track record of prioritizing ambiguous business goals and owning outcomes
  • Experience with dbt or similar transformation tooling, dimensional modeling, or analytics engineering
  • Familiarity with AI evaluations or prompt evaluation
  • Experience with Sigma, Looker, Hex, Tableau, or similar BI tools
  • Experience collaborating closely with analytics or data engineers
  • Linguistics or computational linguistics background
  • EdTech, higher education, or student success experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco
56 Employees
Year Founded: 2008

What We Do

Vizury is a commerce marketing platform and its personalized retargeting stack is used by digital companies to grow marketing ROI and enhance transactions. Vizury’s retargeting platform is unique as it offers an integrated proposition to target and engage with the interested consumers over Programmatic, Social and Notification channels. This platform was launched in 2007 and after achieving global scale and success, the platform and business of Vizury was acquired by Affle in 2018. After the acquisition of the Vizury platform, it has now become an integral product as part of Affle’s Consumer Platform. Affle started in 2005 and is a global technology company with a proprietary consumer intelligence platform that delivers consumer engagement, acquisitions and transactions through relevant Mobile Advertising.

Similar Jobs

Pfizer Logo Pfizer

Senior Data Scientist

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
32 Locations
121990 Employees
163K-272K Annually

PatientPoint Logo PatientPoint

Senior Data Scientist

AdTech • Digital Media • Healthtech • Marketing Tech • Sales • Analytics • Pharmaceutical
Easy Apply
In-Office or Remote
Cincinnati, OH, USA
660 Employees
143K-211K Annually

Square Logo Square

Senior Data Scientist

eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Remote or Hybrid
Seattle, WA, USA
12000 Employees
168K-297K Annually

Square Logo Square

Senior Data Scientist

eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
12000 Employees
168K-297K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account