Data Scientist

Posted 5 Days Ago
Be an Early Applicant
11 Locations
Remote
Junior
Information Technology • Software
The Role
Build and productionize ML and generative models (including LLM-powered pipelines and agents), process large-scale messy datasets, develop reusable Python data science libraries, design evals and observability, debug distributed systems, and collaborate with cross-functional stakeholders to deliver production-grade data products.
Summary Generated by Built In

About the DevSavant

DevSavant is an operating partner for startups and growth-stage companies, helping them turn ambition into execution.

We support founders and leadership teams with product engineering and global staffing, from early prototypes and MVPs to scaling high-performing teams. Our vetted talent across LATAM and Asia embeds directly into client teams, operating as true extensions rather than external vendors.

With over 8 years working in venture-backed ecosystems, DevSavant is trusted to accelerate delivery, scale teams efficiently, and support companies as they reach their next milestone.

About the Role

We're looking for a Data Scientist to build the models and data products that make it all the way to production — from generative models on mixed data sources, to subscriber-behavior predictions, to new models for TV providers (MVPDs). You'll dig into large, messy datasets to find the trends and patterns that turn into shipped features, and you'll build LLM-powered pipelines and agents with the evals to prove they work. You'll work closely with data scientists and engineers to take ideas from first experiment to production at market scale, and collaborate directly with cross-functional stakeholders — including our co-founders.

This is a mid-level, remote role reporting into the Data Science team, requiring advanced (C1) English proficiency for clear, direct communication on complex technical and system design decisions.

Key Responsibilities

  • Build models and data products that make it all the way to production — from generative models on mixed data sources to subscriber-behavior predictions and new models for TV providers (MVPDs).

  • Dig into large, messy datasets to find the trends and patterns that turn into shipped features, and add the functions, classes, and tools to our core Python data science library that the rest of the team builds on.

  • Take on Antenna R&D work: explore new datasets and methods to answer real business questions and present what you find to senior stakeholders.

  • Write clear, well-organized, testable, and efficient code using object-oriented principles, grounded in deep knowledge of core Python and data tools. Because you care about quality, your code is well-documented.

  • Debug complex distributed systems and make code faster and able to handle more data.

  • Build LLM-powered pipelines and agents, and explain the failure modes you hit and the guardrails you added.

  • Treat evals as a core deliverable: validate model responses with provider-enforced structured outputs, build eval sets with clear pass/fail checks, calibrate LLM-as-a-judge rubrics, and use tracing tools to track cost, latency, and quality over time. You can point to an eval that caught a problem human review missed.

  • Use agentic coding tools as part of your daily workflow: plan first, write tests and instructions up front, and review every change before accepting it — while still designing, debugging, and defending your work without AI assistance.

  • Collaborate with cross-functional stakeholders, including our co-founders, and clearly explain complex technical and system design decisions.

Required Qualifications

  • 2+ years of experience building machine learning models and data products in Python, with the engineering skills to take them from prototype to production.

  • Expert in Python with strong object-oriented design, software system design, and experience building high-quality, testable, production-grade code.

  • Hands-on experience with deep learning frameworks (PyTorch or TensorFlow), plus a deep understanding of machine learning concepts, the end-to-end model development lifecycle, and MLOps principles.

  • Hands-on experience with large-scale data processing tools (e.g., Apache Spark/PySpark, Dask) and strong SQL skills working with large, complex datasets.

  • Solid experience with cloud platforms (GCP highly preferred), including deploying, managing, and scaling services (Docker, Cloud Run, GKE) and working with big data systems (Dataproc, BigQuery).

  • Excellent problem-solver, skilled at debugging complex distributed systems and optimizing them for performance and scale.

  • Advanced English proficiency (B2–C1) with strong communication, teamwork, and consulting skills; able to clearly explain complex technical and system design decisions.

  • Daily use of agentic coding tools (Claude Code, Cursor, or Codex CLI): plan first, write tests and instructions up front, and review every change before accepting it. You can still design, debug, and defend your own work without AI assistance — our interview process tests this directly.

  • Experience building and shipping LLM-powered agents or pipelines using an orchestration framework (LangGraph, Pydantic AI, or OpenAI Agents SDK), including custom tool definitions against internal APIs and data systems, agent state and memory, and human review steps. You can explain the failure modes you hit and the guardrails you added.

  • You treat evals as a core deliverable: you validate model responses with provider-enforced structured outputs (Pydantic), build eval sets with clear pass/fail checks, calibrate LLM-as-a-judge rubrics, and use tracing tools (Langfuse, LangSmith, or Braintrust) to track cost, latency, and quality over time. You can describe an eval that caught a problem a human review missed.

Bonus

  • Experience in or passion for the Subscription Economy, especially media and entertainment; experience working with media data or data clean rooms is a plus.

  • Experience with synthetic data generation or advanced generative models (e.g., GANs, VAEs, CTGAN).

  • Experience building Python libraries that others use, or contributions to open-source projects.

  • Knowledge of advanced MLOps practices (like model monitoring) and build automation tools (e.g., Cloud Build, Cloud Run).

  • Experience building custom tool integrations for agents: your own tool definitions and routing against internal APIs and data systems, with clear input schemas, validation, and safe handling of side-effecting actions.

  • Experience with advanced evaluation and observability practices like multi-judge calibration, automated regression suites, and production monitoring of agent quality.

  • Familiarity with RAG and context engineering for grounding model responses in proprietary data.

  • Experience using LLMs for testing pipelines and QA workflows.

Skills Required

  • 2+ years building machine learning models and data products in Python
  • Expert Python skills with strong object-oriented design and production-grade, testable code experience
  • Hands-on experience with deep learning frameworks (PyTorch or TensorFlow)
  • Deep understanding of machine learning concepts, end-to-end model lifecycle, and MLOps principles
  • Hands-on experience with large-scale data processing tools (Apache Spark/PySpark, Dask)
  • Strong SQL skills working with large, complex datasets
  • Experience with cloud platforms (GCP highly preferred) and deploying/managing/scaling services (Docker, Cloud Run, GKE)
  • Experience with big data systems (Dataproc, BigQuery)
  • Skilled at debugging complex distributed systems and optimizing for performance and scale
  • Advanced English proficiency (B2-C1) with strong communication and consulting skills
  • Daily use of agentic coding tools (Claude Code, Cursor, or Codex CLI) and ability to work without AI assistance when needed
  • Experience building and shipping LLM-powered agents or pipelines using orchestration frameworks (LangGraph, Pydantic AI, or OpenAI Agents SDK), including custom tool definitions and human review workflows
  • Experience designing and running evals: provider-enforced structured outputs (Pydantic), eval sets with pass/fail checks, LLM-as-judge calibration, and tracing tools (Langfuse, LangSmith, Braintrust)
  • Experience with synthetic data generation or advanced generative models (GANs, VAEs, CTGAN)
  • Experience in or passion for subscription economy/media, media data or data clean rooms
  • Experience building Python libraries or open-source contributions
  • Knowledge of advanced MLOps practices (model monitoring) and build automation tools (Cloud Build, Cloud Run)
  • Experience with RAG, context engineering, and using LLMs for testing/QA workflows
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Mateo, California
93 Employees
Year Founded: 2020

What We Do

DevSavant provides comprehensive technology solutions to Savant Growth's portfolio companies. Our data scientists and developers are experts in the fields of data analytics, software development, and AI. DevSavant’s engineers work across the spectrum of full-stack technology solutions for the B2B SaaS industry, helping you growing company tackle its toughest challenges and giving you the freedom to focus on what matters most, the future of your business

Similar Jobs

ON.energy Logo ON.energy

Data Scientist

Artificial Intelligence • Energy • Renewable Energy
Remote
12 Locations
165 Employees

RevenueCat Logo RevenueCat

Senior Data Scientist

Fintech • Mobile • Payments • Software
Remote
41 Locations
40 Employees

RevenueCat Logo RevenueCat

Senior Data Scientist

Fintech • Mobile • Payments • Software
Remote
14 Locations
40 Employees

Factored Logo Factored

Senior Data Scientist

Artificial Intelligence • Machine Learning • Analytics
Remote
11 Locations
166 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account