Data Manager

Posted 14 Days Ago
Be an Early Applicant
Somerville, MA
Hybrid
Senior level
Big Data • Machine Learning • Software • Analytics • Biotech
AI-Powered Tools to Engineer Biology
The Role
The Data Manager will own data strategy, management, and governance, ensuring data quality and integration across various scientific datasets whilst collaborating with multiple teams.
Summary Generated by Built In
About Us

At Matterworks we are building AI tools to extract insights from the ever-growing corpora of biological data and to unlock opportunities in therapeutic discovery, development, and manufacturing. We are building large-scale deep learning models of biological data to predict the phenotype and behavior of biological systems.

Position Overview

Matterworks is seeking a Data Manager (Bioinformatics / Cheminformatics) to build our data management practice, owning the strategy, processes, and day-to-day execution that turn complex, messy chemical and biological datasets into high-quality, well-governed training corpora and product-ready data assets.

You’ll be the connective tissue between applied science, AI, product, and our data platform/engineering team helping to answer what data is most valuable, how do we onboard it quickly, and how do we keep it consistently high quality over time.

This role will start as an individual contributor with end-to-end ownership, with a growth path to leading a function as we scale.

Key Responsibilities

  • Data Strategy & E2E Ownership: Partner with scientific, ML, and product stakeholders to define a data roadmap: which datasets move the needle, which should be refreshed, and what “good enough” looks like for each use case. Establish clear success metrics for onboarding speed, dataset quality, and downstream usability (e.g., fewer training/data failures, higher match rates, better coverage, higher-confidence labels).

  • Dataset Sourcing, Discovery, and Intake: Proactively scout and integrate public and client datasets, plus relevant literature and reference materials, to keep our corpora current and comprehensive. Design a repeatable dataset intake workflow including provenance, source tracking, and refresh cadence.

  • Data Curation, Quality and Governance: Define curation standards that make data consistent across sources and modalities, including compound identity management, biological/sample metadata standardization, and schema + conventions mappings. Build a scalable approach to integrating metabolomics now and expanding to additional omics without reinventing everything each time. Develop practical QC/QA frameworks that combine scientific judgment with repeatable checks.

  • Cross-Functional Collaboration: Work closely with leadership in engineering, AI, product, and scientific discovery to align initiatives with company-wide goals. Use experience to keep initiatives moving smoothly. Translate ambiguous questions into crisp data requirements, priorities, and execution plans. Build trust across disciplines by being both scientifically rigorous and pragmatically execution oriented.

About You

  • 6+ years of demonstrated experience owning scientific data work end-to-end (curation, standardization, QC, documentation, governance) in bioinformatics, cheminformatics, computational biology, scientific data engineering, or related roles.

  • Ability to navigate complex chemical and biological datasets, reconcile identifiers/metadata across sources, and make data consistently usable for end users.

  • Strong attention to detail with a keen ability to balance priorities and delivery incremental value while operating with minimal oversight.

  • Comfortable building structure from scratch: you can define processes, set standards, and iterate toward scalable practices in an early-stage environment.

  • Practical proficiency in Python and SQL for data investigation, transformation, QC, and automation.

  • Familiarity with modern data workflows (structured + semi-structured data, pipelines, reproducibility, documentation).

  • Experience with chemical structure representations and normalization (e.g., SMILES/InChI, canonicalization, salt/tautomer handling, stereochemistry considerations).

  • Demonstrated ability to communicate and collaborate with product, machine learning, applied science and engineers while reducing complex business questions into valuable, reliable technical solutions.

  • A passion for contributing to an early-stage startup where autonomy, eagerness to learn, and enthusiasm for solving novel scientific challenges prevail over rigid processes and egos.

Working at Matterworks

Given the cross-disciplinary and innovative nature of our work, effective collaboration and communication are critical to our progress. We operate in a flexible hybrid model that accommodates both fully remote team members and those who work full-time from our Somerville, MA office. While some positions may require regular in-person presence for hands-on work or local collaboration, many roles can be performed remotely with team members distributed across various locations.

Compensation and Benefits

Matterworks offers a competitive base salary, stock options, and benefits (health & dental, vision, long- and short-term disability, life insurance, 401k with company match). Employees enjoy a flexible work & unlimited time away policy, commuter benefits and parking, regular team meals and outings, and company support for continued education/coursework and conference participation.

Matterworks, Inc. is an equal opportunity employer. All candidates for employment at Matterworks are considered without regard to race, color, religion, national origin, age, sex, marital status, ancestry, physical or mental disability, veteran status, gender identity, sexual orientation, or any other category protected by law.

Top Skills

Python
SQL
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Somerville, MA
25 Employees
Year Founded: 2019

What We Do

Matterworks is developing advanced AI-powered software and proprietary chemistry reagents to enable real-time, quantitative metabolomics to accelerate the research, development, and manufacturing of biologic therapeutics.

Metabolomics measures all the small molecules a cell needs to function. In spite of its far-reaching potential impact across life science disciplines, adoption of metabolomics remains limited and lags behind other ‘omic techniques due to a number of inherent challenges. Matterworks is developing a novel alternative methodology that activates deep learning (DL) for analytical chemistry, to remove the constraints imposed by the requirement for human-interpretable data.

Matterworks is on a mission to democratize metabolomics and catalyze its widespread adoption and integration into the life sciences. Our view is very simple: solving metabolomics will advance all of the life sciences industries.

Why Work With Us

We are a passionate team of scientists and engineers committed to building better tools to interrogate biology and accelerate the rate of scientific progress. In pursuit of our ambitious goals, we rely on our core company values to guide our work: Mission First, Communication & Trust, Growth Through Challenges, and Scientific Excellence.

Gallery

Gallery

Similar Jobs

PwC Logo PwC

Data Scientist

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
14 Locations
370000 Employees
124K-280K Annually

Chewy Logo Chewy

Analytics Manager

eCommerce • Healthtech • Pet • Retail • Pharmaceutical
Hybrid
Boston, MA, USA
17800 Employees
130K-207K Annually

WHOOP Logo WHOOP

Senior Program Manager

Fitness • Hardware • Healthtech • Sports • Wearables
Easy Apply
Hybrid
Boston, MA, USA
500 Employees
150K-216K Annually

Capital One Logo Capital One

Manager, Ontology and Data Modeling

Fintech • Machine Learning • Payments • Software • Financial Services
Hybrid
3 Locations
55000 Employees
165K-205K Annually

Similar Companies Hiring

Scotch Thumbnail
Software • Retail • Payments • Fintech • eCommerce • Artificial Intelligence • Analytics
US
25 Employees
Milestone Systems Thumbnail
Software • Security • Other • Big Data Analytics • Artificial Intelligence • Analytics
Lake Oswego, OR
1500 Employees
Fairly Even Thumbnail
Software • Sales • Robotics • Other • Hospitality • Hardware
New York, NY

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account