Data Operations Scientist, Development (52350)

Posted An Hour Ago
Be an Early Applicant
Hiring Remotely in USA
Remote
145K-175K Annually
Mid level
Professional Services • Consulting • Financial Services • Cybersecurity
The Role
Design, train, validate, and deploy traditional ML models and feature pipelines within Microsoft Fabric/Databricks. Partner with Data Engineers to optimize Medallion layers for model training and inference, implement MLOps (model registry, monitoring, retraining), perform EDA on large enterprise datasets, translate analytics to business insights, and document algorithm governance and fairness.
Summary Generated by Built In

Citrin Cooperman offers a dynamic work environment, fostering professional growth and collaboration. We’re continuously seeking talented individuals who bring a problem-solving mindset, fresh perspectives, and sharp technical expertise. We know you have choices, so our team of collaborative, innovative professionals are ready to support your professional development. At Citrin Cooperman, we offer competitive compensation and benefits and most importantly, the flexibility to manage your personal and professional life to focus on what matters most to you!

We are seeking a Data Operations Scientist, Development, to join our Development team within the Information Technology department. While our parallel AI Solutions team focuses on Generative AI and Agentic pilots, we’re seeking a dedicated Data Operations Scientist to own our core predictive analytics, statistical modeling, and traditional Machine Learning (ML) capabilities.

In this role, you’ll be the analytical powerhouse of our “Base Plan.” You’ll work directly with the Database Administrator and Data Engineers to ensure our Medallion architecture (bronze, silver, gold layers) is optimized not just for BI reporting, but for feature engineering and model training at scale. Utilizing Microsoft Fabric’s Synapse and Databricks, you’ll design, train, and deploy robust ML models that solve tangible business problems, including but not limited to customer churn prediction, demand forecasting, and operational optimization. The ideal candidate is a pragmatic statistician and coder who values MLOps discipline, model interpretability, and stable production deployments over experimental hype.

Responsibilities are, but not limited to:

  • Predictive Modeling & Advanced Analytics: Design, train, and validate traditional machine learning models (e.g., regression, classification, clustering, time-series forecasting) using Python, PySpark, and established libraries (Scikit-Learn, XGBoost, LightGBM).
  • Feature Engineering & Data Shaping: Partner closely with Data Engineers to design the “Gold” data layer. Create and manage robust feature pipelines, ensuring data is properly structured, normalized, and optimized for both training and low-latency inference.
  • MLOps & Model Lifecycle Management: Deploy models into production within the Microsoft Fabric ecosystem. Establish the MLOps pipelines required to track model versions (e.g., using MLflow), monitor for concept/data drift, and trigger automated retraining when performance degrades.
  • Exploratory Data Analysis (EDA): Conduct deep-dive statistical analyses on large, complex enterprise datasets (housed in OneLake/SQL) to uncover hidden patterns, validate business hypotheses, and inform strategic decision-making.
  • Collaboration & Translation: Act as the bridge between raw data and business strategy. Translate complex statistical outcomes into clear, actionable insights for non-technical stakeholders, often partnering with BI developers to integrate model outputs into Power BI dashboards.
  • Algorithm Governance: Document model methodologies, assumptions, and limitations to ensure compliance with enterprise data governance and algorithmic fairness standards.
Qualifications

The ideal candidate must:

  • Have a bachelor’s degree in computer science, data engineering, mathematics, or equivalent practical experience.
  • Have 3–5 years of professional experience as a Data Scientist, Machine Learning Engineer, or Advanced Analyst in a corporate environment.
  • Have deep proficiency in Python and SQL, with strong hands-on experience using industry-standard data science and ML libraries (Pandas, NumPy, Scikit-Learn, PyTorch/TensorFlow).
  • Have proven experience with Big Data processing frameworks (Apache Spark, PySpark) and modern cloud data platforms (Microsoft Fabric, Databricks, or Azure Machine Learning heavily preferred).
  • Possess a solid foundation in statistics, probability, and mathematics, with the ability to mathematically justify model selection and evaluation metrics (RMSE, F1-score, AUC-ROC).
  • Have experience implementing MLOps best practices, including model registry management, containerized deployments, and performance monitoring.
  • Possess strong business acumen and the ability to connect statistical improvements directly to business ROI.
  • Be pragmatic problem solver: Chooses the simplest, most explainable model (like a well-tuned random forest) that solves the business problem, rather than over-engineering a complex neural network just for the sake of it.
  • Be rigorous & methodical: Deeply respects data quality and understands that a model is only as good as the pipelines feeding it. Naturally skeptical of “perfect” training results.
  • Be a cross-functional collaborator: Thrives in a team setting. Eager to sit down with a Data Engineer to optimize a Spark query or with a TPM to scope a sprint, rather than working in an isolated research silo.
  • Be Microsoft certified: Azure Data Science Associate (DP-100) (preferred).
  • Be Microsoft certified: Fabric Analytics Engineer Associate (DP-600) (preferred).
  • Be Databricks certified: Machine Learning Associate (PL-300) (preferred).

Skills Required

  • Bachelor's degree in computer science, data engineering, mathematics, or equivalent practical experience.
  • 3-5 years professional experience as a Data Scientist, Machine Learning Engineer, or Advanced Analyst in a corporate environment.
  • Deep proficiency in Python.
  • Strong hands-on experience with SQL.
  • Experience with data science and ML libraries: Pandas, NumPy, Scikit-Learn, XGBoost, LightGBM.
  • Experience with deep learning frameworks (PyTorch or TensorFlow).
  • Proven experience with Big Data processing frameworks (Apache Spark, PySpark).
  • Experience with modern cloud data platforms (Microsoft Fabric, Databricks, or Azure Machine Learning).
  • Experience implementing MLOps best practices, including model registry management, containerized deployments, and performance monitoring (e.g., MLflow).
  • Solid foundation in statistics, probability, and mathematics; ability to justify model selection and evaluation metrics (RMSE, F1, AUC-ROC).
  • Ability to create and manage feature pipelines and optimize data for training and low-latency inference.
  • Experience conducting exploratory data analysis on large enterprise datasets (OneLake/SQL) and translating results for business stakeholders.
  • Strong business acumen and ability to link statistical improvements to business ROI.
  • Demonstrated pragmatic, rigorous, and collaborative working style.
  • Azure Data Science Associate (DP-100) certification.
  • Fabric Analytics Engineer Associate (DP-600) certification.
  • Databricks Machine Learning Associate certification.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
3,502 Employees
Year Founded: 1979

What We Do

Citrin Cooperman is one of the largest professional services firms in the United States, founded in 1979. It provides comprehensive tax, accounting, and advisory services to middle-market companies and high-net-worth individuals. The firm delivers a tailored, integrated business approach and proactive guidance to help clients achieve their business and personal financial goals across various global markets.

Similar Jobs

The Tara Group Logo The Tara Group

Account Executive

AdTech • Digital Media • Marketing Tech • Analytics
Remote
United States
50 Employees
80K-95K Annually

Samsara Logo Samsara

Firmware Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
6 Locations
4000 Employees
162K-290K Annually

Liberty Mutual Insurance Logo Liberty Mutual Insurance

Inside Sales Representative

Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Remote or Hybrid
10 Locations
40000 Employees
45K-85K Annually

Zscaler Logo Zscaler

Account Executive

Cloud • Information Technology • Security • Software • Cybersecurity
Easy Apply
Remote or Hybrid
Location, WV, USA
8697 Employees

Similar Companies Hiring

NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account