AI/ML Data Scientist | Onsite

Posted 5 Hours Ago
Be an Early Applicant
2 Locations
In-Office or Remote
40K-140K Annually
Mid level
Agency • Information Technology
The Role
Design, build, evaluate, and deploy predictive ML models across regression, classification, NLP, computer vision, and time-series problems. Perform data cleaning, feature engineering, EDA, model selection, tuning, explanation (SHAP, feature importance), and collaborate with engineers and stakeholders to operationalize and monitor models in production.
Summary Generated by Built In

Job Description: ML/AI Data Scientist – Predictive Modeling 

Position Overview 

We are seeking an experienced ML/AI Data Scientist to design, build, train, evaluate, and deploy predictive models that solve complex business and operational problems. The ideal candidate combines strong statistical foundations with hands-on experience in data analysis, feature engineering, model development, and performance evaluation. 

This role is especially suited to someone who can build predictive solutions from the ground up and clearly explain the “what,” “why,” and “how” behind their analytical and modeling decisions to both technical and non-technical stakeholders. 

Key Responsibilities 

  • Translate business problems into well-defined machine learning and predictive modeling objectives. 
  • Collect, clean, transform, and analyze structured and unstructured data from multiple sources. 
  • Perform exploratory data analysis to identify trends, relationships, anomalies, biases, and data-quality issues. 
  • Develop predictive models from scratch, including data preparation, feature engineering, training, validation, testing, and optimization. 
  • Build and apply regression models for forecasting, estimation, risk scoring, pricing, demand prediction, and related use cases. 
  • Build and apply classification models for segmentation, fraud detection, churn prediction, recommendation, anomaly detection, and other decision-support applications. 
  • Select appropriate algorithms based on the problem type, data characteristics, business requirements, interpretability needs, and operational constraints. 
  • Compare baseline, linear, tree-based, ensemble, and other appropriate modeling approaches. 
  • Tune model hyperparameters and use appropriate cross-validation strategies to improve generalization. 
  • Experience building and deploying AI solutions using Natural Language Processing (NLP), Computer Vision, and sequence modeling techniques for text, image, video, and time-series data. 
  • Strong knowledge of deep learning architectures including RNNs, LSTMs, GRUs, CNNs, and Transformer-based models, with hands-on experience using TensorFlow or PyTorch. 
  • Ability to evaluate, optimize, and explain AI model performance, including model accuracy, robustness, bias detection, feature interpretation, and production monitoring. 
  • Evaluate model performance using relevant metrics such as RMSE, MAE, R², accuracy, precision, recall, F1 score, ROC-AUC, PR-AUC, log loss, calibration, and lift. 
  • Analyze model errors and identify opportunities for improving data quality, features, sampling strategies, and model assumptions. 
  • Assess model robustness, explainability, fairness, stability, and sensitivity to changing data patterns. 
  • Clearly communicate the rationale behind model selection, including why a particular model was chosen over alternatives. 
  • Explain technical results, assumptions, limitations, and trade-offs to product managers, business leaders, and other stakeholders. 
  • Document analytical methods, data sources, assumptions, experiments, model decisions, and results. 
  • Collaborate with data engineers, software engineers, product teams, domain experts, and business stakeholders to operationalize models. 
  • Support model deployment, monitoring, retraining, and continuous improvement in production environments. 
  • Stay current with developments in machine learning, statistical modeling, AI techniques, and responsible AI practices. 

Required Qualifications 

  • Bachelor’s or master’s degree in Data Science, Statistics, Mathematics, Computer Science, Engineering, Economics, or a related quantitative discipline. 
  • 3+ years of professional experience in data science, machine learning, predictive analytics, or a closely related field. 
  • Strong understanding of statistical analysis, probability, experimental design, and machine learning fundamentals. 
  • Demonstrated experience building predictive models from raw data through final evaluation. 
  • Deep practical expertise in regression and classification algorithms, including: 

          - Linear and polynomial regression 

          - Logistic regression 

          - Regularization methods such as Ridge, Lasso, and Elastic Net 

          - Decision trees 

          - Random forests 

          - Gradient boosting methods 

          - Support vector machines 

          - k-nearest neighbors 

          - Naive Bayes 

          - Ensemble modeling techniques 

          - CNN 

          - Computervision 

          - RNN 

          - NLP 

  • Strong knowledge of supervised learning workflows, including data splitting, cross-validation, feature selection, feature engineering, model tuning, and evaluation. 
  • Proficiency in Python and common data science libraries such as pandas, NumPy, scikit-learn, SciPy, and matplotlib or Seaborn. 
  • Strong SQL skills and experience querying, joining, aggregating, and analyzing data from relational databases. 
  • Experience working with missing data, outliers, imbalanced classes, categorical variables, high-cardinality features, and data leakage risks. 
  • Ability to select and justify appropriate evaluation metrics based on business objectives and model use cases. 
  • Experience explaining model behavior using techniques such as feature importance, partial dependence, SHAP, coefficients, permutation importance, or related methods. 
  • Excellent written and verbal communication skills. 
  • Ability to present complex analytical concepts clearly to audiences with varying levels of technical expertise. 

Preferred Qualifications 

  • Experience deploying machine learning models through APIs, batch pipelines, or cloud-based platforms. 
  • Familiarity with MLflow, Kubeflow, Airflow, Docker, Git, CI/CD, or similar tools. 
  • Experience with cloud platforms such as AWS, Microsoft Azure, or Google Cloud. 
  • Knowledge of time-series forecasting, survival analysis, recommender systems, or anomaly detection. 
  • Experience with deep learning frameworks such as PyTorch or TensorFlow. 
  • Familiarity with model monitoring, data drift, concept drift, model retraining, and performance degradation. 
  • Experience working with distributed data-processing tools such as Spark. 
  • Knowledge of responsible AI, model governance, fairness, privacy, and regulatory requirements. 
  • Experience working in an Agile or cross-functional product development environment. 

Core Competencies 

Analytical Thinking 

Ability to break down ambiguous problems, identify relevant data, test assumptions, and develop rigorous analytical solutions. 

Model Selection and Justification 

Ability to explain why a specific model is appropriate based on accuracy, interpretability, scalability, latency, data volume, feature relationships, regulatory requirements, and business impact. 

Statistical and Technical Expertise 

Strong understanding of statistical concepts and practical machine learning methods, with the ability to distinguish correlation from causation and identify modeling limitations. 

Data Understanding 

Ability to assess data quality, determine whether variables are meaningful, identify bias and leakage, and understand how data-generating processes affect model results. 

Communication 

Ability to communicate model assumptions, results, trade-offs, uncertainty, and limitations in clear and accessible language. 

Business Orientation 

Ability to connect technical modeling outcomes to measurable business goals, operational decisions, customer outcomes, or financial impact. 

Collaboration 

Ability to work effectively with engineering, product, operations, and leadership teams throughout the model lifecycle. 

Expected Modeling Approach 

Successful candidates should be able to demonstrate a structured approach that includes: 

  1. Defining the business problem and prediction target. 
  1. Establishing a simple and interpretable baseline. 
  1. Understanding the data-generating process and identifying potential biases. 
  1. Performing exploratory data analysis. 
  1. Preparing the data and engineering meaningful features. 
  1. Selecting candidate models based on the problem and constraints. 
  1. Training and validating models using appropriate methodology. 
  1. Comparing models using business-relevant metrics. 
  1. Explaining model behavior and identifying limitations. 
  1. Selecting the final model based on accuracy, interpretability, reliability, and operational fit. 
  1. Documenting the decision-making process. 
  1. Monitoring and improving the model after deployment. 

Deliverables and Success Measures 

  • High-quality exploratory analyses that produce actionable insights. 
  • Reliable regression and classification models aligned with business objectives. 
  • Clearly documented modeling decisions and assumptions. 
  • Reproducible data preparation and model-training workflows. 
  • Measurable improvements in forecasting accuracy, decision quality, efficiency, revenue, risk reduction, or customer outcomes. 
  • Models that are appropriately interpretable, robust, maintainable, and production-ready. 
  • Clear communication of model performance, uncertainty, trade-offs, and limitations. 
  • Effective collaboration with stakeholders throughout the analytics and model development lifecycle. 

Compensation, Benefits and Duration

Minimum Compensation: USD 40,000
Maximum Compensation: USD 140,000
Compensation is based on actual experience and qualifications of the candidate. The above is a reasonable and a good faith estimate for the role.
Medical, vision, and dental benefits, 401k retirement plan, variable pay/incentives, paid time off, and paid holidays are available for full-time employees.
This position is available for independent contractors
No applications will be considered if received more than 120 days after the date of this post

Skills Required

  • Bachelor's or master's degree in Data Science, Statistics, Mathematics, Computer Science, Engineering, Economics, or related quantitative discipline.
  • 3+ years of professional experience in data science, machine learning, predictive analytics, or a closely related field.
  • Strong understanding of statistical analysis, probability, experimental design, and machine learning fundamentals.
  • Demonstrated experience building predictive models from raw data through final evaluation, including data preparation, feature engineering, training, validation, testing, and optimization.
  • Deep practical expertise with regression and classification algorithms (linear/polynomial regression, logistic regression, Ridge/Lasso/Elastic Net, decision trees, random forests, gradient boosting, SVM, k-NN, Naive Bayes, ensemble methods).
  • Experience with deep learning architectures and techniques (CNN, RNN, LSTM, GRU, Transformer-based models) and applied work in NLP and computer vision.
  • Proficiency in Python and libraries such as pandas, NumPy, scikit-learn, SciPy, and matplotlib or Seaborn.
  • Strong SQL skills for querying, joining, aggregating, and analyzing relational data.
  • Experience handling missing data, outliers, imbalanced classes, categorical and high-cardinality features, and avoiding data leakage.
  • Ability to select and justify evaluation metrics (RMSE, MAE, R², accuracy, precision, recall, F1, ROC-AUC, PR-AUC, log loss, calibration, lift).
  • Experience explaining model behavior using feature importance, partial dependence, SHAP, coefficients, permutation importance, or related methods.
  • Excellent written and verbal communication skills and ability to present complex analytical concepts to varied audiences.
  • Ability to connect modeling outcomes to business goals and collaborate with engineering, product, and domain teams to operationalize models.
  • Experience deploying, monitoring, or maintaining models in production (retraining, performance monitoring, addressing data drift).
  • Experience deploying models through APIs, batch pipelines, or cloud platforms and familiarity with ML orchestration/CI tools (MLflow, Kubeflow, Airflow, Docker, Git, CI/CD).
  • Experience with cloud platforms (AWS, Microsoft Azure, Google Cloud) and distributed data-processing tools such as Spark.
  • Knowledge of time-series forecasting, survival analysis, recommender systems, anomaly detection, and responsible AI/model governance practices.
  • Experience working in Agile or cross-functional product development environments.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: London
5,017 Employees
Year Founded: 2007

What We Do

Photon.com has emerged as one of the world’s largest and fastest-growing Digital Agencies. We work with 40% of the Fortune 100 on their Digital initiatives and are known for our ability to integrate Strategy Consulting, Creative Design, and Technology at scale. Please visit www.photon.com to learn more about us, how we work, and our customer case studies. Digital Transformation Starts Here.

Similar Jobs

Mondelēz International Logo Mondelēz International

Security Engineer

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
7 Locations
90000 Employees
122K-168K Annually

Square Logo Square

Strategic Programs Lead

eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
12000 Employees
164K-297K Annually

Square Logo Square

Regional Sales Manager

eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Remote or Hybrid
8 Locations
12000 Employees
112K-203K Annually

Wipfli Logo Wipfli

Senior Consultant

Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Remote or Hybrid
United States
3000 Employees
88K-118K Annually

Similar Companies Hiring

Scrunch  Thumbnail
Artificial Intelligence • Information Technology • Marketing Tech • Software • SEO
Salt Lake City, Utah
Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account