InXiteOut- Data Science Lead (NLP & GenAI)

Reposted 18 Days Ago
Be an Early Applicant
Hiring Remotely in India
Remote
Senior level
Artificial Intelligence • HR Tech • Professional Services • Software
The Role
Lead design and delivery of NLP and Generative AI solutions, from problem definition through deployment. Build and fine-tune transformers and LLMs for summarization, QA, classification and chatbots, apply classical and deep learning methods, ensure scalable reproducible ML workflows in cloud environments, mentor junior data scientists, and collaborate with product and engineering to translate business problems into AI solutions.
Summary Generated by Built In

Data Science Lead (NLP & GenAI)

Summary

We are seeking a highly experienced and innovative Data Science Lead with 8+ years of expertise in core data science concepts and around 2+ years of focused, hands-on experience in Natural Language Processing (NLP) and Generative AI (GenAI). You will lead strategic AI/ML initiatives, mentor junior data scientists, and deliver intelligent solutions that drive business value using both classical and modern machine learning techniques.

Key Responsibilities

Lead end-to-end design and delivery of data science solutions, from problem definition to deployment.

Design, build, and fine-tune NLP and GenAI models for tasks such as summarization, classification, question answering, translation, and chatbot applications.

Apply statistical modeling, predictive analytics, and machine learning algorithms on structured and unstructured datasets.

Collaborate with product, engineering, and business teams to translate high-level business problems into data science solutions.

Ensure scalability, reproducibility, and performance optimization in all machine learning workflows.

Work with large-scale data processing tools and frameworks in cloud-based environments.

Mentor and review work of junior data scientists and collaborate on research and experimentation.

Track advancements in GenAI, LLMs, and NLP frameworks and bring innovation to enterprise AI use cases.

Mandatory Skills

Python: Strong proficiency in Python for data science, modeling, and scripting

Machine Learning: Hands-on with classical and ensemble models (e.g., Random Forest, XGBoost)

NLP (2+ years): Experience with transformers, tokenization, embeddings, sentiment analysis

GenAI & LLMs: Working with GPT-like models, fine-tuning, prompt engineering

Deep Learning (PyTorch / TensorFlow): Building and training deep learning models for NLP and other domains

Model Deployment: Deploying models via REST APIs, Docker, or cloud-native services

SQL & Data Manipulation: Strong ability to query, clean, and process data

Statistical Analysis: Applied statistics, hypothesis testing, and A/B testing

Version Control (Git): Experience using Git in collaborative environments

Optional/nice-to-have skills

Vector Databases: Experience with FAISS, Pinecone, or ChromaDB for semantic search

RAG Architecture: Building Retrieval-Augmented Generation pipelines

LLM Orchestration: LangChain, LlamaIndex, or similar frameworks

Cloud Platforms (Azure/GCP/AWS): Cloud-based ML workflows, pipelines, and infrastructure

MLOps: Model tracking, monitoring, CI/CD with MLflow, Kubeflow, etc.

Big Data Tools: Spark, Databricks, or Hadoop ecosystem familiarity

Experiment Tracking: Tools like Weights & Biases, MLflow

Academic Research / Publications: Experience publishing whitepapers or research contributions

Hand-on experience with Databricks, preferably Azure Databricks platform.

Hand-on experience with Delta Lake, preferably Azure Databricks and ADLS Gen2 platforms.

Educational Qualifications

Master’s or PhD in Computer Science, Data Science, AI/ML, Statistics, or a related field.

Certifications (preferred but not mandatory)

Google Cloud or Azure AI Engineer / Data Scientist Associate

Databricks Certified Machine Learning Professional

DeepLearning.AI Generative AI certification

Hugging Face Transformers certification

Skills Required

  • 8+ years of experience in data science
  • 2+ years of focused hands-on NLP and Generative AI experience
  • Proficiency in Python for data science, modeling, and scripting
  • Hands-on experience with classical and ensemble models (Random Forest, XGBoost)
  • Experience with transformers, tokenization, embeddings, and NLP tasks
  • Experience working with GPT-like models, fine-tuning, and prompt engineering
  • Deep learning experience (PyTorch or TensorFlow)
  • Model deployment experience (REST APIs, Docker, or cloud-native services)
  • Strong SQL and data manipulation skills
  • Applied statistical analysis, hypothesis testing, and A/B testing experience
  • Experience using Git in collaborative environments
  • Experience with large-scale data processing tools and cloud-based ML workflows
  • Experience mentoring and reviewing work of junior data scientists
  • Master's or PhD in Computer Science, Data Science, AI/ML, Statistics, or related field
  • Hands-on experience with Databricks, preferably Azure Databricks
  • Hands-on experience with Delta Lake and ADLS Gen2
  • Experience with vector databases (FAISS, Pinecone, ChromaDB)
  • Experience building RAG (Retrieval-Augmented Generation) pipelines
  • Familiarity with LLM orchestration frameworks (LangChain, LlamaIndex, or similar)
  • Experience with cloud platforms (Azure/GCP/AWS) for ML infrastructure
  • MLOps experience (model tracking, monitoring, CI/CD with MLflow, Kubeflow, etc.)
  • Big data tools familiarity (Spark, Databricks, Hadoop ecosystem)
  • Experiment tracking tools (Weights & Biases, MLflow)
  • Academic research or publications in relevant fields
  • Preferred certifications: Google Cloud/Azure AI Engineer, Databricks ML Professional, DeepLearning.AI Generative AI, Hugging Face Transformers
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
100 Employees

What We Do

NextHire Consulting is an AI-driven recruiting platform that streamlines the hiring process for companies. By leveraging AI agents for sourcing, screening, and interviewing, the platform enables teams to focus on pre-qualified finalists. It provides data-driven insights into candidate soft skills and behavioral styles, aiming to disrupt traditional recruitment models with efficient, automated, and science-based talent acquisition solutions for businesses of all sizes.

Similar Jobs

Micron Technology Logo Micron Technology

Sr MSTI Indirect Regional Supplier Manager

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Gujarat, IND
45000 Employees

Optum Logo Optum

Senior Software Engineer

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Remote
Tamil Nadu, IND
160000 Employees

JPMorganChase Logo JPMorganChase

Data Scientist

Financial Services
Remote or Hybrid
2 Locations
289097 Employees
Remote or Hybrid
2 Locations
289097 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account