Senior Data Scientist

Posted Yesterday
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Senior level
Healthtech • Software
The Role
Develop and deploy machine learning, generative AI, NLP, and LLM solutions for healthcare data. Build RAG systems, fine-tune models, engineer prompts, create automated pipelines, and analyze EHR and claims data. Develop models for risk stratification, clinical decision support, patient outcomes, population health, and revenue cycle optimization. Apply healthcare data standards, deploy solutions across cloud platforms, create visualizations, and present recommendations to clinicians, product managers, and executives.
Summary Generated by Built In
We are seeking a Senior Data Scientist – Gen AI with hands-on experience in building and deploying data-driven solutions. You will work closely with cross-functional teams to extract insights, develop machine learning
models, and deploy scalable analytics solutions across various cloud platforms. As a Senior Data Scientist – Gen AI, you’ll play a key role in harnessing data to drive better outcomes, improve performance, and enhance the lives of those served by our clients.

Key Responsibilities:

• The ability to design and develop ML, Gen AI, NLP, LLM Models for AI data pipelines. Model components will include data ingestion, preprocessing, Retrieval Augmented Generation (RAG), NLP/LLM model development, fine-tuning and prompt engineering.

• Analyze large, complex healthcare datasets including electronic health records (EHR) and claims data

• Develop statistical models for patient risk stratification, treatment optimization, population health management, and revenue cycle optimization

• Build models for clinical decision support, patient outcome prediction, care quality improvement, and revenue cycle optimization

 • Create and maintain automated data pipelines for real-time analytics and reporting

• Work with healthcare data standards (HL7 FHIR, ICD-10, CPT, SNOMED CT) and ensure regulatory compliance

• Develop and deploy models in cloud environments while creating visualizations for stakeholders

• Present findings and recommendations to cross-functional teams including clinicians, product managers, and executives

Qualifications required:

• Bachelor's degree in data science, Statistics, Computer Science, Mathematics, or related quantitative field

• 5-7 years of hands-on experience in data science, analytics, or machine learning roles

 • Knowledge of ML Ops practices deploying and operationalize AI models.

• Familiarity with cloud platforms (Azure, AWS, GCP) and containerized deployments (Docker, Kubernetes).

 • Demonstrated experience working with large datasets and statistical modeling

• Proficiency in Python or R for data analysis and machine learning

 • Experience with SQL and database management systems

• Knowledge of machine learning frameworks such as scikit-learn, TensorFlow, PyTorch

• Familiarity with data visualization tools such as Tableau, Power BI, matplotlib, ggplot2

 • Experience with version control systems (Git) and collaborative development practices

• Strong foundation in statistics, hypothesis testing, and experimental design

• Experience with supervised and unsupervised learning techniques

• Knowledge of data preprocessing, feature engineering, and model validation

• Understanding of A/B testing and causal inference methods

• Experience with cloud platforms and big data technologies such as Spark, Hadoop

What You’ll Need to Be Successful (Required Skills):

• Large Language Model (LLM) Experience: At least 5 years of hands-on experience working with pretrained language models (GPT, BERT, T5) including fine-tuning, prompt engineering, and model evaluation techniques. LLM experience should include Claude, Llama models. Understanding of RAG techniques is a critical skill.

• Generative AI Frameworks: Proficiency with generative AI libraries and frameworks such as Hugging Face Transformers, Lang Chain, OpenAI API, or similar platforms for building and deploying AI applications

• Prompt Engineering and Optimization: Experience designing, testing, and optimizing prompts for various use cases including text generation, summarization, classification, and conversational AI applications

• Vector Databases and Embeddings: Knowledge of vector similarity search, embedding models, and vector databases (Pinecone, we aviate, Chroma) for building retrieval-augmented generation (RAG) systems

• AI Model Evaluation: Experience with evaluation methodologies for generative models including BLEU scores, ROUGE metrics, human evaluation frameworks, and bias detection techniques

• Multi-modal AI Systems: Familiarity with multi-modal generative models combining text, images, and other data types, including experience with vision-language models and cross-modal applications

• AI Safety and Alignment: Understanding of responsible AI practices including content filtering, bias mitigation, hallucination detection, and techniques for ensuring AI outputs align with business requirements and ethical guidelines

Preferred Skills:

• At least 1 year of experience working with healthcare data or in healthcare IT environments

• Familiarity with electronic health record (EHR) systems and healthcare workflows

• Understanding of healthcare data privacy regulations (HIPAA, HITECH)

• Knowledge of clinical data standards and interoperability frameworks

• Knowledge of MLOps practices and model deployment pipelines

• Familiarity with natural language processing for clinical text analysis

• Experience with time series analysis for patient monitoring data

Netsmart is proud to be an equal opportunity workplace and is an affirmative action employer, providing equal employment and advancement opportunities to all individuals. We celebrate diversity and are committed to creating an inclusive environment for all associates. All employment decisions at Netsmart, including but not limited to recruiting, hiring, promotion and transfer, are based on performance, qualifications, abilities, education and experience. Netsmart does not discriminate in employment opportunities or practices based on race, color, religion, sex (including pregnancy), sexual orientation, gender identity or expression, national origin, age, physical or mental disability, past or present military service, or any other status protected by the laws or regulations in the locations where we operate.

Netsmart desires to provide a healthy and safe workplace and, as a government contractor, Netsmart is committed to maintaining a drug-free workplace in accordance with applicable federal law. Pursuant to Netsmart policy, all post-offer candidates are required to successfully complete a pre-employment background check, which is provided at Netsmart’s sole expense.

Skills Required

  • Bachelor's degree in data science, statistics, computer science, mathematics, or a related quantitative field
  • 5-7 years of hands-on experience in data science, analytics, or machine learning roles
  • At least 5 years of hands-on experience with pretrained language models, including fine-tuning, prompt engineering, and model evaluation
  • Experience with GPT, BERT, T5, Claude, and Llama models
  • Understanding of Retrieval Augmented Generation techniques
  • Proficiency with generative AI frameworks such as Hugging Face Transformers, LangChain, OpenAI API, or similar
  • Experience designing, testing, and optimizing prompts
  • Knowledge of vector similarity search, embedding models, and vector databases such as Pinecone, Weaviate, or Chroma
  • Experience evaluating generative models using BLEU, ROUGE, human evaluation, and bias detection methods
  • Familiarity with multimodal generative models, vision-language models, and cross-modal applications
  • Understanding of responsible AI, content filtering, bias mitigation, hallucination detection, and AI alignment
  • Knowledge of MLOps practices for deploying and operationalizing AI models
  • Familiarity with Azure, AWS, or GCP and containerized deployments using Docker and Kubernetes
  • Experience working with large datasets and statistical modeling
  • Proficiency in Python or R for data analysis and machine learning
  • Experience with SQL and database management systems
  • Knowledge of scikit-learn, TensorFlow, or PyTorch
  • Familiarity with Tableau, Power BI, matplotlib, or ggplot2
  • Experience with Git and collaborative development practices
  • Strong foundation in statistics, hypothesis testing, and experimental design
  • Experience with supervised and unsupervised learning
  • Knowledge of data preprocessing, feature engineering, and model validation
  • Understanding of A/B testing and causal inference methods
  • Experience with cloud platforms and big data technologies such as Spark or Hadoop
  • At least 1 year of experience working with healthcare data or healthcare IT environments
  • Familiarity with EHR systems and healthcare workflows
  • Understanding of HIPAA and HITECH
  • Knowledge of clinical data standards and interoperability frameworks
  • Familiarity with NLP for clinical text analysis
  • Experience with time series analysis for patient monitoring data

NetSmart Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about NetSmart and has not been reviewed or approved by NetSmart.

  • Wellbeing & Lifestyle Benefits Offerings include a designated mental wellness day, paid volunteer time, telehealth, wellness events, and free 1:1 sessions with counselors and coaches. These wellbeing resources are positioned as stronger than many employers.
  • Healthcare Strength Medical, dental, and vision coverage are paired with short- and long‑term disability and telehealth. The breadth of core health coverage aligns with a comprehensive package.
  • Leave & Time Off Breadth PTO, paid holidays, paid volunteer time, and a designated mental wellness day are part of the offering, with some materials referencing flexible or unlimited PTO. This variety provides multiple avenues for time away.

NetSmart Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Overland Park, KS
1,919 Employees
Year Founded: 1968

What We Do

Netsmart designs, builds and delivers electronic health records (EHRs), solutions and services that are powerful, intuitive and easy-to-use. Our platform provides accurate, up-to-date information that is easily accessible to care team members in behavioral health, care at home, senior living and social services. We make the complex simple and personalized so our clients can concentrate on what they do best: provide services and treatment that support whole-person care. By leveraging the powerful Netsmart network, care providers can seamlessly and securely integrate information across communities, collaborate on the most effective treatments and improve outcomes for those in their care. Our streamlined systems and personalized workflows put relevant information at the fingertips of users when and where they need it. For 50 years, Netsmart has been committed to providing a common platform to integrate care. SIMPLE. PERSONAL. POWERFUL. Our more than 2,200 associates work hand-in-hand with our 600,000+ users in more than 25,000 organizations across the U.S. to develop and deploy technology that automates and coordinates everything from clinical to financial to administrative.

Similar Jobs

JPMorganChase Logo JPMorganChase

Data Scientist

Financial Services
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees

Optum Logo Optum

Senior Data Scientist

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
160000 Employees

EXL Logo EXL

Senior Data Scientist

Information Technology • Database • Consulting
Remote or Hybrid
Karnataka, IND
30246 Employees

EXL Logo EXL

Senior Data Scientist

Information Technology • Database • Consulting
Remote or Hybrid
Karnataka, IND
30246 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account