Data Scientist I

Posted Yesterday
Warrendale, PA, USA
In-Office
Mid level
Artificial Intelligence • Analytics • Business Intelligence • Consulting
The Role
Design, build, and steward predictive and GenAI solutions: feature identification, data acquisition/preparation, hypothesis testing, LLM context creation (embeddings, vectorization, RAG), model development, prompt engineering, integrations (front-end/back-end), testing, CI/CD, and collaboration with stakeholders to deliver and iterate data science projects.
Summary Generated by Built In

About the Role

At Green Cabbage we empower our customers in the procurement and negotiation process with industry data to save money, time and risk. The company is growing rapidly, and we want to set the foundation for accelerating our development team. We are currently designing, building, and optimizing our core data science offerings and looking for a Data Scientist to join our team to steward experimentation and innovation. 

Key Responsibilities: 

  • Analyze and identify the linkages and interactions between the component parts of an entire system.
  • Take ownership of projects, ensuring their successful planning and technical execution. 
  • Partner with team leadership to ensure collective ownership of quality, timelines, and deliverables. 
  • Develop skills outside your comfort zone and encourage others to do the same. 
  • Use the review of work as an opportunity to deepen the expertise of team members. 
  • Address conflicts or issues, engaging in difficult conversations with clients, team members and other stakeholders, escalating where appropriate. 
     

Qualifications:

  • Bachelor's Degree
  • 3 year(s) in a quantitative field (Computer Science, Mathematics, Machine Learning, AI, Statistics, Operational research or equivalent)
  • At least 1 year of direct experience with feature identification in developing predictive models, data acquisition and preparation, and hypothesis testing.
  • At least 1 year experience in Python, R or other relevant language.

You demonstrate experience in: 

  • Heavy contributor in building of AI and GenAI solutions, including but not limited to analytical modeling, prompt engineering, general all-purpose programming (e.g., Python), testing, communication of results, front end and back-end integration, and iterative development with clients 
  • Documenting and analyzing business processes for AI and Generative AI opportunities, including gathering of requirements, creation of initial hypotheses, and development of AI/GenAI solution approach 
  • Experience processing unstructured and structured data to be consumed as context for LLMs, including but not limited to embedding of large text corpus, generative development of SQL queries, building connectors to structured databases; and 
  • Demonstrates abilities and/or a proven record of success learning and performing in functional and technical capacities, including the following areas: 
    • Managing GenAI application including back-end and front-end integrations
    • Using Python (e.g., Pandas, NLTK, Scikit-learn, Keras, etc.), common LLM development frameworks (e.g., Langchain, Semantic Kernel), Relational storage (SQL), Non-relational storage (NoSQL);
    • Experience in analytical techniques such as Machine Learning, Deep Learning and Optimization
    • Vectorization and embedding, prompt engineering, RAG (retrieval augmented generation) workflow development
    • Understanding or hands on experience with Azure, AWS, and / or Google Cloud platforms
    • Experience with Git Version Control, Unit/Integration/End-to-End Testing, CI/CD, release management, etc.

Why Green Cabbage

Green Cabbage, the Global Leader in Procurement Intelligence, provides mid-market and enterprise clients with the data and expertise needed to achieve better deals across technology, third-party labor, marketing, and travel & expense contracts. Our flagship platform, Harvest, empowers procurement teams worldwide to access Green Cabbage’s precise, governed intelligence.

At Green Cabbage, we foster a collaborative, high-energy, and growth-minded culture. We move quickly, value accountability, and celebrate team wins. If you’re passionate about building secure, modern IT systems that keep a business running efficiently and safely, we’d love to hear from you.

We are an Equal Opportunity Employer and make employment decisions based on merit, qualifications, and business needs. We prohibit discrimination and harassment of any kind based on race, color, religion, sex (including pregnancy, sexual orientation, and gender identity), national origin, age, disability, genetic information, veteran status, or any other protected characteristic under federal, state, or local law.

Skills Required

  • Bachelor's Degree
  • 3 years in a quantitative field (Computer Science, Mathematics, Machine Learning, AI, Statistics, Operational Research or equivalent)
  • At least 1 year direct experience with feature identification for predictive models, data acquisition/preparation, and hypothesis testing
  • At least 1 year experience in Python, R, or other relevant language
  • Experience building AI and Generative AI solutions, including analytical modeling, prompt engineering, and client-facing iterative development
  • Documenting and analyzing business processes for AI/GenAI opportunities and developing solution approaches
  • Experience processing unstructured and structured data for LLM context (embeddings, connectors, generative SQL)
  • Managing GenAI applications including back-end and front-end integrations
  • Experience with Python libraries and ML frameworks (Pandas, NLTK, Scikit-learn, Keras)
  • Experience with common LLM development frameworks (Langchain, Semantic Kernel)
  • Experience with relational (SQL) and non-relational (NoSQL) storage and building database connectors
  • Experience in ML techniques: Machine Learning, Deep Learning, and Optimization
  • Experience with vectorization/embeddings, prompt engineering, and RAG workflow development
  • Understanding or hands-on experience with Azure, AWS, and/or Google Cloud platforms
  • Experience with Git version control, unit/integration/end-to-end testing, CI/CD, and release management
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
133 Employees
Year Founded: 2017

What We Do

Green Cabbage is a global leader in procurement intelligence that helps companies transform spend into actionable insights to drive measurable impact. Through its AI-powered Harvest platform, the company specializes in indirect technology, contingency workforce, and marketing categories. It provides spend analytics, market intelligence, and negotiation strategies to help businesses reduce costs and minimize risk across the entire procurement cycle.

Similar Jobs

In-Office
Philadelphia, PA, USA
47 Employees
40K-80K Annually

UL Solutions Logo UL Solutions

Senior Business/Industry Technical Advisor- ComplianceWire®

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
United States
15000 Employees
120K-160K Annually

PNC Bank Logo PNC Bank

Security Analyst: Forensic Investigations

Machine Learning • Payments • Security • Software • Financial Services
Hybrid
Pittsburgh, PA, USA
55000 Employees
75K-112K Annually

PNC Bank Logo PNC Bank

Business Systems Analyst

Machine Learning • Payments • Security • Software • Financial Services
Hybrid
Pittsburgh, PA, USA
55000 Employees

Similar Companies Hiring

Legora Thumbnail
Artificial Intelligence • Legal Tech • Software
New York, New York
700 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account