Data Scientist

Posted 10 Days Ago
Be an Early Applicant
Hiring Remotely in Región de Callao, PER
Remote
Senior level
Artificial Intelligence • Cloud • Machine Learning • Software
The Role
Design, build, and evaluate ML and LLM-based matching systems for company/entity resolution. Develop embeddings, scoring, ranking, and classification models on messy, multilingual data. Define benchmarks, metrics, baselines, and run experiments with robust error analysis while considering inference cost, scalability, and production maintenance. Communicate results and recommendations to engineering and business stakeholders.
Summary Generated by Built In
Job Description

About You

You are an experienced Data Scientist with strong applied Machine Learning expertise and a track record of building and evaluating models using real-world, messy, large-scale data. You are comfortable working with embeddings, semantic similarity, LLMs, NLP, classification, and both supervised and unsupervised learning. You approach ambiguous problems through structured experimentation, clearly defined hypotheses, baselines, metrics, and error analysis.

You are highly autonomous, intellectually honest about experimental results, and able to clearly communicate technical recommendations and trade-offs to engineering and business stakeholders.

You Bring to Applaudo the Following Competencies

  • 5+ years of professional Data Science / Machine Learning experience.
  • Strong applied Machine Learning fundamentals.
  • Excellent Python and SQL skills.
  • Hands-on experience with embeddings and semantic similarity.
  • Practical experience applying LLMs to real-world problems.
  • Experience with supervised and unsupervised learning.
  • Strong experience with classification and NLP.
  • Working knowledge of neural networks and transformer architectures.
  • Hands-on experience with TensorFlow, PyTorch, PyCaret, or equivalent ML frameworks.
  • Experience retraining or maintaining classification models in production.
  • Strong experimental design and model evaluation skills.
  • Experience defining baselines, metrics, test sets, and error-analysis processes.
  • Ability to evaluate model quality and demonstrate measurable improvements.
  • Strong understanding of scalability and ML inference costs.
  • Strong English communication skills.

Nice-to-Have

  • Entity resolution, record linkage, or deduplication experience.
  • Ranking and similarity scoring.
  • Retrieval, clustering, or candidate-generation techniques.
  • LLM/embedding solutions designed for cost and scale constraints.
  • Spark, Snowflake, Databricks, or BigQuery.
  • Experience with company, domain, website, or firmographic data.
  • Experience working with multilingual datasets.

You Will Be Accountable for the Following Responsibilities

  • Build and evaluate ML approaches for company/entity matching.
  • Develop embedding and LLM-based matching approaches.
  • Develop scoring and ranking methodologies to identify true matches and distinguish them from duplicates, lookalikes, and unrelated entities.
  • Work with messy data, including names, aliases, domains, websites, firmographic attributes, multilingual records, and data hierarchies.
  • Define benchmark datasets, metrics, baselines, and error-analysis processes.
  • Design and execute experiments to validate hypotheses.
  • Compare LLM-assisted approaches against lower-cost alternatives.
  • Analyze model behavior, edge cases, and trade-offs.
  • Consider inference economics and scalability from the beginning.
  • Communicate experimental findings and recommendations to engineering and business stakeholders.
  • Independently establish experimental pipelines and research approaches.
  • Clearly document both successful and unsuccessful experiments.

What Sets You Apart

  • Strong analytical and experimental mindset.
  • Intellectual honesty and willingness to communicate negative results.
  • Strong autonomy and self-direction.
  • Excellent written and verbal communication.
  • Ability to defend technical recommendations with stakeholders.
  • Strong problem-solving skills.
  • Comfort working with ambiguity and large-scale datasets.
  • Ability to balance model quality, cost, and scalability.

Additional Information

About Us

We Are Engineered Different.

At Applaudo, talented people design, build, and scale meaningful, AI-powered solutions that create real business impact. As an AI-native organization, we collaborate across design, development, cloud, data, and artificial intelligence to turn ideas into scalable products that transform how companies operate, make decisions, and grow.

We are building a high-performance culture grounded in five values: Empowering Excellence, Collaborative Teamwork, Unsolicited Respect, Consistent Transparency, and Efficient Communication. These define how we work, how we support one another, and how we hold ourselves accountable.

Applaudo is a place for people who want to learn fast, take ownership, and work alongside strong teams they are proud to belong to. Joining us means being part of an organization that is evolving intentionally, investing in modern ways of working, and leading AI-native transformation at scale.

Skills Required

  • 5+ years of professional Data Science / Machine Learning experience.
  • Strong applied Machine Learning fundamentals.
  • Excellent Python and SQL skills.
  • Hands-on experience with embeddings and semantic similarity.
  • Practical experience applying LLMs to real-world problems.
  • Experience with supervised and unsupervised learning.
  • Strong experience with classification and NLP.
  • Working knowledge of neural networks and transformer architectures.
  • Hands-on experience with TensorFlow, PyTorch, PyCaret, or equivalent ML frameworks.
  • Experience retraining or maintaining classification models in production.
  • Strong experimental design and model evaluation skills.
  • Experience defining baselines, metrics, test sets, and error-analysis processes.
  • Ability to evaluate model quality and demonstrate measurable improvements.
  • Strong understanding of scalability and ML inference costs.
  • Strong English communication skills.
  • Entity resolution, record linkage, or deduplication experience.
  • Ranking and similarity scoring.
  • Retrieval, clustering, or candidate-generation techniques.
  • LLM/embedding solutions designed for cost and scale constraints.
  • Spark, Snowflake, Databricks, or BigQuery.
  • Experience with company, domain, website, or firmographic data.
  • Experience working with multilingual datasets.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Salvador
471 Employees
Year Founded: 2013

What We Do

We are a Nearshore digital solutions company powered by our LATAM based tech talent. Our specialties include Digital Transformation, Web and Mobile Development, Cloud Computing, AI, and Machine Learning, among others. We are committed to delivering high-quality software solutions that not only are scalable and dependable but also future proof. We accelerate our customers digital roadmap by leveraging our 10 years of experience building custom digital solutions, augmenting our clients' teams, and reducing time-to-market. Our Vision: Code that changes lives; has made us believe that the power of innovation can change the world. Contact us to see how we can help you achieve your business goals. www.applaudo.com We are hiring! We are looking for diverse and talented professionals around the world, who share our commitment to making a positive impact on those surrounding us. Working hand in hand to develop our skills, we will continue daring each other to reach new horizons and overcoming barriers. Visit our Jobs tab to see the multiple open positions waiting for you. Apply now and discover why Applaudo is the Best Place to Code.

Similar Jobs

DevSavant Logo DevSavant

Data Scientist

Information Technology • Software
Remote
11 Locations
93 Employees

ON.energy Logo ON.energy

Data Scientist

Artificial Intelligence • Energy • Renewable Energy
Remote
12 Locations
165 Employees

RevenueCat Logo RevenueCat

Senior Data Scientist

Fintech • Mobile • Payments • Software
Remote
41 Locations
40 Employees

RevenueCat Logo RevenueCat

Senior Data Scientist

Fintech • Mobile • Payments • Software
Remote
14 Locations
40 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account