Staff Machine Learning Engineer

Posted One Month Ago
Be an Early Applicant
Hiring Remotely in Singapore, SGP
In-Office or Remote
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
The Role
Lead production-grade ML systems: build and operate end-to-end ML pipelines, fine-tune large models (LoRA/QLoRA/SFT/DPO/distillation), design scalable inference, ensure GPU optimization, implement evaluation for performance/robustness/safety, and collaborate with product and application teams to ship reliable, low-latency AI features.
Summary Generated by Built In
About the Role

There are over 5 billion users using basic applications today such email, notes, tasks, calendar and they're not AI-native. Our mission is to build proactive applications for anyone in the world, who are not used to complex prompting. We aim to bring intelligence to conversations, errands, organising and workflows, with minimal to no prompting.

Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. We believe products will greatly reduce hallucinations

Our objective is to organise anyone's life, allowing us all to spend time on valuable and meaningful things

As Staff Machine Learning Engineer, you own the execution layer of our intelligence, turning research and model capabilities into reliable, scalable production systems.

You will work across the model lifecycle: data, training, evaluation, inference, and deployment. This is a hands-on leadership role for someone who wants to operate at the intersection of research, systems, and product.

 
What You'll Own
  • Own the end-to-end ML systems powering our company, from data and training to evaluation, inference, and deployment.

  • Build and evolve training and fine-tuning pipelines for large models.

  • Design evaluation systems that measure capability, robustness, safety, and real-world product performance.

  • Architect high-performance inference systems, optimizing latency, GPU utilization, memory, cost, and reliability.

  • Build data pipelines and systems for high-quality real-world and synthetic training data.

  • Establish reliable production infrastructure for deploying, monitoring, and continuously improving models.

  • Partner closely with research and application engineering to turn model capabilities into product improvements.

  • Make pragmatic technical trade-offs and rapidly iterate based on real-world performance.

 
What We're Looking For
  • Experience building and shipping ML systems used in production, not just research prototypes.

  • Strong understanding of modern large-model training, fine-tuning, evaluation, and inference.

  • Strong software engineering and systems fundamentals.

  • Experience operating ML workloads at meaningful scale, particularly GPU-based systems.

  • Strong technical judgment and the ability to navigate ambiguous problems independently.

  • A bias toward experimentation, measurement, and shipping.

  • High standards for correctness, reliability, and production quality.

 
Outcomes
  • Research and models reliably translate into production-ready solutions with clear performance and quality targets.

  • ML pipelines, training loops, and inference systems are stable, efficient, and maintainable.

  • Production issues are detected, debugged, and resolved quickly, minimizing user impact.

  • Team members are supported, aligned, and able to deliver high-impact ML work with minimal friction.

  • Iterations on models and systems are measurable, safe, and improve user experience over time.

 
Tech Stack
  • Python

  • PyTorch / JAX

  • GPU-based training and inference system

 
Ideal Experience
  • You have built or shipped real ML systems used by people, not just demos.

  • You are comfortable working with large models and understanding their failure modes.

  • You write strong, production-grade code and care about system correctness.

 
How We Work

We are a small, high-talent-density, hands-on team. Engineers have broad ownership and are expected to exercise strong judgment and execute independently.

We make decisions quickly, work closely together, and balance speed with engineering fundamentals. We care less about process and more about building something exceptional.

 
Interview process

If there appears to be a fit, we'll reach to schedule 3, but no more than 4 interviews.

Applications are evaluated by our technical team members. Interviews will be conducted via virtual meetings and/or onsite.

We value transparency and efficiency, so expect a prompt decision. If you've demonstrated the exceptional skills and mindset we're looking for, we'll extend an offer to join us. This isn't just a job offer; it's an invitation to be part of a team that's bringing AI to have practical benefits to billions globally.

Skills Required

  • Proficiency in Python
  • Experience with PyTorch or JAX
  • Experience with GPU-based training and inference systems
  • Experience fine-tuning and adapting large models using LoRA, QLoRA, SFT, DPO, and distillation
  • Built or shipped production ML systems (data pipelines, training workflows, evaluation, inference, deployment)
  • Experience optimizing GPU performance, memory efficiency, and latency reduction
  • Designing and maintaining data systems for synthetic and real-world training data
  • Implementing evaluation pipelines for performance, robustness, safety, and bias
  • Strong production-grade coding and system correctness practices
  • Ability to collaborate with application engineering to integrate ML into backend, mobile, and desktop products
  • Self-directed, pragmatic, takes ownership and communicates clearly in small teams
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Petaling Jaya
253 Employees
Year Founded: 2019

What We Do

Our mission is to develop technology based solutions to improve financial inclusion. We develop new & innovative platforms & services globally. For example, we are the first platform to simplify and digitise comprehensive life and medical insurance, supported by AI agent. BJAK is the largest insurance platform in Southeast Asia. If you enjoy building cutting edge platform-ecosystems that gives equal access to financial services to everyone at scale, join us

Similar Jobs

Mondelēz International Logo Mondelēz International

Analytics Manager

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
HarbourFront, SGP
90000 Employees

CSC Logo CSC

Senior Transaction Manager (Legal)

Fintech • Legal Tech • Software • Financial Services • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
8500 Employees

Citadel Logo Citadel

Machine Learning Researcher - PhD Intern (Asia)

Information Technology • Software • Financial Services • Big Data Analytics
In-Office or Remote
2 Locations
4000 Employees

Atlassian Logo Atlassian

Solution Engineer- Mandarin Speaking

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
In-Office or Remote
Singapore, SGP
11000 Employees

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account