Research Intern

Posted 10 Days Ago
Be an Early Applicant
Singapore, SGP
In-Office
Internship
Artificial Intelligence • Software
The Role
Conduct research on next-generation video generation models, focusing on distillation, inference efficiency, reward modeling, preference optimization, evaluation, and scalable training infrastructure. Interns will formulate hypotheses, run large-scale experiments, analyze results, develop research tooling, and communicate findings through reports, presentations, demonstrations, and potential conference submissions under senior mentorship.
Summary Generated by Built In

About Cantina

Cantina Labs is a social AI company developing a suite of advanced video generation models. We bring characters to life, transforming how people tell stories, connect, and create. We build and power ecosystems. Cantina, our flagship social AI platform, is just the beginning.

About the Internship

Cantina is growing its research lab in Singapore, and we are looking for exceptional research interns to work with us on the next generation of video models in October 2026.

This is a three month onsite internship designed to give you meaningful ownership of a well-defined research or engineering problem. You will be matched with a project based on your background and interests, working closely with a senior mentor from initial problem formulation through experimentation, evaluation, and, where appropriate, submission to a leading AI conference.

Projects may focus on post-training and inference efficiency for video generation models, reward modeling and preference-based optimization multimodal data systems, or scalable infrastructure for video model training. The primary focus will be your core project, with opportunities to contribute to applied or product-adjacent work where relevant.

What You’ll Work On

Depending on your project, you may:

  • Research and develop distillation methods for large-scale diffusion and flow-based video generation models, including guidance and adversarial distillation

  • Explore techniques that reduce inference cost while preserving or improving generation quality

  • Develop reward models and preference-based optimization methods to improve aesthetics, motion quality, temporal consistency, and prompt adherence

  • Study how base-model behavior affects post-training outcomes and use experimental findings to inform model development

  • Design rigorous evaluations and conduct large-scale experiments on generative video models

  • Contribute to evaluation harnesses, model integrations, research tooling, or other product-adjacent projects related to your core work

  • Document and communicate your findings through research reports, internal presentations, demonstrations, and potential conference submissions

You may be a good fit if you

  • Are currently pursuing a PhD or are a final-year master’s student in computer science, machine learning, computer vision, or a related field

  • Have research experience in generative modeling, computer vision, multimodal learning, or video generation

  • Have hands on experience with diffusion models, flow-based models, model distillation, reinforcement learning, preference optimization, or related post-training techniques

  • Can formulate hypotheses, design controlled experiments, analyze results, and communicate conclusions clearly

  • Are proficient in Python and have hands-on experience with PyTorch, JAX, or another modern machine learning framework

  • Are comfortable working independently on an open-ended research problem while collaborating closely with a mentor and the broader team

Experience with video, image, audio, or other multimodal data is valuable. Publications at leading venues such as NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, or AAAI are a plus, but are not required. We care most about the quality of your thinking, the depth of your technical work, and your ability to learn quickly.

What You Can Expect

  • A defined project and named senior mentor before your first day

  • Weekly one-on-one meetings and clear project milestones

  • A meaningful compute allocation for your research

  • The opportunity to own a complete research or engineering result

  • First-author positioning by default where your contribution supports a publication

  • Timely internal review of research intended for submission

  • Support for conference travel if your paper is accepted

  • Opportunities to demonstrate your work and receive credit for product contributions

  • A competitive monthly stipend

  • Visa and travel support for eligible international candidates

  • Housing support for qualifying international interns in Singapore

  • Equipment and resources needed to complete your work

Internship Details

  • Location: Singapore

  • Duration: Three months

  • Working model: Onsite

  • Start dates: First batch starts in October 2026, second batch starts in January 2027

Skills Required

  • Currently pursuing a PhD or final-year master's degree in computer science, machine learning, computer vision, or a related field
  • Research experience in generative modeling, computer vision, multimodal learning, or video generation
  • Hands-on experience with diffusion models, flow-based models, model distillation, reinforcement learning, preference optimization, or related post-training techniques
  • Ability to formulate hypotheses, design controlled experiments, analyze results, and communicate conclusions clearly
  • Proficiency in Python and hands-on experience with PyTorch, JAX, or another modern machine learning framework
  • Ability to work independently on open-ended research problems while collaborating with a mentor and broader team
  • Experience with video, image, audio, or other multimodal data
  • Publications at leading venues such as NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, or AAAI
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, California
364 Employees
Year Founded: 2023

What We Do

Cantina Labs, founded by Sean Parker, is a new social platform with the most advanced AI character creator. Build, share, and interact with AI bots and your friends directly in the Cantina or across the internet. Cantina bots are lifelike, social creatures, capable of interacting wherever humans go on the internet. Recreate yourself using powerful AI, imagine someone new, or choose from thousands of existing characters. Bots are a new media type that offer a way for creators to share infinitely scalable and personalized content experiences combined with seamless group chat across voice, video, and text.

Similar Jobs

In-Office
Singapore, SGP
929 Employees
In-Office
Singapore, SGP
883 Employees

V-Key Logo V-Key

Security Research (Intern)

Security • Software
In-Office
Singapore, SGP
121 Employees

Yotta Labs Logo Yotta Labs

Research Engineer Intern - AI Systems

Artificial Intelligence • Information Technology • Software
In-Office or Remote
4 Locations
16 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account