← Back to results

Research Intern

Join Cantina as a Research Intern to work on next-gen video models in a hands-on, mentorship-driven environment.

Location
Singapore
Compensation
Not disclosed
Level
intern
Type
internship · On-site

Posted by employer 4 days ago

First seen on Joblaze 3 days ago

Last verified on the company career page 1 day ago

Apply at Cantina → Save job Scanned from cantina.com

AI in the day-to-day

Cantina develops advanced video generation models and integrates AI into their research projects.

Requirements

Education
Master's degree
Visa
Sponsorship available

Not disclosed in this posting: compensation, years of experience.

Benefits

Housing Support Conference Travel Support Competitive Monthly Stipend Relocation Assistance

Joblaze summary

The Research Intern at Cantina engages in hands-on research and development of advanced video generation models, focusing on areas like post-training efficiency and reward modeling. Candidates should possess a strong background in machine learning, particularly with generative models and frameworks like PyTorch or JAX, and be adept at conducting experiments and analyzing results. This role is ideal for PhD students or final-year master's candidates with relevant research experience, looking to contribute to cutting-edge AI projects in a collaborative environment.

Joblaze insights

Quick facts

Is the Research Intern role remote?
No — this is an on-site role in Singapore.
Where is the role based?
Cantina is hiring for this position in Singapore.
What's the tech stack?
Joblaze extracted these technologies from the posting: Computer Vision, JAX, PyTorch, Python, generative modeling, multimodal learning.
Does Cantina sponsor work visas for this role?
Yes — the posting indicates visa sponsorship is available for the right candidate.
What seniority level is this role?
Cantina targets intern candidates for this position.
Is this full-time or contract?
Internship for this Research Intern role at Cantina.

From the original posting

About Cantina

Cantina Labs is a social AI company developing a suite of advanced video generation models. We bring characters to life, transforming how people tell stories, connect, and create. We build and power ecosystems. Cantina, our flagship social AI platform, is just the beginning.

About the Internship

Cantina is growing its research lab in Singapore, and we are looking for exceptional research interns to work with us on the next generation of video models in October 2026.

This is a three month onsite internship designed to give you meaningful ownership of a well-defined research or engineering problem. You will be matched with a project based on your background and interests, working closely with a senior mentor from initial problem formulation through experimentation, evaluation, and, where appropriate, submission to a leading AI conference.

Projects may focus on post-training and inference efficiency for video generation models, reward modeling and preference-based optimization multimodal data systems, or scalable infrastructure for video model training. The primary focus will be your core project, with opportunities to contribute to applied or product-adjacent work where relevant.

What You’ll Work On

Depending on your project, you may:

  • Research and develop distillation methods for large-scale diffusion and flow-based video generation models, including guidance and adversarial distillation

  • Explore techniques that reduce inference cost while preserving or improving generation quality

  • Develop reward models and preference-based optimization methods to improve aesthetics, motion quality, temporal consistency, and prompt adherence

  • Study how base-model behavior affects post-training outcomes and use experimental findings to inform model development

  • Design rigorous evaluations and conduct large-scale experiments on generative video models

  • Contribute to evaluation harnesses, model integrations, research tooling, or other product-adjacent projects related to your core work

  • Document and communicate your findings through research reports, internal presentations, demonstrations, and potential conference submissions

You may be a good fit if you

  • Are currently pursuing a PhD or are a final-year master’s student in computer science, machine learning, computer vision, or a related field

  • Have research experience in generative modeling, computer vision, multimodal learning, or video generation

  • Have hands on experience with diffusion models, flow-based models, model distillation, reinforcement learning, preference optimization, or related post-training techniques

  • Can formulate hypotheses, design controlled experiments, analyze results, and communicate conclusions clearly

  • Are proficient in Python and have hands-on experience with PyTorch, JAX, or another modern machine learning framework

  • Are comfortable working independently on an open-ended research problem while collaborating closely with a mentor and the broader team

Experience with video, image, audio, or other multimodal data is valuable. Publications at leading venues such as NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, or AAAI are a plus, but are not required. We care most about the quality of your thinking, the depth of your technical work, and your ability to learn quickly.

What You Can Expect

  • A defined project and named senior mentor before your first day

  • Weekly one-on-one meetings and clear project milestones

  • A meaningful compute allocation for your research

  • The opportunity to own a complete research or engineering result

  • First-author positioning by default where your contribution supports a publication

  • Timely internal review of research intended for submission

  • Support for conference travel if your paper is accepted

  • Opportunities to demonstrate your work and receive credit for product contributions

  • A competitive monthly stipend

  • Visa and travel support for eligible international candidates

  • Housing support for qualifying international interns in Singapore

  • Equipment and resources needed to complete your work

Internship Details

  • Location: Singapore

  • Duration: Three months

  • Working model: Onsite

  • Start dates: First batch starts in October 2026, second batch starts in January 2027

Similar positions

Cantina
Machine Learning Intern
Cantina · Singapore
Cantina
Cantina
Research Scientist, Video Foundation Models
Cantina · California, United States
Cantina