← Back to results

Conversational Modelling Research Engineer

Join Tavus as a Conversational Modelling Research Engineer to advance AI Humans through innovative multimodal conversational models.

Location
Remote
Compensation
Not disclosed
Level
mid
Type
full time · Hybrid

Posted by employer 4 months ago

First seen on Joblaze 1 week ago

Last verified on the company career page 1 day ago

Apply at Tavus → Save job Scanned from tavus.io

AI in the day-to-day

Conduct research on Large Multimodal Models for Conversational Avatars and partner with Applied ML team to move prototypes to production.

Requirements

Education
PhD

Not disclosed in this posting: compensation, years of experience, visa sponsorship.

Joblaze summary

In this role, the Conversational Modelling Research Engineer at Tavus focuses on advancing Foundation Multimodal Conversational Models, particularly for real-time avatar interactions. The position requires expertise in large multimodal models, deep learning, and proficiency in PyTorch, emphasizing the development of expressive and controllable AI systems. Ideal candidates are those with a PhD or equivalent experience, who thrive in dynamic startup settings and possess a strong research background. Tavus, a Series B company, is at the forefront of creating AI that enhances human-machine communication.

Joblaze insights

Quick facts

Is the Conversational Modelling Research Engineer role remote?
It's hybrid — Tavus expects some on-site time in Remote.
Where is the role based?
Tavus is hiring for this position in Remote.
What's the tech stack?
Joblaze extracted these technologies from the posting: Deep Learning, Large Multimodal Models, PyTorch, generative models.
What seniority level is this role?
Tavus targets mid-level candidates for this position.
Is this full-time or contract?
Full-time for this Conversational Modelling Research Engineer role at Tavus.

From the original posting

About Us

Tavus is a research lab pioneering human computing. We’re building AI Humans: a new interface that closes the gap between people and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real—enabling meaningful, face-to-face conversations. AI Humans combine the emotional intelligence of humans with the reach and reliability of machines, making them capable, trusted agents available 24/7, in every language, on our terms.

Imagine a friend who can discuss any topic with you. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. With Tavus, individuals, enterprises, and developers can all build AI Humans to connect, understand, and act with empathy at scale.

We’re a Series B company backed by world-class investors including Sequoia Capital, Y Combinator, and Scale Venture Partners.

Be part of shaping a future where humans and machines truly understand each other.

The Role

We’re looking for an AI Researcher to join our core AI team and push the boundaries of Foundation Multimodal Conversational Models. If you thrive in fast-moving startup environments, enjoy experimenting with new ideas, and love seeing your work come to life in production then you’ll feel right at home.

Your Mission 🚀

  • Conduct research on Large Multimodal Models in the context of Conversational Avatars (e.g. Neural Avatars, Talking-Heads).

  • Develop methods to model both verbal and non-verbal aspects of conversation, adapting and controlling avatar behavior in real time, with low-latency.

  • Experiment with fine-tuning, adaptation, and conditioning techniques to make AudioVisual Multimodal Models, more expressive, controllable, and task-specific.

  • Partner with the Applied ML team to take research from prototype to production.

  • Stay up to date with cutting-edge advancements — and help define what comes next.

You’ll Be Great At This If You Have:

  • A PhD (or near completion) in a relevant field, or equivalent research experience.

  • Hands-on experience with Large Multimodal Models and a strong foundation in generative (language) models. This could be in the context of tasks such as VQA, Audio/Video understanding tasks, captioning behavioral analysis, Translation tasks, Speech to Speech systems.

  • Experience in fine-tuning/adapting VLMs for control, conditioning, or downstream tasks.

  • Solid background in deep learning and foundation modes.

  • Strong PyTorch skills and comfort building deep learning pipelines.

Nice-to-Haves

  • Knowledge of large-scale model training and optimization.

  • Experience in duplex-conversational model.

  • Broader understanding of generative AI across modalities.

  • Exposure to software development best practices.

  • A flexible, experimental mindset i.e. comfortable working across research and engineering.

  • (Bonus) Publications at EMNLP, COLING, NeurIPS, ICLR, CVPR, ICCV.

Location


Preferred: San Francisco (hybrid) or London.

Remote within the U.S. or Europe available for exceptional candidates.

Similar positions

Tavus
Senior Software Engineer (CVI)
Tavus · San Francisco
Tavus
Data Engineer /ML Ops
Tavus · Remote
Cartesia
Applied Researcher, Audio
Cartesia · *HQ - San Francisco, CA
Tavus