"multimodal modeling" Jobs

724 open tech roles matching “multimodal modeling”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 724 results

MaintainX

Own the Generation half of Document Intelligence, transforming multimodal inputs into structured maintenance knowledge.

MaintainX San Francisco Published 1 week ago
Flexible on stack
Ambient

Join Ambient.ai as a Senior Applied Research Scientist to develop cutting-edge foundation models for computer vision in a hybrid work environment.

Ambient Redwood City Published 1 year ago
Flexible on stack 70% coding
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to fine-tune multimodal models for clinical prediction in drug discovery.

Iambic Therapeutics Boston Office Published 1 week ago
Flexible on stack
MaintainX

Own the Generation half of Document Intelligence, turning multimodal primitives into structured maintenance knowledge.

MaintainX Canada Published 3 weeks ago
Descript

Join Descript as an Applied Research Scientist to develop multimodal understanding models for innovative AI editing features.

Descript San Francisco, CA or Remote, US $197k–$262.5k/yr Published 2 weeks ago
Flexible on stack
Twelve Labs

Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 5 months ago
Flexible on stack
Synthesia

Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.

Synthesia Europe Published 1 month ago
Flexible on stack
Pinterest

Join Pinterest as a Machine Learning Engineer II to advance vision-centric LLMs and contribute to innovative AI solutions.

Pinterest San Francisco, CA, US; Remote, US $138.9k–$286.0k/yr Published 6 months ago
Flexible on stack
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
Together AI

Own the interfaces for Model APIs and Developer Experience at Together AI, focusing on multimodal APIs and ecosystem compatibility.

Together AI San Francisco $200k–$280k/yr Published 3 weeks ago
Gamma

Join Gamma as a Research Engineer to fine-tune vision-language models for exceptional visual communication.

Gamma San Francisco $180k–$340k/yr Published 10 months ago
Flexible on stack
Sesame

Join Sesame as a Data Engineer to build and maintain data pipelines for AI models in a team of experts from leading tech companies.

Sesame San Francisco Published 2 months ago
Flexible on stack
Cantina

Join Cantina as a Senior Machine Learning Engineer to develop innovative AI image generation models for lifelike AI bots.

Cantina Bay Area or Remote $200k–$265k/yr Published 6 months ago
Flexible on stack
Pika

Join Pika as a Research Intern to work on Videogen models and contribute to generative AI and multimedia technology.

Pika Palo Alto HQ Published 2 months ago
Flexible on stack
Tavus

Lead research in audio-visual avatar generation at Tavus, shaping the future of human-AI interaction.

Tavus San Francisco Published 4 months ago
Flexible on stack
Cantina

Join Cantina as an ML Engineer in Singapore to build and scale systems for processing large-scale video and multimodal data.

Cantina Singapore Published 4 months ago
Flexible on stack
Pinterest

Join Pinterest as a Sr. Machine Learning Engineer to develop innovative generative models for visualization features.

Pinterest San Francisco, CA, US; Remote, US $161.3k–$332.0k/yr Published 1 year ago
Flexible on stack
Cartesia

Join Cartesia as a Software Engineer to shape data infrastructure for cutting-edge AI models in a collaborative, in-office environment.

Cartesia *HQ - San Francisco, CA Published 2 months ago
Flexible on stack
Twelve Labs

Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.

Twelve Labs Seoul, South Korea Published 3 weeks ago
Flexible on stack