"large multimodal models" Jobs

604 open tech roles matching “large multimodal models”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 604 results

Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to innovate AI-based drug discovery with multimodal models.

Iambic Therapeutics UK Office Published 2 weeks ago
Flexible on stack
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to innovate AI-based drug discovery with multimodal models.

Iambic Therapeutics Boston Office Published 2 weeks ago
Flexible on stack
Mind Robotics

Join Mind Robotics as a Research & Modeling Engineer to build and train core models for real-world robotic systems.

Mind Robotics Palo Alto Published 7 months ago
AI-first team
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Mirelo AI

Join Mirelo AI as a Research Scientist to develop cutting-edge multimodal models for audio generation in a rapidly growing company.

Mirelo AI Berlin Published 9 months ago
Flexible on stack 70% coding
Cantina

Join Cantina as a Research Scientist to develop next-generation video foundation models and shape the future of AI-driven creativity.

Cantina California, United States $200k–$320k/yr Published 2 months ago
Pika

Join Pika as a lead Research Scientist to advance real-time multimodal foundation models for creative technology.

Pika Palo Alto HQ Published 3 months ago
Flexible on stack
Ambient

Join Ambient.ai as a Senior Applied Research Scientist to develop cutting-edge foundation models for computer vision in a hybrid work environment.

Ambient Redwood City Published 1 year ago
Flexible on stack 70% coding
Cantina

Join Cantina as a Machine Learning Engineer to develop cutting-edge speech and audio generation systems in a collaborative environment.

Cantina Remote (U.S. or Europe) $200k–$220k/yr Published 1 month ago
Flexible on stack 70% coding
Cartesia

Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.

Cartesia *HQ - San Francisco, CA Published 10 months ago
Cartesia

Join Cartesia as a Researcher in London to advance AI through innovative neural network architecture design.

Cartesia London Published 11 months ago
Flexible on stack
Xaira Therapeutics

Join Xaira Therapeutics as an AI Scientist to develop multimodal foundation models for drug discovery using advanced AI techniques.

Xaira Therapeutics South San Francisco, California, United States $170k–$240k/yr Published 4 months ago
Flexible on stack
Google DeepMind
Google DeepMind Los Angeles, California, US; Mountain View, California, US $174k–$252k/yr Published 5 months ago
Tavus

Join Tavus as a Conversational Modelling Research Engineer to advance AI Humans through innovative multimodal conversational models.

Tavus Remote Published 4 months ago
Flexible on stack
Cantina

Join Cantina as an ML Engineer in Singapore to build and scale systems for processing large-scale video and multimodal data.

Cantina Singapore Published 4 months ago
Flexible on stack
Cartesia

Lead a data team at Cartesia to enhance the quality of multimodal AI training data and infrastructure.

Cartesia *HQ - San Francisco, CA Published 1 month ago
Heavy meetings
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
Iambic Therapeutics

Join Iambic Therapeutics as a Machine Learning Scientist to fine-tune multimodal models for clinical prediction in drug discovery.

Iambic Therapeutics Boston Office Published 1 week ago
Flexible on stack
Twelve Labs

Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 5 months ago
Flexible on stack