"multimodal understanding" Jobs

886 open tech roles matching “multimodal understanding”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 886 results

Perplexity

Join Perplexity's Multimodal AI team to design and build innovative human-AI interaction systems.

Perplexity San Francisco Published 2 weeks ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a backend engineer to design and scale distributed systems for real-time voice interactions.

Perplexity AI San Francisco Published 2 weeks ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Kodiak Robotics

Lead the architecture direction for multimodal transformers at Kodiak Robotics, focusing on AI-powered autonomous technology.

Kodiak Robotics Mountain View, CA $230k–$300k/yr Published 1 month ago
Cartesia

Join Cartesia as a Research Engineer to design high-quality datasets and engineer data pipelines for cutting-edge AI models.

Cartesia *HQ - San Francisco, CA Published 8 months ago
Cartesia

Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.

Cartesia *HQ - San Francisco, CA Published 10 months ago
Cantina

Join Cantina as a Research Scientist to develop next-generation video foundation models and shape the future of AI-driven creativity.

Cantina California, United States $200k–$320k/yr Published 2 months ago
Cantina

Join Cantina as a Machine Learning Engineer to develop cutting-edge speech and audio generation systems in a collaborative environment.

Cantina Remote (U.S. or Europe) $200k–$220k/yr Published 1 month ago
Flexible on stack 70% coding
Tavus

Join Tavus as a Conversational Modelling Research Engineer to advance AI Humans through innovative multimodal conversational models.

Tavus Remote Published 4 months ago
Flexible on stack
Pika

Join Pika as a lead Research Scientist to advance real-time multimodal foundation models for creative technology.

Pika Palo Alto HQ Published 3 months ago
Flexible on stack
Cartesia

Lead research in realtime audio understanding and human AI interaction at a pioneering AI startup.

Cartesia *HQ - San Francisco, CA Published 12 months ago
Mind Robotics

Join Mind Robotics as a Research & Modeling Engineer to build and train core models for real-world robotic systems.

Mind Robotics Palo Alto Published 7 months ago
AI-first team
Twelve Labs

Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 5 months ago
Flexible on stack
Mirelo AI

Join Mirelo AI as a Research Scientist to develop cutting-edge multimodal models for audio generation in a rapidly growing company.

Mirelo AI Berlin Published 9 months ago
Flexible on stack 70% coding
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
MaintainX

Own the Generation half of Document Intelligence, transforming multimodal inputs into structured maintenance knowledge.

MaintainX San Francisco Published 1 week ago
Flexible on stack
MaintainX

Own the Generation half of Document Intelligence, turning multimodal primitives into structured maintenance knowledge.

MaintainX Canada Published 3 weeks ago
Cartesia

Lead a data team at Cartesia to enhance the quality of multimodal AI training data and infrastructure.

Cartesia *HQ - San Francisco, CA Published 1 month ago
Heavy meetings
Descript

Join Descript as an Applied Research Scientist to develop multimodal understanding models for innovative AI editing features.

Descript San Francisco, CA or Remote, US $197k–$262.5k/yr Published 2 weeks ago
Flexible on stack