"multimodal systems" Jobs
414 open tech roles matching “multimodal systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 414 results
Join Cantina as a Machine Learning Engineer to develop cutting-edge speech and audio generation systems in a collaborative environment.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.
Join Applied Intuition as a System Designer to create scalable design systems for complex software experiences.
Join Cantina as a Research Scientist to develop next-generation video foundation models and shape the future of AI-driven creativity.
Lead research in realtime audio understanding and human AI interaction at a pioneering AI startup.
Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.
Join Cartesia as a Researcher in London to advance AI through innovative neural network architecture design.
Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.
Join Ambient.ai as a Senior Applied Research Scientist to develop cutting-edge foundation models for computer vision in a hybrid work environment.
Join Kodiak Robotics as a Senior AI Infrastructure Engineer to optimize model training for autonomous technology.
Lead fine-tuning and development of multimodal models for document translation at DeepL.
Join Hedra as a Senior Full-Stack Engineer to design and scale visual intelligence products in a collaborative environment.
Join Tavus as a Senior Software Engineer to drive the development of a multimodal real-time conversational product.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.
Lead research in audio-visual avatar generation at Tavus, shaping the future of human-AI interaction.