"multimodal large language models" Jobs
136 open tech roles matching “multimodal large language models”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Machine Learning. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 136 results
Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.
Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.
Join Tavus as a Conversational Modelling Research Engineer to advance AI Humans through innovative multimodal conversational models.
Lead research in realtime audio understanding and human AI interaction at a pioneering AI startup.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.
Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.
Join Cartesia as a Research Engineer to design high-quality datasets and engineer data pipelines for cutting-edge AI models.
Join Descript as an Applied Research Scientist to develop multimodal understanding models for innovative AI editing features.
Join Pinterest as a Machine Learning Engineer II to advance vision-centric LLMs and contribute to innovative AI solutions.
Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.
Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.
Join Cartesia as an Applied Researcher to enhance generative audio models by bridging customer needs with innovative research.
Lead research in audio-visual avatar generation at Tavus, shaping the future of human-AI interaction.
Related searches