"multimodal learning" Jobs
415 open tech roles matching “multimodal learning”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 415 results
Join Cantina as a Machine Learning Engineer to develop cutting-edge speech and audio generation systems in a collaborative environment.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.
Join Mirelo AI as a Research Scientist to develop cutting-edge multimodal models for audio generation in a rapidly growing company.
Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.
Join Cantina as a Research Scientist to develop next-generation video foundation models and shape the future of AI-driven creativity.
Join Cantina as a Senior Machine Learning Engineer to develop innovative AI image generation models for lifelike AI bots.
Join Pinterest as a Sr. Machine Learning Engineer to develop innovative generative models for visualization features.
Join Reddit as a Senior ML Engineer to develop advanced embedding models for advertising in a fully remote role.
Join Reddit as a Senior ML Engineer to develop advanced embedding models for advertising in a fully remote role.
Lead research in realtime audio understanding and human AI interaction at a pioneering AI startup.
Join Ambient.ai as a Senior Applied Research Scientist to develop cutting-edge foundation models for computer vision in a hybrid work environment.
Join Protege as a Machine Learning Researcher to lead the evaluation and optimization of audio data quality for AI training.
Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.
Lead fine-tuning and development of multimodal models for document translation at DeepL.
Join Sarvam AI as a researcher to develop vision-language models that impact AI applications in India.