"large multimodal models" Jobs
604 open tech roles matching “large multimodal models”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 604 results
Lead research in realtime audio understanding and human AI interaction at a pioneering AI startup.
Lead the design and development of generative world models for autonomous driving at Kodiak Robotics.
Lead the architecture direction for multimodal transformers at Kodiak Robotics, focusing on AI-powered autonomous technology.
Join Cartesia as an Applied Researcher to enhance generative audio models by bridging customer needs with innovative research.
Join Kodiak Robotics as a Senior AI Infrastructure Engineer to optimize model training for autonomous technology.
Join Cartesia as a Research Engineer to design high-quality datasets and engineer data pipelines for cutting-edge AI models.
Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.
Join Ideogram as an Applied ML Engineer to bridge research and product, turning generative models into production features.
Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.
Join Pinterest as a Machine Learning Engineer II to advance vision-centric LLMs and contribute to innovative AI solutions.
Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join Pika as a Research Intern to work on Videogen models and contribute to generative AI and multimedia technology.
Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.
Join Hedra as a Research Engineer to lead the development of action-conditioned world models in a pioneering Physical AI team.
Own the interfaces for Model APIs and Developer Experience at Together AI, focusing on multimodal APIs and ecosystem compatibility.
Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Cantina as a Member of Technical Staff to build and scale data pipelines for large video generation models.