"multimodal understanding" Jobs
588 open tech roles matching “multimodal understanding”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, SQL. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 588 results
Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.
Join Cartesia as a Software Engineer to shape data infrastructure for cutting-edge AI models in a collaborative, in-office environment.
Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.
Own the interfaces for Model APIs and Developer Experience at Together AI, focusing on multimodal APIs and ecosystem compatibility.
Lead research in audio-visual avatar generation at Tavus, shaping the future of human-AI interaction.
Join Pinterest as a Sr. Machine Learning Engineer to develop innovative generative models for visualization features.
Own and build internal products to support model evaluation and data management in a hybrid role at Twelve Labs.
Build the platform behind Managed Agents, shaping the latency, reliability, and developer experience of real-time voice agents.
Lead the application layer at Twelve Labs, building AI-native products that transform user experiences in video technology.
Coordinate global mobility operations, ensuring compliance and supporting employees through relocations and international assignments.
Join Anthropic as a Software Engineer to build secure and scalable infrastructure for AI interpretability research.
Join Cartesia as a lead researcher to design evaluation frameworks for next-generation AI models.
Join Tavus as a Senior Software Engineer to drive the development of a multimodal real-time conversational product.
Join Cartesia as an Enterprise Engineer to build real-time multimodal intelligence and deliver voice AI solutions for enterprise customers.
Join Pinterest as a Staff Machine Learning Engineer to develop state-of-the-art visual AI models in a collaborative environment.
Join Hedra as a Research Engineer to lead the development of action-conditioned world models in a pioneering Physical AI team.