"vision language models" Jobs
1437 open tech roles matching “vision language models”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Go. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 1437 results
Join Synthesia as a Staff Research Engineer to shape the future of interactive multimodal systems in AI video communication.
Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.
Join Skydio as a Deep Learning Infrastructure Engineer to build and scale AI solutions for autonomous drones.
Join Verkada as a Senior Software Engineer to develop AI and machine learning models for advanced video analytics.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced machine learning models.
Join Salient as an Applied AI Engineer to build and improve production speech models for voice agents in the financial services sector.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.
Join Cantina as a Machine Learning Engineer to build advanced speech systems and contribute to innovative AI technology.
Join Genmo as a Research Engineer to advance visual generative AI in a fast-paced startup environment.
Join Reflection AI as a Forward Deployed Engineer to fine-tune models and work directly with enterprise customers in a dynamic startup environment.
Join Hedra as a Research Scientist to lead innovative research in generative AI and physical systems with access to large-scale compute.
Join Skydio as a Deep Learning Infrastructure Engineer to build and scale systems for autonomous flight and computer vision.
Serve as the primary language resource for voice AI in English, managing quality and providing linguistic expertise.
Join Preference Model as a Research Engineer to advance self-directed learning in machine learning with a focus on RL environments.
Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.
Join LlamaIndex as an AI Research Engineer to enhance document understanding systems through applied research and engineering.
Lead fine-tuning and development of multimodal models for document translation at DeepL.