"multimodal large language models" Jobs
213 open tech roles matching “multimodal large language models”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 213 results
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Join Pinterest as a Machine Learning Engineer II to advance vision-centric LLMs and contribute to innovative AI solutions.
Join Cloudflare as a Senior Machine Learning Engineer to optimize and productionize ML models for a global serverless inference platform.
Join Cartesia as a Researcher to enhance multimodal models through innovative post-training methods and alignment techniques.
Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Join Pika as a Research Intern to work on Videogen models and contribute to generative AI and multimedia technology.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced machine learning models.
Design and optimize infrastructure for large-scale AI model training at a leading generative AI company.
Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.
Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.
Join Cartesia as an Applied Researcher to enhance generative audio models by bridging customer needs with innovative research.
Join Mirage as a Research Engineer to build and scale systems for cutting-edge video generation models in a dynamic AI-focused environment.
Lead research in audio-visual avatar generation at Tavus, shaping the future of human-AI interaction.
Join Iambic Therapeutics as a Software Engineer to develop agentic data pipelines for biomedical data using LLMs.
Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.