"multimodal modeling" Jobs

724 open tech roles matching “multimodal modeling”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 724 results

Kodiak Robotics

Join Kodiak Robotics as a Senior AI Infrastructure Engineer to optimize model training for autonomous technology.

Kodiak Robotics Mountain View, CA $190k–$260k/yr Published 2 months ago
Flexible on stack
Cantina

Join Cantina as a Member of Technical Staff to build and scale data pipelines for large video generation models.

Cantina Remote (U.S. or Europe) $200k–$260k/yr Published 5 months ago
Flexible on stack
Hedra

Join Hedra as a Research Engineer to lead the development of action-conditioned world models in a pioneering Physical AI team.

Hedra San Francisco Published 5 months ago
Flexible on stack
Mirage

Join Mirage as a Research Engineer to advance agentic systems for creative tasks in an AI-native video platform.

Mirage Union Square, New York City Published 1 week ago
Ideogram

Join Ideogram as an Applied ML Engineer to bridge research and product, turning generative models into production features.

Ideogram Toronto Published 2 months ago
Flexible on stack
Sesame

Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.

Sesame San Francisco Published 2 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 5 months ago
Flexible on stack
Mirage

Join Mirage as a Research Engineer to build and scale systems for cutting-edge video generation models in a dynamic AI-focused environment.

Mirage Union Square, New York City Published 3 weeks ago
Flexible on stack
Cantina

Join Cantina as a Machine Learning Intern to work on advanced video generation models in a hands-on research environment.

Cantina Singapore Published 5 days ago
Flexible on stack
Pika

Join Pika as a Research Scientist, Data to architect and scale data engineering systems for multimodal foundation models.

Pika Palo Alto HQ Published 2 months ago
Flexible on stack
Cantina

Join Cantina as a Machine Learning Engineer to build advanced speech systems and contribute to innovative AI technology.

Cantina Remote (U.S. or Europe) $200k–$220k/yr Published 1 month ago
Flexible on stack 70% coding
Reddit

Join Reddit as a Senior ML Engineer to develop advanced embedding models for advertising in a fully remote role.

Reddit Remote - The Netherlands Published 2 months ago
Flexible on stack 70% coding
Reddit

Join Reddit as a Senior ML Engineer to develop advanced embedding models for advertising in a fully remote role.

Reddit Remote - United Kingdom Published 2 months ago
Flexible on stack 70% coding
Cartesia

Build the platform behind Managed Agents, shaping the latency, reliability, and developer experience of real-time voice agents.

Cartesia *HQ - San Francisco, CA Published 4 months ago
AI-first team
Cartesia

Join Cartesia as a Software Engineer to design and build scalable AI model inference systems in a collaborative, in-office environment.

Cartesia *HQ - San Francisco, CA Published 2 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.

Inworld AI UK £140k–£200k/yr Published 5 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 3 years ago
Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.

Inworld AI Germany Published 5 months ago
Flexible on stack
Twelve Labs

Join Twelve Labs as a Senior AI Engineer to build the integration layer for their multimodal video AI system.

Twelve Labs San Francisco Published 1 month ago
Flexible on stack