Direct from source · No middlemen

Cuda In Jobs

131 open positions · Updated 5 months ago

Average salary (USD/year): 162.8k–226.8k/yr

Showing 20 of 131 positions

Search with filters →
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate infrastructure management for AI clusters.

Together AI London & Amsterdam Published 3 weeks ago
Flexible on stack
Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic environment.

Inworld AI Serbia Published 5 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focused on ML performance to optimize large language models in a fast-paced startup environment.

baseten San Francisco Published 2 years ago
Flexible on stack
Inferact

Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.

Inferact Remote Published 7 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Cowboy Space Corporation

Develop real-time DSP algorithms for space-based communications at Cowboy Space Corporation, a pioneering energy startup.

Cowboy Space Corporation San Carlos, California $167k–$214k/yr Published 1 week ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.

Perplexity AI London Published 5 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $139k–$204k/yr Published 7 months ago
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
ElevenLabs

Join ElevenLabs as an HPC Infrastructure Engineer to optimize GPU clusters for AI research in a fully remote environment.

ElevenLabs United States Published 1 week ago
Flexible on stack
Hippocratic AI

Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.

Hippocratic AI Menlo Park, CA Published 3 weeks ago
Flexible on stack
Tavus

Join Tavus as a Senior Software Engineer to enhance the infrastructure behind real-time AI conversations.

Tavus Remote Published 4 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.

Inworld AI Germany Published 5 months ago
Flexible on stack