Direct from source · No middlemen

Cuda In Jobs

131 open positions · Updated 1 week ago

Average salary (USD/year): 188.3k–327.5k/yr

Showing 20 of 131 positions

Search with filters →
Genmo

Join Genmo as a GPU Performance Engineer to optimize video generation models and achieve significant performance improvements.

Genmo San Francisco HQ Published 1 year ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding
baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Fundamental

Join Fundamental as a Senior Applied Research Engineer to tackle technical challenges in AI model development for enterprise decision-making.

Fundamental Barcelona Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Clockwork Systems

Design and implement low-level systems software for GPU clusters in a pioneering AI infrastructure company.

Clockwork Systems On Site, Palo Alto, California $150k–$230k/yr Published 7 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago
Pika

Join Pika as a Senior/Staff ML Engineer to enhance AI-driven products through advanced inference acceleration and GPU optimization.

Pika Palo Alto HQ Published 2 months ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $92k–$135k/yr Published 10 months ago
Coreweave

Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.

Coreweave Sunnyvale, CA / Bellevue, WA $188k–$275k/yr Published 4 months ago
Flexible on stack
Sarvam AI

Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.

Sarvam AI Bengaluru Published 1 month ago
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate GPU infrastructure management.

Together AI San Francisco $240k–$280k/yr Published 1 month ago
Flexible on stack
Page 1 of 7 Next →