Direct from source · No middlemen
131 open positions · Updated 1 week ago
Showing 20 of 131 positions
Search with filters →Join Genmo as a GPU Performance Engineer to optimize video generation models and achieve significant performance improvements.
Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.
Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.
Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.
Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.
Join Fundamental as a Senior Applied Research Engineer to tackle technical challenges in AI model development for enterprise decision-making.
Design and implement low-level systems software for GPU clusters in a pioneering AI infrastructure company.
Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.
Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.
Join Pika as a Senior/Staff ML Engineer to enhance AI-driven products through advanced inference acceleration and GPU optimization.
Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.
Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.
Join Together AI as a Staff Software Engineer to build systems that automate GPU infrastructure management.