"gpu kernels" Jobs
44 open tech roles matching “gpu kernels”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, CUDA, GPU. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 44 results
Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.
Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.
Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.
Join CoreWeave as a Staff Engineer to design and implement the backend architecture for the innovative molab notebook service.
Join Chai Discovery as an AI Research Engineer to advance AI drug design with a team of experts.
Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.
Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.
Join fal as a Senior Site Reliability Engineer to enhance the reliability of our generative media infrastructure in San Francisco.
Join Xaira Therapeutics as a Senior/Staff AI Research Engineer to develop AI models that accelerate drug discovery and improve human health.
Join Together AI as a Research Engineer to optimize large-scale training infrastructure for cutting-edge AI models.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.