"gpu kernels" Jobs

44 open tech roles matching “gpu kernels”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, CUDA, GPU. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 44 results

baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Preference Model

Join Preference Model as a Machine Learning Engineer to develop low-level reinforcement learning environments in a fast-paced startup.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Genmo

Join Genmo as a GPU Performance Engineer to optimize video generation models and achieve significant performance improvements.

Genmo San Francisco HQ Published 1 year ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 1 year ago
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to lead GPU Networking efforts and optimize distributed systems for AI applications.

baseten San Francisco Published 6 months ago
Flexible on stack
Fal

Join fal as a Software Engineer to build and maintain systems for managing a large fleet of GPU servers in a growing AI platform.

Fal San Francisco $180k–$250k/yr Published 6 months ago
Flexible on stack 70% coding
krea.ai

Join Krea as an ML Researcher to train diffusion models for image and video generation in a creative AI-focused environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago
World Labs

Join World Labs as a Performance Engineer to optimize AI models for speed and efficiency in a cutting-edge research environment.

World Labs San Francisco $200k–$300k/yr Published 4 months ago
Flexible on stack 70% coding
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.

baseten San Francisco Published 11 months ago
Kodiak Robotics

Join Kodiak Robotics as a Staff Machine Learning Engineer to design and deploy ML systems for autonomous trucking.

Kodiak Robotics San Francisco Bay Area $200k–$265k/yr Published 4 months ago
Flexible on stack
baseten

Join Baseten as a Post-Training Research Engineer to build in-house tooling for efficient and high-quality machine learning models.

baseten San Francisco Published 5 months ago