"gpu programming" Jobs

447 open tech roles matching “gpu programming”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 447 results

Genmo

Join Genmo as a GPU Performance Engineer to optimize video generation models and achieve significant performance improvements.

Genmo San Francisco HQ Published 1 year ago
Flexible on stack
baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Pika

Join Pika as a Senior/Staff ML Engineer to enhance AI-driven products through advanced inference acceleration and GPU optimization.

Pika Palo Alto HQ Published 2 months ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
70% coding
Fundamental

Join Fundamental as a Senior Applied Research Engineer to tackle technical challenges in AI model development for enterprise decision-making.

Fundamental Barcelona Published 5 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago
Kodiak Robotics

Join Kodiak Robotics as a Senior AI Infrastructure Engineer to optimize model training for autonomous technology.

Kodiak Robotics Mountain View, CA $190k–$260k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Cursor

Join Cursor as a Technical Program Manager to drive infrastructure efficiency and resource allocation in a dynamic startup environment.

Cursor San Francisco Published 3 months ago
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Preference Model

Join Preference Model as a Machine Learning Engineer to develop low-level reinforcement learning environments in a fast-paced startup.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Software Engineer to design and build secure, high-performance environments for AI-driven workloads.

Coreweave Livingston, NJ / Sunnyvale, CA / Bellevue, WA $153k–$204k/yr Published 1 week ago
Flexible on stack
Genmo

Join Genmo as a Research Engineer to advance visual generative AI in a fast-paced startup environment.

Genmo San Francisco HQ Published 1 month ago
Flexible on stack
Graphcore

Lead cross-functional programs for next-generation AI networking infrastructure at Graphcore.

Graphcore Milpitas, California, United States Published 1 week ago
World Labs

Join World Labs as a Performance Engineer to optimize AI models for speed and efficiency in a cutting-edge research environment.

World Labs San Francisco $200k–$300k/yr Published 4 months ago
Flexible on stack 70% coding