"nvidia gpu" Jobs

143 open tech roles matching “nvidia gpu”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 143 results

baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
baseten

Join Baseten as a Software Engineer to lead GPU Networking efforts and optimize distributed systems for AI applications.

baseten San Francisco Published 6 months ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to optimize GPU kernels for high-performance AI applications in a rapidly growing environment.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 2 months ago
70% coding
Fal

Join fal as a Software Engineer to build and maintain systems for managing a large fleet of GPU servers in a growing AI platform.

Fal San Francisco $180k–$250k/yr Published 6 months ago
Flexible on stack 70% coding
Fal

Build high-performance compute environments for AI products at fal, focusing on Kubernetes and infrastructure automation.

Fal Remote - USA Published 4 weeks ago
Flexible on stack
baseten

Join Baseten as a hands-on Operations Manager to optimize the health and utilization of our GPU fleet in a fast-growing AI company.

baseten San Francisco Published 1 week ago
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Fal

Join fal as a hands-on Software Engineer to build and maintain systems for managing a large fleet of GPU servers.

Fal Remote - Global Published 6 months ago
Flexible on stack 70% coding
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
baseten

Lead the global GPU fleet management at Baseten, ensuring optimal capacity and uptime for AI workloads.

baseten San Francisco Published 4 days ago
Flexible on stack 70% coding
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.

Inworld AI UK £140k–£200k/yr Published 5 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve top-ranked realtime voice models.

Inworld AI Mountain View, California, USA $270k–$500k/yr Published 5 months ago
Flexible on stack
Kodiak Robotics

Join Kodiak Robotics as a Senior AI Infrastructure Engineer to optimize model training for autonomous technology.

Kodiak Robotics Mountain View, CA $190k–$260k/yr Published 2 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Staff/Principal Machine Learning Engineer to optimize and serve state-of-the-art voice models in a fully remote role.

Inworld AI Switzerland Published 5 months ago
Flexible on stack
Inworld AI

Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic environment.

Inworld AI Serbia Published 5 months ago
Flexible on stack