"gpu" Jobs

288 open tech roles matching “gpu”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 288 results

Together AI

Join Together AI as a Staff Software Engineer to build systems that automate GPU infrastructure management.

Together AI San Francisco $240k–$280k/yr Published 1 month ago
Flexible on stack
World Labs

Join World Labs as a Performance Engineer to optimize AI models for speed and efficiency in a cutting-edge research environment.

World Labs San Francisco $200k–$300k/yr Published 4 months ago
Flexible on stack 70% coding
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
baseten

Join Baseten as a Software Engineer focused on ML performance to optimize large language models in a fast-paced startup environment.

baseten San Francisco Published 2 years ago
Flexible on stack
baseten

Join Baseten as an Infrastructure Software Engineer to build and maintain components of our ML inference platform for AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Databricks

Join Databricks as a Senior Software Engineer to build and scale a managed GPU training platform for AI models.

Databricks Mountain View, California; San Francisco, California $160k–$225k/yr Published 3 months ago
Flexible on stack 70% coding
Genmo

Join Genmo as a Research Engineer to advance visual generative AI in a fast-paced startup environment.

Genmo San Francisco HQ Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.

Anthropic San Francisco, CA | New York City, NY $350k–$850k/yr Published 4 days ago
Flexible on stack
baseten

Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.

baseten San Francisco Published 3 days ago
baseten

Lead the delivery of on-premises data center builds and GPU cloud programs in a high-growth AI infrastructure company.

baseten San Francisco Published 1 week ago
baseten

Join Baseten as an AI Engineer to design and automate workflows for AI-powered capacity management in a fast-growing company.

baseten San Francisco Published 3 days ago
Flexible on stack
baseten

Join Baseten as a Post-Training Research Engineer to build in-house tooling for efficient and high-quality machine learning models.

baseten San Francisco Published 5 months ago
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Databricks

Lead a high-performing engineering team at Databricks to enhance the AI Runtime product for GPU training infrastructure.

Databricks Mountain View, California; San Francisco, California $228.6k–$297.1k/yr Published 2 months ago
Flexible on stack Heavy meetings