"gpu computing" Jobs

206 open tech roles matching “gpu computing”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 206 results

baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
Cursor

Join Cursor as a Technical Program Manager to drive infrastructure efficiency and resource allocation in a dynamic startup environment.

Cursor San Francisco Published 3 months ago
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
Fal

Join fal as a Software Engineer to build large-scale distributed systems for AI products in a growth-focused environment.

Fal San Francisco $180k–$250k/yr Published 1 year ago
Flexible on stack
baseten

Lead the capacity management for Baseten's TPU fleet, ensuring optimal performance and reliability in AI workloads.

baseten San Francisco Published 2 days ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
krea.ai

Join Krea as an ML Researcher to train diffusion models for image and video generation in a creative AI-focused environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 2 days ago
Flexible on stack
Together AI

Own and improve the qualification process for new compute capacity at Together AI, ensuring technical standards are met.

Together AI San Francisco $200k–$250k/yr Published 1 month ago
Flexible on stack
Together AI

Own the cost side of data center build-outs at Together AI, partnering with Infrastructure Strategy to evaluate new opportunities.

Together AI San Francisco $138k–$175k/yr Published 3 weeks ago
Atoms

Join Atoms as a Staff Cluster Infrastructure Engineer to optimize and manage GPU compute clusters for real-world AI applications.

Atoms San Francisco, CA $224k–$284k/yr Published 2 months ago
Flexible on stack
Atoms

Join Atoms as a Staff HPC Network Engineer to design and optimize high-performance networks for GPU compute environments.

Atoms San Francisco, CA $224k–$284k/yr Published 2 months ago
Flexible on stack
baseten

Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.

baseten San Francisco Published 2 days ago
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Coreweave
Coreweave Livingston, NJ / New York, NY / San Francisco, CA / Sunnyvale, CA / Bellevue, WA $83k–$110k/yr Published 1 year ago
Coreweave

Join CoreWeave as a Finance Manager to lead GPU capacity planning and drive financial performance in a fast-growing AI cloud company.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $127k–$168k/yr Published 3 weeks ago
Flexible on stack AI-first team
baseten

Join Baseten as an Infrastructure Software Engineer to build and maintain components of our ML inference platform for AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 1 year ago