"multi gpu" Jobs

574 open tech roles matching “multi gpu”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 574 results

baseten

Lead the capacity management for Baseten's TPU fleet, ensuring optimal performance and reliability in AI workloads.

baseten San Francisco, United States Published 1 day ago
Flexible on stack
Sarvam AI

Join Sarvam AI as a Senior Performance Engineer to optimize GPU kernels for high-performance ML systems.

Sarvam AI Bengaluru Published 1 month ago
Mirelo AI

Join Mirelo AI as a Training Infrastructure Engineer to optimize and design scalable systems for training generative AI models.

Mirelo AI Berlin Published 9 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $500k–$850k/yr Published 1 year ago
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Graphcore

Lead the engineering behind manufacturing test for advanced AI hardware at Graphcore.

Graphcore Bristol, UK; Cambridge, UK Published 1 month ago
Flexible on stack
Sarvam AI

Build and enhance the AI infrastructure platform at Sarvam, focusing on GPU scheduling and multi-tenancy for machine learning workloads.

Sarvam AI Bengaluru Published 2 months ago
70% coding
Coreweave
Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA / San Francisco, CA $139k–$242k/yr Published 6 months ago
krea.ai

Join Krea to build innovative AI tools in a hands-on role focused on supercomputing and distributed systems.

krea.ai San Francisco Published 5 months ago
Flexible on stack
Together AI

Operate and optimize multi-petabyte storage systems for AI workloads at Together AI.

Together AI Bangalore, India Published 1 day ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
Volta

Own end-to-end program delivery for GPU and data center infrastructure at a rapidly growing AI infrastructure company.

Volta Palo Alto, CA Published 1 month ago
Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Mithril

Join Mithril as a Software Engineer to build and maintain backend systems for a cutting-edge AI infrastructure platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack
Coreweave
Coreweave Livingston, NJ / New York, NY / San Francisco, CA / Sunnyvale, CA / Bellevue, WA $83k–$110k/yr Published 1 year ago
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack