"gpu cluster architecture" Jobs

193 open tech roles matching “gpu cluster architecture”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Go. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 193 results

ChipAgents

Join ChipAgents as an ML Systems Engineer to optimize LLM inference systems for leading semiconductor companies.

ChipAgents San Jose $150k–$350k/yr Published 3 months ago
Flexible on stack
Reflection AI

Lead the Compute Platform team at Reflection AI, focusing on multi-cloud scheduling and GPU deployments while mentoring a team of systems engineers.

Reflection AI San Francisco, CA Published 4 weeks ago
Flexible on stack
Coreweave
Coreweave Bellevue, WA / Sunnyvale, CA $185k–$275k/yr Published 6 months ago
Graphcore

Lead cross-functional programs for next-generation AI networking infrastructure at Graphcore.

Graphcore Milpitas, California, United States Published 1 week ago
baseten

Lead the capacity management for Baseten's TPU fleet, ensuring optimal performance and reliability in AI workloads.

baseten San Francisco, United States Published 2 days ago
Flexible on stack
Together AI

Own and improve the qualification process for new compute capacity at Together AI, ensuring technical standards are met.

Together AI San Francisco $200k–$250k/yr Published 1 month ago
Flexible on stack
Coreweave

Design and build scalable storage systems for AI workloads at a rapidly growing public cloud company.

Coreweave New York $207k–$303k/yr Published 1 month ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
Flexible on stack
SpaceX

Lead and manage a sourcing team for datacenter compute infrastructure to support SpaceXAI's AI training clusters.

SpaceX Austin, TX Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Palo Alto, CA $155k–$215k/yr Published 1 month ago
Coreweave

Design and build high-performance storage systems for AI workloads at a rapidly growing public cloud company.

Coreweave California $207k–$303k/yr Published 1 month ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$405k/yr Published 4 months ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago
Volta

Join Volta as a Senior Network Engineer to design and operate large-scale GPU compute infrastructure for AI workloads.

Volta Palo Alto, CA Published 1 month ago
Flexible on stack
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Austin, TX Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure at SpaceX, focusing on AI training clusters.

SpaceX Palo Alto, CA $185k–$260k/yr Published 1 month ago