"gpu cluster architecture" Jobs
193 open tech roles matching “gpu cluster architecture”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Go. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 193 results
Join Triomics as a Platform Engineer to build backend services and manage cloud infrastructure for processing millions of clinical documents.
Join Fireworks AI as a Member of Technical Staff to design and build systems infrastructure for AI workloads at scale.
Own end-to-end program delivery for GPU and data center infrastructure at a rapidly growing AI infrastructure company.
Lead the Execution Sandbox team at Databricks to architect and launch a new service for non-Spark compute workloads.
Own end-to-end program delivery for GPU and data center infrastructure at a rapidly growing AI infrastructure company.
Lead the global teams operating Graphcore's engineering labs and data center infrastructure while ensuring reliability and efficiency.
Join Anthropic as a Hardware Systems Architect to lead the design and architecture of cutting-edge AI hardware systems.
Lead the security engineering function at Reflection AI, ensuring robust protection for sensitive assets in a fast-paced research environment.
Join Clockwork Systems as a Senior Software Engineer to design high-performance network acceleration software.
Join Atoms as a Staff HPC Network Engineer to design and optimize high-performance networks for GPU compute environments.
Join Clockwork Systems as a Senior Software Engineer to build high-performance network observability platforms for advanced distributed computing.
Lead the platform engineering organization at Volta, building infrastructure for frontier AI deployments with a focus on strategy and execution.
Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.
Lead a high-performing engineering team at Databricks to enhance the AI Runtime product for GPU training infrastructure.
Lead the delivery of on-premises data center builds and GPU cloud programs in a high-growth AI infrastructure company.