"gpu cloud infrastructure" Jobs
164 open tech roles matching “gpu cloud infrastructure”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 164 results
Lead the infrastructure for AI computing services at CoreWeave, driving innovation and managing high-performing engineering teams.
Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.
Own and improve the qualification process for new compute capacity at Together AI, ensuring technical standards are met.
Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.
Join fal as a Staff Security Engineer to secure cutting-edge AI infrastructure and design systems from first principles.
Lead a team of cloud platform engineers to build scalable and reliable infrastructure for AI products at Baseten.
Join Cursor as a Technical Program Manager to drive infrastructure efficiency and resource allocation in a dynamic startup environment.
Join CoreWeave as a Finance Manager to lead GPU capacity planning and drive financial performance in a fast-growing AI cloud company.
Lead the Compute Platform team at Reflection AI, focusing on multi-cloud scheduling and GPU deployments while mentoring a team of systems engineers.
Join Atoms as a Staff Cluster Infrastructure Engineer to optimize and manage GPU compute clusters for real-world AI applications.
Join Together AI as a Product Manager to drive product work across AI infrastructure with a focus on observability and GPU clusters.
Design and operate data systems that power Decagon's AI products, ensuring high reliability and performance.
Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.
Lead the Runtime Fabric team at Baseten to build container runtimes tailored for AI inference workloads.