"cluster management" Jobs

833 open tech roles matching “cluster management”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Go. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 833 results

baseten

Lead the global GPU fleet management at Baseten, ensuring optimal capacity and uptime for AI workloads.

baseten San Francisco Published 4 days ago
Flexible on stack 70% coding
Yugabyte

Join Yugabyte as a Staff Engineer to design and scale core distributed storage and transaction systems for YugabyteDB.

Yugabyte Sunnyvale, CA $150k–$250k/yr Published 7 months ago
Flexible on stack
Together AI

Design and maintain high-performance networks as a Senior Network Engineer at Together AI, contributing to cutting-edge AI infrastructure.

Together AI San Francisco $190k–$270k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
MongoDB
MongoDB Austin; Boston; Chicago; Denver; Miami; New York City; San Francisco; Seattle; United States $127k–$249k/yr Published 5 months ago
Fal

Build and operate the network infrastructure for fal's generative media ecosystem, focusing on performance and automation.

Fal Remote - USA Published 4 weeks ago
Flexible on stack
Cresta

Join Cresta as a Senior Infrastructure Software Engineer to design and build core infrastructure for AI-driven customer experiences.

Cresta Germany (Remote) Published 1 year ago
Flexible on stack
Horizon3.ai

Join Horizon3.ai as a Staff Technical Support Engineer to lead complex technical issue resolution for enterprise customers in a fully remote environment.

Horizon3.ai US, Remote $117k–$145k/yr Published 1 month ago
Flexible on stack
SpaceX

Join SpaceX as a Global Supply Manager to lead ODM/CM management for compute and network hardware in a fast-paced environment.

SpaceX Austin, TX Published 3 months ago
Fal

Build high-performance compute environments for AI products at fal, focusing on Kubernetes and infrastructure automation.

Fal Remote - USA Published 4 weeks ago
Flexible on stack
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
Perplexity

Join Perplexity's Cloud Infrastructure team to design and operate secure cloud solutions for enterprise customers.

Perplexity San Francisco Published 3 months ago
Flexible on stack
Volta

Join Volta as a Network Modeling/Automation Engineer to build and operate large-scale GPU compute infrastructure for AI workloads.

Volta Palo Alto, CA Published 1 month ago
Flexible on stack
Cresta

Join Cresta as a senior Infrastructure Software Engineer to design and build core infrastructure for AI-driven customer experiences.

Cresta Romania (Remote) Published 1 year ago
Flexible on stack
Fal

Join fal as a Senior Site Reliability Engineer to enhance the reliability of customer-facing systems in a growing generative media ecosystem.

Fal Remote - Global Published 6 months ago
Flexible on stack
Coreweave
Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA $165k–$242k/yr Published 11 months ago
Anthropic

Lead a team to build Anthropic's scheduling platform and improve fleet efficiency in a fast-paced AI environment.

Anthropic San Francisco, CA | New York City, NY $405k–$485k/yr Published 1 week ago
Flexible on stack Heavy meetings
Together AI

Operate and optimize multi-petabyte storage systems for AI workloads at Together AI.

Together AI Bangalore India Published 1 day ago
Flexible on stack