"gpu monitoring" Jobs

30 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, GPU. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 30 results

Datadog

Contribute to GPU Monitoring capabilities within the Datadog Agent while working on eBPF infrastructure.

Datadog Portugal, Remote Published 2 months ago
Flexible on stack
Datadog

Contribute to GPU Monitoring capabilities within the Datadog Agent while working with eBPF and Linux kernel infrastructure.

Datadog Denmark, Remote; France, Remote; Germany, Remote; Ireland, Remote; Italy, Remote; Poland, Remote; Spain, Remote; Sweden, Remote; Switzerland, Remote; United Kingdom, Remote Published 2 months ago
Flexible on stack
Together AI

Join Together AI as a Senior Software Engineer to design and implement a scalable observability platform for our generative AI lifecycle.

Together AI San Francisco $200k–$280k/yr Published 8 months ago
Flexible on stack
Databricks

Join Databricks as a Senior Software Engineer to build and scale a managed GPU training platform for AI models.

Databricks Mountain View, California; San Francisco, California $160k–$225k/yr Published 1 month ago
Flexible on stack 70% coding
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Starbase, TX Published 2 weeks ago
Flexible on stack
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Hawthorne, CA $165k–$230k/yr Published 2 weeks ago
Flexible on stack
Together AI

Design and maintain high-performance networks as a Senior Network Engineer at Together AI, contributing to cutting-edge AI infrastructure.

Together AI San Francisco $190k–$270k/yr Published 3 weeks ago
Flexible on stack
Together AI

Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.

Together AI San Francisco $220k–$260k/yr Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Austin, TX Published 5 days ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure at SpaceX, focusing on AI training clusters.

SpaceX Palo Alto, CA $185k–$260k/yr Published 5 days ago
Together AI

Join Together AI as an AI Infrastructure Engineer (SRE) to ensure the reliability and scalability of user-facing services.

Together AI Bangalore, India Published 2 weeks ago
Flexible on stack
Together AI

Join Together AI as a senior Infrastructure Design Engineer to lead the design and execution of whitespace environments for AI data centers.

Together AI San Francisco $210k–$250k/yr Published 2 months ago
SpaceX

Build tooling for classified environments as a Senior AI Engineer at SpaceX, ensuring effective operations for engineers in critical settings.

SpaceX Washington, DC $220k–$350k/yr Published 2 months ago