"gpu monitoring" Jobs

72 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 72 results

Datadog

Contribute to GPU Monitoring capabilities within the Datadog Agent while working on eBPF infrastructure.

Datadog Portugal, Remote Published 2 months ago
Flexible on stack
Datadog

Contribute to GPU Monitoring capabilities within the Datadog Agent while working with eBPF and Linux kernel infrastructure.

Datadog Denmark, Remote; France, Remote; Germany, Remote; Ireland, Remote; Italy, Remote; Poland, Remote; Spain, Remote; Sweden, Remote; Switzerland, Remote; United Kingdom, Remote Published 2 months ago
Flexible on stack
Datadog

Own the go-to-market strategy for Infrastructure Monitoring for AI and containerized workloads at Datadog.

Datadog New York, New York, USA $123k–$164k/yr Published 5 days ago
Together AI

Join Together AI as a Senior Software Engineer to design and implement a scalable observability platform for our generative AI lifecycle.

Together AI San Francisco $200k–$280k/yr Published 8 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
Reddit

Join Reddit as a Staff Machine Learning Engineer to enhance ML efficiency and drive performance improvements across the platform.

Reddit Remote - The Netherlands Published 1 month ago
Flexible on stack
Reddit

Join Reddit as a Staff Machine Learning Engineer to enhance ML efficiency and drive performance improvements across the company's ML ecosystem.

Reddit Remote - United Kingdom Published 1 month ago
Flexible on stack
Datadog

Lead engineering for Cloud Observability at Datadog, managing a team of ~40 engineers in a hybrid work environment.

Datadog Boston, Massachusetts, USA; New York, New York, USA $296k–$370k/yr Published 1 week ago
Databricks

Join Databricks as a Senior Software Engineer to build and scale a managed GPU training platform for AI models.

Databricks Mountain View, California; San Francisco, California $160k–$225k/yr Published 1 month ago
Flexible on stack 70% coding
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Hawthorne, CA $165k–$230k/yr Published 2 weeks ago
Flexible on stack
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Starbase, TX Published 2 weeks ago
Flexible on stack