"gpu monitoring" Jobs

82 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Grafana. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 82 results

SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Hawthorne, CA $165k–$230k/yr Published 2 months ago
Flexible on stack
Wonderful

Build and own the infrastructure and pipelines for machine learning models in production at a fast-growing startup.

Wonderful Tel Aviv, HQ Published 1 month ago
Flexible on stack
Volta

Join Volta as a Security Engineer focusing on compliance to implement and maintain security controls for ISO 27001 and SOC 2.

Volta Palo Alto, CA Published 1 month ago
Flexible on stack
Together AI

Design and maintain high-performance networks as a Senior Network Engineer at Together AI, contributing to cutting-edge AI infrastructure.

Together AI San Francisco $190k–$270k/yr Published 2 months ago
Flexible on stack
Together AI

Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.

Together AI San Francisco $220k–$260k/yr Published 3 months ago
SpaceX

Lead sourcing strategy and supplier management for server mechanicals and cooling at SpaceX, driving innovation and cost efficiency.

SpaceX Austin, TX Published 1 week ago
Together AI

Join Together AI as a Technical Support Engineer to tackle complex technical challenges in a fast-paced AI environment.

Together AI Remote $160k–$230k/yr Published 1 month ago
Flexible on stack
Applied Intuition

Join Applied Intuition as a Software Engineer to optimize application-layer software for embedded systems in autonomous driving.

Applied Intuition Sunnyvale Published 1 year ago
Fal

Join fal as a Senior Site Reliability Engineer to enhance the reliability of customer-facing systems in a growing generative media ecosystem.

Fal Remote - Global Published 6 months ago
Flexible on stack
Doctronic

Join Doctronic as a DevOps Engineer to own and enhance our infrastructure for AI-driven healthcare solutions.

Doctronic New York City $180k–$240k/yr Published 2 months ago
70% coding
Commure

Join Commure as a Senior Backend Engineer to build a next-generation ambient AI platform for healthcare.

Commure Mountain View, CA Published 3 weeks ago
Flexible on stack
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Austin, TX Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure at SpaceX, focusing on AI training clusters.

SpaceX Palo Alto, CA $185k–$260k/yr Published 1 month ago
Skydio

Join Skydio as a Senior Autonomy Engineer to enhance deep learning infrastructure for autonomous drones.

Skydio San Mateo, California, United States $170k–$277.5k/yr Published 9 months ago
Flexible on stack
Fal

Join fal as a Senior Site Reliability Engineer to enhance the reliability of our generative media infrastructure in San Francisco.

Fal San Francisco Published 6 months ago