"gpu monitoring" Jobs

215 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Terraform. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 215 results

Exa

Join Exa as a Security Engineer to build protective systems for massive-scale AI infrastructure in San Francisco.

Exa San Francisco, California Published 2 weeks ago
AI-first team
SpaceX

Lead sourcing strategy and supplier management for server mechanicals and cooling at SpaceX, driving innovation and cost efficiency.

SpaceX Austin, TX Published 1 week ago
Together AI

Join Together AI as a Technical Support Engineer to tackle complex technical challenges in a fast-paced AI environment.

Together AI Remote $160k–$230k/yr Published 1 month ago
Flexible on stack
Databricks

Develop and run the research stack that powers Databricks AI Research, enabling rapid large-scale experiments.

Databricks New York City, New York; San Francisco, California $199k–$270k/yr Published 4 months ago
Flexible on stack
Databricks
Databricks New York City, New York; San Francisco, California $190k–$270k/yr Published 4 months ago
Applied Intuition

Join Applied Intuition as a Software Engineer to optimize application-layer software for embedded systems in autonomous driving.

Applied Intuition Sunnyvale Published 1 year ago
Mithril

Join Mithril as a Software Engineer to build and maintain backend systems for a cutting-edge AI infrastructure platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack
Fal

Join fal as a Senior Site Reliability Engineer to enhance the reliability of customer-facing systems in a growing generative media ecosystem.

Fal Remote - Global Published 6 months ago
Flexible on stack
Doctronic

Join Doctronic as a DevOps Engineer to own and enhance our infrastructure for AI-driven healthcare solutions.

Doctronic New York City $180k–$240k/yr Published 2 months ago
70% coding
Commure

Join Commure as a Senior Backend Engineer to build a next-generation ambient AI platform for healthcare.

Commure Mountain View, CA Published 3 weeks ago
Flexible on stack
Valar Atomics

Lead the technical architecture of Valar's software platform, focusing on telemetry, streaming, and simulation systems in a high-impact startup.

Valar Atomics Torrance, California, United States $195k–$270k/yr Published 2 weeks ago
Flexible on stack
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Austin, TX Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure at SpaceX, focusing on AI training clusters.

SpaceX Palo Alto, CA $185k–$260k/yr Published 1 month ago
Inferact

Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack AI-first team
Inferact

Join Inferact as a cloud orchestration engineer to build reliable systems for AI model deployment at scale.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Skydio

Join Skydio as a Senior Autonomy Engineer to enhance deep learning infrastructure for autonomous drones.

Skydio San Mateo, California, United States $170k–$277.5k/yr Published 9 months ago
Flexible on stack