"gpu monitoring" Jobs

215 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Terraform. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 215 results

Volta

Join Volta as an IT Analyst to support a fast-growing GPU compute infrastructure company with hands-on IT operations.

Volta Palo Alto, CA Published 1 month ago
Datadog

Lead engineering for Cloud Observability at Datadog, managing a team of ~40 engineers in a hybrid work environment.

Datadog Boston, Massachusetts, USA; New York, New York, USA $296k–$370k/yr Published 1 month ago
Databricks

Join Databricks as a Senior Software Engineer to build and scale a managed GPU training platform for AI models.

Databricks Mountain View, California; San Francisco, California $160k–$225k/yr Published 3 months ago
Flexible on stack 70% coding
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Hawthorne, CA $165k–$230k/yr Published 2 months ago
Flexible on stack
SpaceX

Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.

SpaceX Starbase, TX Published 2 months ago
Flexible on stack
Wonderful

Build and own the infrastructure and pipelines for machine learning models in production at a fast-growing startup.

Wonderful Tel Aviv, HQ Published 1 month ago
Flexible on stack
Omnifold

Lead the infrastructure team at Omnifold, focusing on AI model training and deployment in a fast-paced startup environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
Databricks

Join Databricks as a Staff Software Engineer to drive the architecture of a managed GPU training platform for large-scale AI models.

Databricks Mountain View, California; San Francisco, California $190k–$265k/yr Published 3 months ago
Fundamental

Lead and mentor a team of DevOps engineers at an AI company focused on enterprise decision-making.

Fundamental Europe Published 7 months ago
Flexible on stack AI-first team
Volta

Join Volta as a Security Engineer focusing on compliance to implement and maintain security controls for ISO 27001 and SOC 2.

Volta Palo Alto, CA Published 1 month ago
Flexible on stack
ElevenLabs

Join ElevenLabs as an HPC Infrastructure Engineer to optimize GPU clusters for AI research in a fully remote environment.

ElevenLabs United States Published 1 week ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.

baseten San Francisco Published 3 months ago
Flexible on stack
Fireworks AI

Design and optimize infrastructure for large-scale AI model training at a leading generative AI company.

Fireworks AI San Mateo Published 1 month ago
Flexible on stack
SpaceX

Lead and manage a sourcing team for datacenter compute infrastructure to support SpaceXAI's AI training clusters.

SpaceX Austin, TX Published 1 month ago
SpaceX

Lead and manage the sourcing team for datacenter compute infrastructure to support SpaceXAI's large-scale AI training clusters.

SpaceX Palo Alto, CA $155k–$215k/yr Published 1 month ago
Together AI

Design and maintain high-performance networks as a Senior Network Engineer at Together AI, contributing to cutting-edge AI infrastructure.

Together AI San Francisco $190k–$270k/yr Published 2 months ago
Flexible on stack
Together AI

Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.

Together AI San Francisco $220k–$260k/yr Published 3 months ago
Genesis Molecular AI

Join Genesis Molecular AI as a Machine Learning Infrastructure Engineer to build innovative data systems for drug discovery.

Genesis Molecular AI San Mateo, CA Published 9 months ago
Flexible on stack