"gpu monitoring" Jobs
215 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Terraform. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 215 results
Lead engineering for Cloud Observability at Datadog, managing a team of ~40 engineers in a hybrid work environment.
Join Databricks as a Senior Software Engineer to build and scale a managed GPU training platform for AI models.
Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.
Join SpaceX as a Sr. HPC Systems Engineer to manage HPC clusters and support engineering teams in a fast-paced environment.
Build and own the infrastructure and pipelines for machine learning models in production at a fast-growing startup.
Lead the infrastructure team at Omnifold, focusing on AI model training and deployment in a fast-paced startup environment.
Join Databricks as a Staff Software Engineer to drive the architecture of a managed GPU training platform for large-scale AI models.
Lead and mentor a team of DevOps engineers at an AI company focused on enterprise decision-making.
Join ElevenLabs as an HPC Infrastructure Engineer to optimize GPU clusters for AI research in a fully remote environment.
Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.
Design and optimize infrastructure for large-scale AI model training at a leading generative AI company.
Design and maintain high-performance networks as a Senior Network Engineer at Together AI, contributing to cutting-edge AI infrastructure.
Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.
Join Genesis Molecular AI as a Machine Learning Infrastructure Engineer to build innovative data systems for drug discovery.