"gpu monitoring" Jobs

65 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Terraform. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 65 results

Fal

Join fal as a Software Engineer to build and maintain systems for managing a large fleet of GPU servers in a growing AI platform.

Fal San Francisco $180k–$250k/yr Published 6 months ago
Flexible on stack 70% coding
baseten

Join Baseten as a hands-on Operations Manager to optimize the health and utilization of our GPU fleet in a fast-growing AI company.

baseten San Francisco Published 1 week ago
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to build next-generation observability systems for large-scale AI infrastructure.

Anthropic London, UK £325k–£390k/yr Published 1 week ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
krea.ai

Join Krea as an ML Researcher to train diffusion models for image and video generation in a creative AI-focused environment.

krea.ai San Francisco Published 1 week ago
Flexible on stack
Together AI

Join Together AI as a Senior Software Engineer to design and implement a scalable observability platform for our generative AI lifecycle.

Together AI San Francisco $200k–$280k/yr Published 10 months ago
Flexible on stack
Reflection AI

Lead the Compute Platform team at Reflection AI, focusing on multi-cloud scheduling and GPU deployments while mentoring a team of systems engineers.

Reflection AI San Francisco, CA Published 1 month ago
Flexible on stack
Meter

Lead the development and market direction of Meter's first data center switching line, collaborating closely with hardware and engineering teams.

Meter San Francisco $205k–$270k/yr Published 5 days ago
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Omnifold

Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.

Omnifold San Francisco HQ Published 6 months ago
Flexible on stack
baseten

Join Baseten as an Infrastructure Software Engineer to build and maintain components of our ML inference platform for AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack
Cursor

Join Cursor as a Software Engineer on the ML Platform to build infrastructure that enhances machine learning models and supports product engineers.

Cursor San Francisco Published 1 week ago
Flexible on stack
Postman

Lead AI reliability engineering efforts to ensure the performance and scalability of Postman's AI-powered API services.

Postman San Francisco, California, United States $256k–$276k/yr Published 11 months ago
Neuralink

Join Neuralink as a Network & Systems Engineer to design and operate cutting-edge data center environments.

Neuralink South San Francisco, California, United States $126k–$190.8k/yr Published 4 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack
baseten

Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.

baseten San Francisco Published 11 months ago
Flexible on stack