"gpu monitoring" Jobs

30 open tech roles matching “gpu monitoring”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Prometheus, Grafana. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 30 results

Coreweave

Join CoreWeave as a Senior Software Engineer to enhance network observability for GPU cloud services in a fast-growing company.

Coreweave Sunnyvale, CA / New York City, NY / Livingston, NJ $153k–$204k/yr Published 2 weeks ago
Flexible on stack
Reflection AI

Join Reflection AI's Compute Platform team to enhance multi-cloud scheduling and GPU infrastructure in a mission-driven environment.

Reflection AI New York, NY Published 5 months ago
Triomics

Join Triomics as a Platform Engineer to build backend services and manage cloud infrastructure for processing millions of clinical documents.

Triomics New York Office Published 2 months ago
Flexible on stack AI-first team
Coreweave

Join CoreWeave as a Senior Software Engineer to build software for managing large-scale GPU data center infrastructure.

Coreweave New York, NY / Sunnyvale, CA $153k–$242k/yr Published 3 months ago
Coreweave

Join CoreWeave as a Staff Software Engineer to build and operate Go-based services for large-scale GPU data center infrastructure.

Coreweave New York, NY / Sunnyvale, CA $207k–$275k/yr Published 3 weeks ago
Coreweave
Coreweave Plano, TX / Washington, DC / Livingston, NJ $83k–$110k/yr Published 9 months ago
Coreweave
Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA $110k–$176k/yr Published 1 year ago
Nomic AI

Join Nomic AI as a Senior Platform Engineer to own and scale our infrastructure stack in the AEC industry.

Nomic AI New York HQ Published 2 months ago
Flexible on stack
Datadog

Lead engineering for Cloud Observability at Datadog, managing a team of ~40 engineers in a hybrid work environment.

Datadog Boston, Massachusetts, USA; New York, New York, USA $296k–$370k/yr Published 1 month ago
Databricks

Develop and run the research stack that powers Databricks AI Research, enabling rapid large-scale experiments.

Databricks New York City, New York; San Francisco, California $199k–$270k/yr Published 4 months ago
Flexible on stack
Databricks
Databricks New York City, New York; San Francisco, California $190k–$270k/yr Published 4 months ago
Doctronic

Join Doctronic as a DevOps Engineer to own and enhance our infrastructure for AI-driven healthcare solutions.

Doctronic New York City $180k–$240k/yr Published 2 months ago
70% coding
Mirage

Join Mirage as a Research Engineer to build and scale systems for cutting-edge video generation models in a dynamic AI-focused environment.

Mirage Union Square, New York City Published 3 weeks ago
Flexible on stack
Reflection AI

Join Reflection AI as a Research Software Engineer to bridge research and production in cutting-edge AI training systems.

Reflection AI New York, NY Published 6 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff+ Software Engineer to build production systems for capacity engineering in a hybrid work environment.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 2 months ago
Flexible on stack 70% coding
Coreweave
Coreweave Livingston, NJ / New York, NY $109k–$160k/yr Published 6 months ago
Coreweave

Join CoreWeave as a Staff Software Engineer to design and develop automation platforms for one of the largest GPU clouds in the world.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA $207k–$275k/yr Published 2 months ago
Flexible on stack 60% coding
Datadog

Join Datadog as a Manager of Strategic Sourcing to lead AI spend and vendor negotiations in a hybrid work environment.

Datadog New York, New York, USA $128k–$187k/yr Published 2 months ago