"large scale observability" Jobs

1648 open tech roles matching “large scale observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, Go. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 1648 results

Anthropic

Join Anthropic as a Research Engineer to build evaluation instruments for AI development in a remote-friendly environment.

Anthropic Remote-Friendly (Travel Required) | San Francisco, CA $350k–$850k/yr Published 5 days ago
Flexible on stack
Vanta

Lead the migration of Vanta’s data layer to a new architecture while ensuring system reliability and customer integration.

Vanta Remote U.S. Published 2 weeks ago
Flexible on stack
Datadog

Lead a team of engineers to build the Experimentation App at a fast-growing startup.

Datadog New York, New York, USA $192k–$240k/yr Published 2 months ago
AI-first team Heavy meetings
Datadog
Datadog New York, New York, USA $110k–$120k/yr Published 8 months ago
Okta

Join Okta as a Senior Site Reliability Engineer to build reliable cloud services and improve operational excellence.

Okta Bengaluru, India Published 5 days ago
Flexible on stack
Harvey AI

Join Harvey AI as a Senior Software Engineer to build and operate core infrastructure for leading law firms and enterprises.

Harvey AI San Francisco $161.3k–$241.9k/yr Published 1 month ago
Flexible on stack
Fal

Lead data center operations at fal, overseeing strategy, capacity, and build-out for a high-performance GPU fleet.

Fal San Francisco Published 1 month ago
Perplexity AI

Shape the architecture and drive the technical direction of Perplexity's data ecosystem as a senior member of the Data Platform team.

Perplexity AI San Francisco Published 3 months ago
Flexible on stack 60% coding
PlanetScale

Join PlanetScale as a Customer Support Engineer to help developers and businesses effectively use our database platform.

PlanetScale EMEA Published 1 week ago
Datadog
Datadog Boston, Massachusetts, USA; Denver, Colorado, USA; New York, New York, USA $140k–$180k/yr Published 1 year ago
Multiverse IO

Lead the Diagnose & Prescribe team at Multiverse, focusing on AI-native solutions for learner and employer needs.

Multiverse IO London Published 2 months ago
Flexible on stack Heavy meetings
Sprig

Join Sprig as a Senior Platform Engineer to modernize CI/CD processes and enhance developer experience in an AI-driven environment.

Sprig San Francisco, CA $180k–$260k/yr Published 1 week ago
Flexible on stack 70% coding
Ambient

Design and optimize AI infrastructure for real-time intelligence at Ambient.ai, enhancing security through advanced AI models.

Ambient Redwood City Published 2 months ago
Flexible on stack 70% coding
Clockwork Systems

Design and implement low-level systems software for GPU clusters in a pioneering AI infrastructure company.

Clockwork Systems On Site, Palo Alto, California $150k–$230k/yr Published 7 months ago
Flexible on stack
MongoDB

Take ownership of MongoDB's global network infrastructure, designing and deploying network solutions while mentoring engineers.

MongoDB Palo Alto $118k–$231k/yr Published 1 week ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.

baseten San Francisco Published 3 months ago
Flexible on stack
Okta

Join Okta as a Staff Software Reliability Engineer to design and build scalable data platform services in a hybrid work environment.

Okta Toronto, Ontario, Canada CA$160k–CA$220k/yr Published 1 month ago
Flexible on stack 60% coding
Cylake

Join a small team to architect and deliver testing frameworks for next-gen Cyber Security products.

Cylake Sunnyvale $150k–$250k/yr Published 3 days ago
Flexible on stack
Together AI

Build production AI agent systems for one of the world's largest GPU fleets at Together AI.

Together AI San Francisco $250k–$300k/yr Published 2 days ago
Flexible on stack