"production observability" Jobs

913 open tech roles matching “production observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 913 results

Anthropic

Join Anthropic as a Staff Software Engineer to build next-generation observability systems for large-scale AI infrastructure.

Anthropic London, UK £325k–£390k/yr Published 1 week ago
Flexible on stack
baseten

Join Baseten as a Software Engineer on the Observability team to enhance the reliability of AI product systems.

baseten San Francisco Published 1 month ago
Flexible on stack
Postman

Join Postman as a Staff Engineer to enhance observability capabilities and drive platform-wide improvements in a collaborative environment.

Postman Bengaluru, Karnataka, India Published 2 months ago
Flexible on stack
Pinterest

Join Pinterest as a Staff Software Engineer to lead the development of observability solutions for large-scale distributed systems.

Pinterest San Francisco, CA, US; Remote, US $177.2k–$364.8k/yr Published 3 months ago
Flexible on stack
Vercel

Join Vercel as a Software Engineer on the Observability team to enhance application monitoring and developer experience.

Vercel Hybrid - San Francisco, New York City, London Published 1 year ago
Flexible on stack
Figma
Figma San Francisco, CA • New York, NY • United States $258k–$376k/yr Published 6 months ago
MaintainX

Own the Generation half of Document Intelligence, transforming multimodal inputs into structured maintenance knowledge.

MaintainX San Francisco Published 1 week ago
Flexible on stack
MaintainX

Join MaintainX as a Site Reliability Engineer to enhance platform reliability and developer autonomy in a collaborative environment.

MaintainX San Francisco Published 5 days ago
Flexible on stack
Meter

Join Meter as a Product Manager to lead the development of their wireless network solutions and enhance customer experience.

Meter San Francisco $205k–$275k/yr Published 2 weeks ago
HappyRobot

Join HappyRobot as a Site Reliability Engineer to enhance operational resilience and improve system uptime in a high-impact role.

HappyRobot San Francisco Published 3 months ago
Doss

Design and build scalable backend systems for an AI-native operations platform in a collaborative team environment.

Doss San Francisco Published 1 year ago
Datadog
Datadog Boston, Massachusetts, USA; Denver, Colorado, USA; New York, New York, USA; San Francisco, California, USA $151.5k–$222k/yr Published 6 months ago
Doss

Join Doss as a Backend Engineer to design and build scalable data models and tooling for an AI-native operations platform.

Doss San Francisco Published 1 year ago
Flexible on stack
Together AI

Join Together AI as a Product Manager to drive product work across AI infrastructure with a focus on observability and GPU clusters.

Together AI San Francisco $175k–$220k/yr Published 2 months ago
PlanetScale

Join PlanetScale as a Software Engineer to build a customer-facing database observability product for Vitess and PostgreSQL.

PlanetScale San Francisco Bay Area or Remote $120k–$290k/yr Published 8 months ago
Flexible on stack AI-first team
Braintrust Data

Join Braintrust as a product engineer to build user-friendly tools for AI observability in a startup environment.

Braintrust Data San Francisco Published 2 years ago
Flexible on stack
Harvey AI

Join Harvey AI as a Staff Software Engineer to build and operate core infrastructure for AI workloads in a fast-growing company.

Harvey AI San Francisco $231k–$340k/yr Published 1 month ago
Flexible on stack
Doss

Join Doss as a Fullstack Engineer to build an AI-native platform for physical product businesses in a collaborative San Francisco team.

Doss San Francisco Published 2 years ago
Flexible on stack
MaintainX

Lead a team of developers to deliver reliable software in a hybrid work environment at MaintainX.

MaintainX San Francisco Published 3 days ago
Heavy meetings
baseten

Join Baseten as a hands-on Operations Manager to optimize the health and utilization of our GPU fleet in a fast-growing AI company.

baseten San Francisco Published 1 week ago