"large scale observability" Jobs

1639 open tech roles matching “large scale observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, Go. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 1639 results

Reflection AI

Join Reflection AI as a staff engineer to build and operate a foundational platform that accelerates engineering teams.

Reflection AI New York, NY Published 6 months ago
Acceldata

Lead SDET for the Pulse team, focusing on quality and performance of large-scale data systems through test automation and collaboration.

Acceldata Bengaluru Published 1 year ago
Descript

Join Descript as a Senior Software Engineer to own and enhance our infrastructure platform, impacting AI model training and deployment.

Descript San Francisco, CA or Remote, US $220k–$292k/yr Published 1 week ago
Flexible on stack
Rox

Join Rox as a Core Engineer to design and operate foundational infrastructure for autonomous revenue agents in a fast-growing AI company.

Rox San Francisco Published 4 months ago
Cockroach Labs

Lead the Production Orchestration team at Cockroach Labs, focusing on reliability, scalability, and operational excellence.

Cockroach Labs New York, NY $194k–$257.3k/yr Published 3 months ago
Flexible on stack Heavy meetings
Orkes

Lead the Cloud & Platform engineering function at Orkes, focusing on distributed systems and team mentorship.

Orkes Cupertino, CA $250k–$300k/yr Published 4 months ago
Heavy meetings
MongoDB

Lead a team to enhance MongoDB's internal developer platform, focusing on automation and collaboration across engineering teams.

MongoDB Gurugram Published 1 month ago
Flexible on stack
Grafana Labs

Join Grafana Labs as a Staff Backend Engineer to influence the roadmap and build scalable alerting systems in a fully remote environment.

Grafana Labs United Kingdom (Remote) £104.0k–£124.8k/yr Published 2 months ago
Flexible on stack
Braintrust Data

Join Braintrust as a backend engineer to build infrastructure for cutting-edge AI development tools in a fast-paced environment.

Braintrust Data San Francisco Published 1 month ago
Flexible on stack
Sierra

Join Sierra as a Software Engineer on the Site Reliability team to enhance the reliability and scalability of AI-driven infrastructure.

Sierra San Francisco, CA Published 10 months ago
Flexible on stack
Legora

Lead the founding SRE team at Legora's NYC hub, driving reliability and infrastructure architecture across multiple teams.

Legora New York City Published 1 week ago
AI-first team
Fal

Join fal as a Software Engineer to build large-scale distributed systems for AI products in a growth-focused environment.

Fal San Francisco $180k–$250k/yr Published 1 year ago
Flexible on stack
Fal

Join fal as a senior software engineer to build large-scale distributed systems for AI products.

Fal Remote - Global Published 6 months ago
Arize AI

Join Arize AI as a Senior AI Product Engineer to build scalable backend systems for ML observability in a hybrid work environment.

Arize AI Remote (United States) $125k–$225k/yr Published 1 year ago
Flexible on stack 70% coding
LaunchDarkly

Lead the Experimentation pillar at LaunchDarkly to shape the future of AI-driven experimentation platforms.

LaunchDarkly Remote - US $301k–$414k/yr Published 1 month ago
Grafana Labs

Join Grafana Labs as a Staff Backend Engineer to influence the roadmap and build scalable alerting systems in a fully remote team.

Grafana Labs Germany (Remote) €109.7k–€131.7k/yr Published 2 months ago
Flexible on stack
Grafana Labs

Join Grafana Labs as a Staff Backend Engineer to drive technical strategy and lead cross-functional projects in a fully remote environment.

Grafana Labs United States (Remote) $175.0k–$210.0k/yr Published 2 months ago
Flexible on stack
Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago