"large scale observability" Jobs
650 open tech roles matching “large scale observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 650 results
Join Anthropic as a Product Manager to lead the development of monetization strategies for AI platforms across various industries.
Join Decagon as a Senior Software Engineer to build and operate core infrastructure for AI-driven customer support solutions.
Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.
Lead the Finance Transformation initiatives for Elastic's Hyperscaler Marketplace business, optimizing financial systems and processes.
Join Supabase as a Platform Engineer to ensure compute capacity meets demand across regions in a fully remote environment.
Join Elastic as a Senior Java Developer to enhance the core infrastructure of Elasticsearch in a fully remote environment.
Join Airbnb's Service Mesh team as a Senior Software Engineer to build and operate critical systems for service-to-service communication.
Lead cross-functional programs to enhance CoreWeave's cloud infrastructure for AI, ensuring scalable and reliable performance.
Provide technical leadership for automated deployments at Cloudflare, impacting the reliability and velocity of its global network.
Join a team building backend services for restaurant point-of-sale systems in a fast-growing online food delivery market.
Join GitLab as a Senior Backend Engineer to enhance Git data storage and improve repository management in a fully remote environment.
Join Commure as a Senior Backend Engineer to build a next-generation ambient AI platform for healthcare.
Join Fin as a Senior Engineer to enhance core technologies, focusing on Elasticsearch and scaling systems for a leading AI Customer Agent.
Lead the Platform Engineering teams at Ethos, shaping architecture and strategy while ensuring system reliability and performance.