"large scale observability" Jobs
261 open tech roles matching “large scale observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 261 results
Join Braze as a Senior Site Reliability Engineer to enhance MongoDB infrastructure and improve developer experience in a hybrid work environment.
Own the system architecture for a startup transforming engineering with AI, working directly with civil engineers.
Join MongoDB's SLS team as a Senior Software Engineer to build scalable cloud storage services for petabytes of data.
Lead the Capacity Engineering team at Anthropic, ensuring efficient allocation and utilization of infrastructure resources.
Join Profound as a Software Engineer to design and scale the data platform infrastructure for AI-driven marketing solutions.
Join Reflection AI as a hands-on technical staff member to enhance model performance through data-driven evaluations and feedback loops.
Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.
Join Datadog as a Senior Software Engineer to enhance a petabyte-scale distributed storage system for the AI era.
Join Datadog as a Software Engineering Intern to tackle real-world engineering challenges in a hybrid work environment.
Lead the GTM Enablement Operations at Datadog, building operational foundations for scalable go-to-market execution.
Join Datadog as a Developer Advocate to leverage your engineering and storytelling skills in service management and incident response.
Join Anthropic as a Staff Software Engineer to enhance deployment infrastructure for AI systems in a collaborative environment.
Lead the development of data ingestion and normalization for advertising platforms at Databricks.
Join CoreWeave as a Senior Software Engineer to build software for managing large-scale GPU data center infrastructure.
Join Reflection AI's Compute Platform team to enhance multi-cloud scheduling and GPU infrastructure in a mission-driven environment.
Join MongoDB as a Senior Site Reliability Engineer to enhance our continuous deployment infrastructure and support engineering teams.
Join Traversal as an AI Platform Engineer to build scalable systems that enhance AI performance in complex production environments.