"observability systems" Jobs
3871 open tech roles matching “observability systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 3871 results
Drive the strategy and execution for telemetry collection and remote management capabilities as a Principal Product Manager at Elastic.
Join Grafana Labs as an Associate Observability Architect to lead technical partnerships and drive customer success in a fully remote environment.
Join Obsidian Security as an AI Security Engineer to mitigate risks across AI platforms and protect enterprise customers.
Join GitLab as a Staff SRE to design and implement observability and anomaly detection for monetization systems in a fully remote environment.
Join Vercel as a Product Manager to lead the strategy for observability products in a hybrid role based in New York.
Join Mind Robotics as a Software Engineer to build and productionize edge software for data collection rigs in real factories.
Join Apex as an Electro-Optical Engineer to lead sensor product development from concept to production in a fast-growing aerospace startup.
Join Orkes as a Site Reliability Engineer to enhance the reliability and performance of cloud-based production systems.
Drive the product vision for Journey Monitoring at Datadog, bridging technical metrics and business outcomes.
Join Grafana Labs as a Senior Observability Architect to lead technical partnerships and drive customer success in a fully remote environment.
Join Graphcore as a Senior QA Engineer to enhance the reliability of its observability platform for AI infrastructure.
Join DataHub as a Senior Product Manager to lead the development of observability and data quality solutions in a fully remote environment.
Join Atoms as a Senior Systems Engineer to lead diagnostics and failure analysis for real-world robotic systems.
Join HappyRobot as a Site Reliability Engineer to enhance operational resilience and improve system uptime in a high-impact role.