"ai observability" Jobs
4017 open tech roles matching “ai observability”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 4017 results
Lead the Data Visualizations Explorations team at Datadog, focusing on product direction and team development in a hybrid work environment.
Join Obsidian Security as a Site Reliability Engineer to enhance the stability and performance of our leading SaaS security platform.
Design and build scalable backend systems for an AI-native operations platform in a collaborative team environment.
Join Doss as a Backend Engineer to design and build scalable data models and tooling for an AI-native operations platform.
Join Obsidian Security as a Staff Software Engineer to lead impactful projects across the engineering stack in a fast-growing SaaS security company.
Join Obsidian Security as a Staff IT Systems Engineer to lead identity and endpoint management in a rapidly growing SaaS security company.
Join MongoDB's Networking & Observability team to enhance distributed database communication and observability features.
Join Together AI as a Software Engineer to build customer-facing visibility systems and enhance analytics capabilities.
Join GitLab as a Backend Engineer to develop AI-driven features and enhance software delivery systems in a fully remote environment.
Join MongoDB as an Operations Business Analyst to drive insights and reporting for business decision-making in a hybrid work environment.
Join Braintrust as an Open Source Engineer to build SDKs that enhance AI observability and developer experience.
Lead a team in managing engineering execution for Datadog's Observability Pipelines product while fostering team growth and collaboration.
Join HappyRobot as a Site Reliability Engineer to enhance operational resilience and improve system uptime in a high-impact role.
Design and build scalable backend systems for an AI-native operations platform in a collaborative San Francisco team.
Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.
Lead platform observability and proactive monitoring efforts to enhance customer experience and platform stability at Databricks.
Join Orkes as a Site Reliability Engineer to enhance the reliability and performance of cloud-based production systems.
Lead engineering for Cloud Observability at Datadog, managing a team of ~40 engineers in a hybrid work environment.
Join Graphcore as a Data Scientist to transform telemetry into predictive insights for AI infrastructure.