"inference serving" Jobs
373 open tech roles matching “inference serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 373 results
Join Sierra as a Software Engineer on the Site Reliability team to enhance the reliability and scalability of AI-driven infrastructure.
Lead hyperscaler partnerships at Together AI, driving complex commercial negotiations and strategic collaborations.
Join CoreWeave as a Senior Specialist Field Engineer to lead GPU cluster delivery and ensure high-performance AI workloads for customers.
Join Baseten as a Software Engineer to own the internal tooling that drives AI product operations and improve high-stakes workflows.
Join HoneyBook as a Senior Fraud Analyst to lead the technical evolution of fraud defenses using AI and advanced analytics.
Join Anthropic as a Staff+ Software Engineer to build production systems for capacity engineering in a hybrid work environment.
Join Cursor as a Technical Program Manager to drive infrastructure efficiency and resource allocation in a dynamic startup environment.
Join PagerDuty as a junior AI/ML Engineer to build and ship AI systems at scale, collaborating with senior engineers.
Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.
Own cost-of-revenue analysis to improve gross margin and unit economics in a growth-stage voice AI company.
Join Baseten as a Workplace Operations team member to enhance the experience in our rapidly growing San Francisco office.
Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing AI company.
Lead the sales organization for digital native companies at AssemblyAI, focusing on voice AI solutions.
Join Baseten as a Product Manager to shape enterprise readiness for AI products in a fast-growing company.
Join Anthropic as a Staff+ Software Engineer to design and build systems for cloud portability in AI infrastructure.
Drive adoption of Fireworks' generative AI platform by engaging with technical founders and product teams in a fast-paced environment.