"inference serving" Jobs
961 open tech roles matching “inference serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 961 results
Join Sarvam as a Backend Engineer to build and maintain production services for a cutting-edge AI media platform.
Own procurement and contract strategy for data center construction at SpaceX, enabling the build-out of high-density, high-reliability data centers.
Join fal as a Deal Desk & Pricing Operations professional to streamline enterprise deal structures and support sales growth.
Join Horizon3.ai as a Senior Machine Learning Engineer to operationalize AI models in a remote-first cybersecurity environment.
Join Harvey AI as a Senior Software Engineer to build and operate core infrastructure for leading law firms and enterprises.
Join Fireworks AI as a new grad to work on real AI systems and ship production code on a leading inference platform.
Join Baseten as a Software Engineer to build continuous deployment infrastructure for AI products in a fast-growing team.
Join Inworld AI as a Staff/Principal Platform Engineer to build and scale AI products with a focus on cloud infrastructure.
Join Harvey AI as a Staff Software Engineer to build and operate core infrastructure for AI workloads in a fast-growing company.
Drive sourced pipeline and revenue through the Microsoft Azure channel in a high-ownership sales role at Fireworks AI.
Join AMP Sortation as a Machine Learning Engineer to develop cutting-edge AI solutions for modernizing recycling infrastructure.
Support the CTO and Head of Engineering in a highly operational role at a rapidly growing AI company.
Join Harvey as a Staff Product Manager to shape the infrastructure of a leading legal AI platform with a focus on reliability and scalability.
Join Horizon3.ai as a Senior Software Engineer to build an autonomous pentesting agent for cybersecurity.
Own the reliability and security of fal's generative media model APIs in a hybrid ML Engineering/SRE role.
Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.