"inference serving" Jobs
391 open tech roles matching “inference serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 391 results
Own the serving infrastructure for healthcare AI, optimizing LLM inference systems to enhance patient experiences.
Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.
Join Roboflow as a Machine Learning Engineer to enhance our inference engine and contribute to impactful computer vision projects.
Lead complex, cross-functional programs for inference platform delivery at a rapidly growing AI cloud company.
Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.
Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.
Join Perplexity as a technical program manager to drive the core inference platform and coordinate between model providers and engineering teams.
Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.
Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.
Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.
Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.
Join Together AI as a Technical Support Engineer to tackle complex technical challenges in a fast-paced AI environment.
Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.
Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.