"inference serving" Jobs
886 open tech roles matching “inference serving”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 886 results
Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.
Join Anthropic as a Staff Software Engineer to enhance deployment infrastructure for AI systems in a collaborative environment.
Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.
Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.
Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.
Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.
Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.
Join Baseten as a Data Engineer to build and scale the internal data platform for AI-driven decision-making.
Join Together AI as a Technical Support Engineer to tackle complex technical challenges in a fast-paced AI environment.
Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.
Join Inferact as a staff engineer to build distributed systems for AI inference at global scale.
Join Baseten as an AI Inference Engineer to architect and deploy high-scale production AI applications while collaborating with customers.
Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.