"inference serving systems" Jobs
735 open tech roles matching “inference serving systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 735 results
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Lead the perception model team for autonomous vehicles at a rapidly growing AI infrastructure company.
Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.
Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.
Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.
Join Applied Intuition as an Embedded AI Engineer to develop on-device intelligence for Android Automotive platforms.
Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.
Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.
Join Periodic Labs as an ML Systems Engineer to build and optimize large-scale training and reinforcement learning infrastructure.
Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.
Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.
Own the Graph Platform at Enterpret, leading data systems and architecture decisions for a customer feedback intelligence platform.
Join Applied Intuition as a Senior Software Engineer to design and implement ML infrastructure for deep learning model training.