"vllm" Jobs
52 open tech roles matching “vllm”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, vLLM, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 52 results
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Own the inference systems that power frontier AI models in production and research at a tech-first startup.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.
Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.
Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.
Join LlamaIndex as an AI Research Engineer to enhance document understanding systems through applied research and engineering.
Join Stuut as a Member of the Technical Staff to design and deploy AI-powered systems for financial operations.
Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.
Join Baseten as a senior software engineer to develop cutting-edge AI training products and enhance user workflows.
Join Abridge as a Machine Learning Infrastructure Engineer to optimize AI model inference infrastructure in a fast-paced healthcare startup.
Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.
Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.