16 Vllm roles across 6 companies, most in AI/ML; 13% fully remote; typical advertised salary $192k
Who is hiring (6 companies)
Role types
Work arrangement: 2 fully remote · 4 hybrid · 3 on-site · 7 not stated
Advertised salaries: p25 $188k · median $192k · p75 $200k (from 13 disclosed annual salaries, USD)
Counts are open roles Joblaze currently tracks on company career pages for this exact skill/location; salaries are advertised minimums, annual, converted to USD.
Showing 16 of 16 positions
Search with filters →Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Own the inference systems that power frontier AI models in production and research at a tech-first startup.
Join Stuut as a Member of the Technical Staff to design and deploy AI-powered systems for financial operations.
Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.