Direct from source · No middlemen
52 open positions · Updated 1 week ago
52 Vllm roles across 24 companies, most in AI/ML; 10% fully remote; typical advertised salary $190k
Who is hiring (24 companies)
Role types
Work arrangement: 5 fully remote · 11 hybrid · 11 on-site · 25 not stated
Advertised salaries: p25 $157.8k · median $190k · p75 $200k (from 25 disclosed annual salaries, USD)
Counts are open roles Joblaze currently tracks on company career pages for this exact skill/location; salaries are advertised minimums, annual, converted to USD.
Showing 20 of 52 positions
Search with filters →Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join ChipAgents as an ML Systems Engineer to optimize LLM inference systems for leading semiconductor companies.
Own Sarvam's production serving path for large distributed models, integrating and optimizing performance across a multi-node stack.
Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.
Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.
Join Handshake as a Senior Software Engineer to build scalable ML infrastructure for a fast-growing AI data business.
Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.