"vllm in" Jobs

131 open tech roles matching “vllm in”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, vLLM. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 131 results

Coreweave

Join CoreWeave as an Applied AI Engineer to enhance the performance of our inference platform through benchmarking and optimization.

Coreweave Bellevue, WA/ San Francisco, CA/ Sunnyvale, CA $188k–$275k/yr Published 6 months ago
Flexible on stack
Coreweave

Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.

Coreweave Bellevue, WA - US $188k–$303k/yr Published 9 months ago
Heavy meetings
Arena

Join Arena as a Site Reliability Engineer to build core infrastructure for AI model evaluations at scale.

Arena Bay Area Published 1 month ago
Flexible on stack
Dialpad

Join Dialpad as a Software Engineer to build and improve ML inference systems for AI models at scale.

Dialpad Buenos Aires, Argentina Published 2 months ago
Flexible on stack 70% coding
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $139k–$204k/yr Published 7 months ago
Arena

Join Arena as an Infrastructure Engineer to build scalable systems for real-world AI model evaluation.

Arena Bay Area Published 1 month ago
Flexible on stack
Coreweave

Join CoreWeave as a Senior Engineer to build performance insights and observability systems for AI infrastructure.

Coreweave Sunnyvale, CA / Bellevue, WA $182k–$242k/yr Published 1 month ago
Flexible on stack
Base Power Company

Join Base Power Company as a Server Architect to build a distributed GPU fleet for the AI industry.

Base Power Company Austin, TX Published 2 weeks ago
Flexible on stack
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Coreweave

Drive the adoption of AI runtime services at CoreWeave, leveraging your expertise in distributed systems and AI infrastructure.

Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $207k–$275k/yr Published 2 months ago
Flexible on stack
Arena

Join Arena as a Backend Engineer to build APIs and services for AI model evaluations in a fast-paced startup environment.

Arena Bay Area Published 1 month ago
Dialpad

Join Dialpad as a Senior Software Engineer to build and improve the AI/ML inference platform for enterprise-scale applications.

Dialpad Buenos Aires, Argentina Published 6 days ago
Flexible on stack 70% coding
Coreweave

Serve as a technical partner for federal customers, designing AI workloads that meet mission requirements and compliance standards.

Coreweave Washington, D.C. $165k–$220k/yr Published 2 months ago
Flexible on stack
Cartesia

Join Cartesia as an Inference Engineer to design and build low latency, scalable model inference for cutting-edge AI applications.

Cartesia *HQ - San Francisco, CA Published 1 year ago
Flexible on stack
Handshake

Join Handshake as a Senior Software Engineer to build scalable ML infrastructure for a fast-growing AI data business.

Handshake San Francisco, CA Published 2 months ago
Flexible on stack
Inflection AI

Lead the architecture and technical direction of agentic AI systems at Inflection AI, building production AI agents.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Flexible on stack
baseten

Join Baseten as a Technical Program Manager to build and optimize the core algorithms for high-performance AI inference.

baseten San Francisco Published 2 weeks ago
baseten

Join Baseten as a Forward Deployed Engineer to solve complex AI challenges for leading companies.

baseten San Francisco Published 3 weeks ago
Flexible on stack