3 jobs tagged LLM inference
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.
We use cookies to keep the site working and, with your consent, to analyze usage with privacy-friendly analytics. See our Cookie Policy for details.