Direct from source · No middlemen
131 open positions · Updated 1 week ago
Showing 20 of 131 positions
Search with filters →Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Intel as an AI Framework Software Engineer to work on cutting-edge AI technologies and frameworks.
Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.
Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.
Join Inferact as a staff engineer to build distributed systems for AI inference at global scale.
Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.
Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.