Showing 4 of 4 positions
Search with filters →Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.
Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.
Join Inferact as a staff engineer to work on optimizing AI inference across the vLLM stack in a fully remote role.