10 jobs tagged TGI
Join Inferact as an inference runtime engineer to innovate AI inference engines for large models in a fully remote role.
Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.
Own the end-to-end lifecycle of production ML serving systems for a top-performing AI Shopping Agent.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as an inference runtime engineer to optimize AI model execution across diverse hardware and architectures.
Join Inferact as a Developer Relations Engineer to shape how developers learn and build with vLLM, the AI inference engine.