"llm inference" Jobs
20 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 20 results
Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.
Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.
Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.
Join Fireworks AI as a Member of Technical Staff to build innovative AI solutions on a large inference platform.
Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Join Fireworks AI as an AI Field Engineer to build production systems for generative AI with large organizations across EMEA.
Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a collaborative environment.
Join DeepL as a Senior Software Engineer to build innovative real-time voice translation solutions in a dynamic, cross-functional team.
Lead the establishment and growth of Perplexity's London office, shaping its engineering culture and technical direction.
Lead the establishment and growth of Perplexity's London office, shaping its engineering culture and technical direction.
Lead scientific innovation in speech and translation models for real-time voice products at DeepL.
Join Together AI to build production AI agents and foundational systems for one of the largest GPU fleets in the world.
Join Together AI as a Solutions Architect to drive customer success through innovative Generative AI applications.