"inference optimization" Jobs
24 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 24 results
Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.
Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.
Join Fireworks AI as an AI Field Engineer to build production systems for generative AI with large organizations across EMEA.
Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.
Join CuspAI as an intern to develop machine learning force fields for advanced materials simulations.
Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a collaborative environment.
Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.
Join Anthropic as a Staff Software Engineer to build next-generation observability systems for large-scale AI infrastructure.
Join Neko Health as an ML Ops Engineer to lead the ML infrastructure for preventive care and early detection.
Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.
Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.
Join CoreWeave as an Account Solution Architect to support AI labs and enterprises with GPU infrastructure and MLOps solutions across Northern EMEA.
Join Fireworks AI as a Field Marketing Manager to shape innovative in-person experiences and drive brand engagement in EMEA.