"inference optimization" Jobs

24 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 24 results

Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Volta

Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.

Volta Palo Alto, CA Published 1 month ago
Fireworks AI

Join Fireworks AI as an AI Field Engineer to build production systems for generative AI with large organizations across EMEA.

Fireworks AI London Published 1 month ago
Flexible on stack 70% coding
Perplexity AI

Join Perplexity AI as an AI Infrastructure Engineer to build and optimize large-scale AI training and inference clusters.

Perplexity AI London Published 5 months ago
Flexible on stack
CuspAI

Join CuspAI as an intern to develop machine learning force fields for advanced materials simulations.

CuspAI London, UK Published 1 month ago
Fireworks AI

Join Fireworks AI as an Applied Machine Learning Engineer to bridge AI research and real-world applications in a collaborative environment.

Fireworks AI London Published 1 week ago
Flexible on stack
DeepL

Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.

DeepL London Published 1 month ago
Heavy meetings
Graphcore

Join Graphcore as an AI Research Engineer to advance AI research and collaborate on innovative AI hardware solutions.

Graphcore London, UK Published 2 months ago
Flexible on stack
Graphcore

Join Graphcore as a Senior Machine Learning Engineer to advance AI technology on cutting-edge hardware.

Graphcore London, UK Published 9 months ago
Flexible on stack
Synthesia

Join Synthesia as a Senior Applied Research Engineer to develop cutting-edge human-centric video generation models.

Synthesia Europe Published 2 weeks ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to build next-generation observability systems for large-scale AI infrastructure.

Anthropic London, UK £325k–£390k/yr Published 1 week ago
Flexible on stack
Neko Health

Join Neko Health as an ML Ops Engineer to lead the ML infrastructure for preventive care and early detection.

Neko Health London Published 2 months ago
Flexible on stack
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
Lilt

Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.

Lilt London, UK Published 1 month ago
Flexible on stack
Coreweave

Join CoreWeave as an Account Solution Architect to support AI labs and enterprises with GPU infrastructure and MLOps solutions across Northern EMEA.

Coreweave London, UK Published 3 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Field Marketing Manager to shape innovative in-person experiences and drive brand engagement in EMEA.

Fireworks AI London Published 1 week ago