"inference runtimes" Jobs

8 open tech roles matching “inference runtimes”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, Rust. Every listing is re-checked daily and closed roles are removed.

Showing 8 of 8 results

Perplexity AI

Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.

Perplexity AI London Published 5 months ago
Flexible on stack
Together AI

Join Together AI as a Staff Software Engineer to build systems that automate infrastructure management for AI clusters.

Together AI London & Amsterdam Published 3 weeks ago
Flexible on stack
DeepL

Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.

DeepL London Published 1 month ago
Heavy meetings
Volta

Join Volta as a Platform Engineer to build and operate large-scale GPU compute infrastructure for AI workloads.

Volta London, UK Published 1 month ago
Flexible on stack
Volta

Join Volta as a Platform Engineer to build and operate large-scale GPU compute infrastructure for AI workloads.

Volta Palo Alto, CA Published 1 month ago
Flexible on stack
Wiz

Join Wiz as a Deal Desk Analyst to support international sales teams and help structure high-volume deals in a fast-paced environment.

Wiz London, UK; Remote - United Kingdom Published 1 month ago