"llm inference systems" Jobs

166 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 166 results

OpenRouter

Conduct original research on large language models to advance understanding and routing optimization at OpenRouter.

OpenRouter Remote (US) Published 2 months ago
Flexible on stack
Preference Model

Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.

Preference Model San Francisco, United States Published 2 days ago
Flexible on stack
Lilt

Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.

Lilt London, UK Published 1 month ago
Flexible on stack
Preference Model

Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
DeepL

Lead research on fine-tuning and steerability of LLM-based translation models in a collaborative AI-focused environment.

DeepL London Published 1 month ago
Flexible on stack
Coreweave
Coreweave Sunnyvale, CA / Bellevue, WA $165k–$242k/yr Published 11 months ago
Instacart

Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.

Instacart United States - Remote $201k–$253.5k/yr Published 3 months ago
Flexible on stack
Perplexity

Join Perplexity as a Machine Learning Engineer to enhance search quality through innovative ranking solutions.

Perplexity Belgrade Published 1 month ago
Fundamental

Join Fundamental as a Model Serving Engineer to optimize and scale the NEXUS model for enterprise decision-making.

Fundamental Europe Published 5 months ago
Flexible on stack
Together AI

Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.

Together AI Singapore Published 1 month ago
Flexible on stack 70% coding
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Machine Learning Engineer to enhance search quality through innovative ranking solutions.

Perplexity AI Belgrade Published 1 month ago
Cylake

Join a small team to build state-of-the-art AI capabilities for Cylake's next-generation cybersecurity platform.

Cylake Sunnyvale $150k–$250k/yr Published 1 month ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Senior Applied AI Engineer to shape agent capabilities and enhance user experiences with cutting-edge AI technologies.

Perplexity AI San Francisco Published 3 days ago
Cloudflare

Join Cloudflare as a Senior Systems Engineer to build core AI Gateway systems for high-volume inference traffic.

Cloudflare In-Office Published 1 month ago
Illumio

Architect high-scale distributed systems and lead the development of autonomous AI agents in a dynamic cybersecurity environment.

Illumio HQ - Sunnyvale (Office) Published 3 months ago
Flexible on stack
Freenome

Join Freenome as a Senior Machine Learning Engineer to develop and optimize deep learning pipelines for cancer detection.

Freenome Remote $173.8k–$246.8k/yr Published 1 month ago
Flexible on stack
Twelve Labs

Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.

Twelve Labs Seoul, South Korea Published 3 weeks ago
Flexible on stack