"llm inference systems" Jobs

392 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 392 results

Together AI

Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.

Together AI San Francisco $220k–$280k/yr Published 3 months ago
Flexible on stack 60% coding
Genesis Molecular AI

Join Genesis Molecular AI as an ML Research Engineer to develop cutting-edge foundation models for drug discovery.

Genesis Molecular AI San Mateo, CA Published 1 year ago
Flexible on stack 70% coding
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and impact AI applications globally.

Inworld AI Germany Published 5 months ago
Anthropic

Join Anthropic as a Staff Software Engineer to design and optimize backend services for cloud inference at scale.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 months ago
Flexible on stack
Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Genesis Molecular AI

Join a world-class team to lead transformative research in generative AI for drug discovery at Genesis Molecular AI.

Genesis Molecular AI San Mateo, CA Published 1 year ago
Flexible on stack
Fundamental

Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.

Fundamental Europe Published 2 months ago
Flexible on stack
Stuut

Join Stuut as a Member of the Technical Staff to design and deploy AI-powered systems for financial operations.

Stuut San Francisco Published 1 month ago
Flexible on stack
Wispr Flow
ML Engineer Hybrid Visa

Join Wispr Flow as a ML Engineer to build a scalable voice interface for millions of users.

Wispr Flow San Francisco Published 1 year ago
Flexible on stack
HappyRobot

Join HappyRobot as a Machine Learning Engineer to build AI models for human-like conversations and shape the future of AI infrastructure.

HappyRobot San Francisco Published 1 month ago
Flexible on stack
Inworld AI

Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and AI applications.

Inworld AI Serbia Published 5 months ago
OpenRouter

Conduct original research on large language models to advance understanding and routing optimization at OpenRouter.

OpenRouter Remote (US) Published 2 months ago
Flexible on stack
Preference Model

Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.

Preference Model San Francisco, United States Published 1 day ago
Flexible on stack
Lilt

Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.

Lilt London, UK Published 1 month ago
Flexible on stack
Coreweave

Join CoreWeave as a Staff Software Engineer to lead the development of a Kubernetes-native inference platform for AI workloads.

Coreweave Sunnyvale, CA / Bellevue, WA $188k–$275k/yr Published 4 months ago
Flexible on stack
Inflection AI

Lead the development of Inflection's realtime Voice AI stack, shaping emotionally intelligent AI for enterprise voice interactions.

Inflection AI Palo Alto, California, United States $400k–$550k/yr Published 2 months ago
Sarvam AI

Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.

Sarvam AI Delhi Published 4 months ago
Flexible on stack