"llm inference systems" Jobs

144 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 144 results

Twelve Labs

Lead the development of next-generation multimodal models at Twelve Labs, impacting thousands of customers worldwide.

Twelve Labs Seoul, South Korea Published 1 week ago
Flexible on stack
Stuut

Join Stuut as a Member of the Technical Staff to design and deploy AI-powered systems for financial operations.

Stuut San Francisco Published 1 month ago
Flexible on stack
Wispr Flow
ML Engineer Hybrid Visa

Join Wispr Flow as a ML Engineer to build a scalable voice interface for millions of users.

Wispr Flow San Francisco Published 1 year ago
Flexible on stack
HappyRobot

Join HappyRobot as a Machine Learning Engineer to build AI models for human-like conversations and shape the future of AI infrastructure.

HappyRobot San Francisco Published 1 month ago
Flexible on stack
Preference Model

Join Preference Model as a Senior Machine Learning Engineer to design RL environments for advancing ML capabilities.

Preference Model San Francisco, United States Published 3 days ago
Flexible on stack
Preference Model

Join Preference Model as a senior ML Engineer to design RL environments for advancing machine learning capabilities.

Preference Model San Francisco Published 2 weeks ago
Flexible on stack
Preference Model

Join Preference Model as a new graduate Machine Learning Engineer to design and build reinforcement learning environments.

Preference Model San Francisco Published 4 months ago
Flexible on stack
Databricks
Databricks San Francisco, California $54–$60/hr Published 2 years ago
Reflection AI

Join Reflection AI to build systems that transform pre-trained models into aligned agents in a fast-paced startup environment.

Reflection AI San Francisco, CA Published 1 month ago
Together AI

Join Together AI as a Research Engineer to develop a platform for customizing open-source models with user data.

Together AI San Francisco $200k–$290k/yr Published 2 months ago
Flexible on stack
Harvey

Lead the design and development of systems powering AI requests at Harvey, a fast-scaling company in the legal tech space.

Harvey San Francisco $231k–$340k/yr Published 1 week ago
Flexible on stack
Hilbert

Join Hilbert as an AI Engineer to build production-grade AI systems that drive enterprise outcomes in a fast-paced startup environment.

Hilbert San Francisco Published 6 months ago
Flexible on stack 70% coding
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to enhance deployment infrastructure for AI systems in a collaborative environment.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 2 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to optimize and scale AI inference across major cloud platforms.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Senior Applied AI Engineer to shape agent capabilities and enhance user experiences with cutting-edge AI technologies.

Perplexity AI San Francisco Published 4 days ago
Perplexity

Join Perplexity as a staff Applied AI Engineer to shape agent capabilities and enhance user experiences with cutting-edge AI technologies.

Perplexity San Francisco Published 4 days ago
Omnifold

Lead a research team at Omnifold to develop advanced forecasting and optimization models in a startup environment.

Omnifold San Francisco HQ Published 4 days ago
Harvey AI

Lead the design and development of systems powering AI requests at Harvey, collaborating with multiple teams to ensure reliability and efficiency.

Harvey AI San Francisco $236k–$290k/yr Published 1 month ago
Flexible on stack
Twelve Labs

Build and operate production ML systems for Pegasus, focusing on reliability and performance in a hybrid work environment.

Twelve Labs Seoul, South Korea Published 3 weeks ago
Flexible on stack