"llm inference systems" Jobs

144 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 144 results

Twelve Labs

Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 5 months ago
Flexible on stack
PagerDuty

Join PagerDuty as a junior AI/ML Engineer to build and ship AI systems at scale, collaborating with senior engineers.

PagerDuty Lisbon Published 3 weeks ago
Flexible on stack
baseten

Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.

baseten San Francisco Published 11 months ago
Twelve Labs

Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.

Twelve Labs Seoul, South Korea Published 1 week ago
krea.ai

Join Krea to build innovative AI tools in a hands-on role focused on supercomputing and distributed systems.

krea.ai San Francisco Published 5 months ago
Flexible on stack
baseten

Join the Base Labs Fellowship to conduct cutting-edge AI research with mentorship and funding in San Francisco.

baseten San Francisco $15k–$15k/mo Published 2 months ago
PagerDuty

Join PagerDuty as a Senior AI/ML Engineer to design and build AI-powered features for high-volume, real-time event streams.

PagerDuty Lisbon Published 3 weeks ago
Flexible on stack
Anthropic

Design and operate backend systems for Claude's safety systems, ensuring low latency and high reliability.

Anthropic San Francisco, CA $320k–$485k/yr Published 5 days ago
LangChain

Join LangChain as a Research Engineer to enhance the capabilities of the LangSmith Engine for AI agents.

LangChain New York, NY Published 1 month ago
Flexible on stack
baseten

Join Baseten as a Software Engineer to lead GPU Networking efforts and optimize distributed systems for AI applications.

baseten San Francisco Published 6 months ago
Flexible on stack
Databricks
Databricks San Francisco, California $280k–$350k/yr Published 4 months ago
baseten

Join Baseten as a Product Manager to shape the future of AI infrastructure and enhance production inference capabilities.

baseten San Francisco Published 5 months ago
Sprinter Health

Join Sprinter Health as a Staff Machine Learning Engineer to build and lead the ML engineering function in a hybrid work environment.

Sprinter Health San Francisco, CA Published 1 month ago
Twelve Labs

Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.

Twelve Labs Seoul, South Korea Published 1 month ago
Flexible on stack
Abridge

Lead product strategy for foundational models and post-training at a growing healthcare AI startup in San Francisco.

Abridge SF Office Published 2 weeks ago
Anthropic

Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.

Anthropic Remote-Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NY $320k–$485k/yr Published 1 week ago
Flexible on stack
Sierra

Join Sierra as a Software Engineer, Infrastructure, to design and maintain core systems for our AI platform in a collaborative environment.

Sierra San Francisco, CA Published 3 months ago
Flexible on stack
Exa

Join Exa's Infrastructure Team to build massive-scale ML systems that redefine how AI consumes information.

Exa San Francisco, California Published 1 year ago