"llm inference systems" Jobs
144 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 144 results
Drive research on Pegasus's complex problems in a hybrid role at a growing AI company focused on video understanding.
Join PagerDuty as a junior AI/ML Engineer to build and ship AI systems at scale, collaborating with senior engineers.
Join Baseten as a Software Engineer focusing on Model APIs to enhance AI model performance and developer experience.
Drive technical direction for training infrastructure and operations within Pegasus at a growing AI company focused on video understanding.
Join the Base Labs Fellowship to conduct cutting-edge AI research with mentorship and funding in San Francisco.
Design and operate backend systems for Claude's safety systems, ensuring low latency and high reliability.
Join LangChain as a Research Engineer to enhance the capabilities of the LangSmith Engine for AI agents.
Join Baseten as a Product Manager to shape the future of AI infrastructure and enhance production inference capabilities.
Join Sprinter Health as a Staff Machine Learning Engineer to build and lead the ML engineering function in a hybrid work environment.
Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.
Lead product strategy for foundational models and post-training at a growing healthcare AI startup in San Francisco.
Join Anthropic as a Staff+ Site Reliability Engineer to ensure safe AI model launches and automate deployment processes.
Join Sierra as a Software Engineer, Infrastructure, to design and maintain core systems for our AI platform in a collaborative environment.
Join Exa's Infrastructure Team to build massive-scale ML systems that redefine how AI consumes information.