"llm inference systems" Jobs
392 open tech roles matching “llm inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, PyTorch. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 392 results
Join Exa's Infrastructure Team to build massive-scale ML systems that redefine how AI consumes information.
Design and operate data systems that power Decagon's AI products, ensuring high reliability and performance.
Join Reddit as a Senior Security Engineer to secure AI-powered products and build reusable security primitives.
Own the Graph Platform at Enterpret, leading data systems and architecture decisions for a customer feedback intelligence platform.
Join Sierra as a Software Engineer on the Site Reliability team to enhance the reliability and scalability of AI-driven infrastructure.
Join CoreWeave as a Staff Applied ML Engineer to tackle challenges in continuous learning for AI agents with a highly autonomous team.
Own the technical vision of Faire’s ML platform, leading cross-functional initiatives to enhance data science velocity at scale.
Join Sesame as a Research Engineer to innovate in NLP, Speech, and Computer Vision with a focus on deep learning.
Lead scientific innovation in speech and translation models for real-time voice products at DeepL.
Join Tavus as a Senior Software Engineer to enhance the infrastructure behind real-time AI conversations.
Lead the Forward Deployed Engineering function at Reflection AI, bridging cutting-edge research with enterprise deployments.
Join EigenLayer as a Senior Agentic AI Engineer to build core systems for AI agents in a remote-friendly environment.
Lead the Forward Deployed Engineering function at Reflection AI, bridging cutting-edge AI research with enterprise deployments.
Join AssemblyAI as a Senior Design Engineer to shape and build exceptional product experiences in a fast-growing Voice AI company.
Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.
Lead a team of engineers to build and operate CoreWeave's next-generation Kubernetes-native inference platform.