"llm inference" Jobs
451 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 451 results
Join Fireworks AI as a Member of Technical Staff to build innovative AI solutions on a large inference platform.
Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.
Lead the financial strategy for AI products at Perplexity, optimizing model spend and driving pricing decisions.
Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models within a fully remote team in Switzerland.
Join a world-class team to lead transformative research in generative AI for drug discovery at Genesis Molecular AI.
Fine-tune state-of-the-art LLMs and develop AI products to enhance the travel experience at Airbnb.
Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and impact AI applications globally.
Own the model lifecycle for defence and strategic sector deployments as an MLOps Engineer at Sarvam AI.
Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.
Join Anthropic as a Staff Software Engineer to design and optimize backend services for cloud inference at scale.
Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.
Lead the economics of AI products at Perplexity AI, optimizing model spend and driving pricing and margin decisions.
Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.
Lead the design and development of core ML models for Instacart’s ads ecosystem in a fully remote role.
Join Freenome as a Senior Machine Learning Engineer to develop and optimize deep learning pipelines for cancer detection.
Join Anthropic's Inference team to design and maintain distributed systems that serve AI models to millions of users worldwide.