"llm inference" Jobs
165 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 165 results
Join Baseten as a Software Engineer to build and optimize large-scale LLM inference systems in a collaborative environment.
Join Perplexity as a technical program manager to drive the core inference platform and coordinate between model providers and engineering teams.
Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Preference Model as a Senior ML Infrastructure Engineer to build scalable infrastructure for post-training research on large language models.
Join Perplexity AI as a technical program manager to drive the core inference platform and coordinate across teams and model providers.
Join Anthropic's Inference team to build and maintain systems that serve AI models to millions of users worldwide.
Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.
Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI workloads.
Join Wispr Flow as a ML Engineer to build a scalable voice interface for millions of users.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.
Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.
Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.
Lead the financial strategy for AI products at Perplexity, optimizing model spend and driving pricing decisions.
Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.