"llm inference" Jobs
451 open tech roles matching “llm inference”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 451 results
Join Perplexity AI as a technical program manager to drive the core inference platform and coordinate across teams and model providers.
Join Anthropic's Inference team to build and maintain systems that serve AI models to millions of users worldwide.
Join Fundamental as an ML Researcher to tackle groundbreaking challenges in AI model development for enterprise decision-making.
Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.
Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.
Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.
Join Perplexity AI as an AI Inference Engineer to optimize and develop our inference engine for various model architectures.
Join Krea as an ML Researcher to finetune diffusion models and enhance AI creative tools in a collaborative environment.
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI workloads.
Join Inworld AI as a Staff/Principal Research Scientist to innovate in real-time voice models and AI applications.
Conduct original research on large language models to advance understanding and routing optimization at OpenRouter.
Join Wispr Flow as a ML Engineer to build a scalable voice interface for millions of users.
Join Genesis Molecular AI as an ML Research Engineer to develop cutting-edge foundation models for drug discovery.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.
Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.
Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.
Join LILT as a Forward Deployed Engineer to integrate AI solutions for complex clients and enhance global communication.