"inference infrastructure" Jobs
88 open tech roles matching “inference infrastructure”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, Kubernetes. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 88 results
Lead a team of engineers to optimize Anthropic's inference infrastructure for AI systems.
Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.
Lead the infrastructure team at Omnifold, focusing on AI model training and deployment in a fast-paced startup environment.
Join Anthropic as a Tech Lead to build reliable AI evaluation systems in a hybrid work environment.
Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic environment.
Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.
Lead product strategy for foundational models and post-training at a growing healthcare AI startup in San Francisco.
Lead the development of AI-powered features for an end-to-end intelligence platform in public safety.
Lead the Runtime Fabric team at Baseten to build container runtimes tailored for AI inference workloads.
Lead HR and People Operations at Inferact, scaling infrastructure in a fast-paced startup environment.
Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.
Join Perplexity AI as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions.
Join Perplexity as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions in a fast-paced environment.
Lead the perception model team for autonomous vehicles at a rapidly growing AI infrastructure company.
Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and AI applications.
Lead a team of cloud platform engineers to build scalable and reliable infrastructure for AI products at Baseten.
Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and impact AI applications globally.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.