"inference optimization" Jobs
46 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: AI/ML, Python, Machine Learning. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 46 results
Lead the financial strategy for AI products at Perplexity, optimizing model spend and driving pricing decisions.
Lead the economics of AI products at Perplexity AI, optimizing model spend and driving pricing and margin decisions.
Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic environment.
Join Inworld AI as a Lead Machine Learning Engineer to optimize and serve state-of-the-art voice models in a dynamic startup environment.
Lead product strategy for inference infrastructure and token-serving capabilities in a rapidly growing AI infrastructure company.
Join Perplexity as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions in a fast-paced environment.
Join Perplexity AI as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions.
Lead a team focused on optimizing machine learning models for ads efficiency at Reddit.
Lead the post-training and evaluation capabilities for large language models in a dynamic AI research lab.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.
Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and AI applications.
Lead the perception model team for autonomous vehicles at a rapidly growing AI infrastructure company.
Join Inworld AI as a Lead Research Scientist to innovate in real-time voice models and impact AI applications globally.
Lead the infrastructure team at Omnifold, focusing on AI model training and deployment in a fast-paced startup environment.
Lead a team of MLOps engineers at an AI company transforming enterprise decision-making.
Lead performance marketing efforts to optimize paid acquisition and build brand awareness for a top AI voice model company.
Lead a team of cloud platform engineers to build scalable and reliable infrastructure for AI products at Baseten.
Lead the Production Inference team at DeepL, focusing on performance-critical model serving systems in a fast-paced AI environment.
Lead a team of engineers to build an AI-native platform for catalog enrichment in a fully remote environment.