"inference optimization" Jobs
214 open tech roles matching “inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AI/ML. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 214 results
Join Perplexity AI as an AI Infrastructure Engineer to design and optimize large-scale AI training and inference clusters.
Join Baseten as a Technical Program Manager to build and optimize the core algorithms for high-performance AI inference.
Join Baseten as a Forward Deployed Engineer to solve complex AI challenges for leading companies.
Join Goodfire as a Machine Learning Engineer to build interpretable AI systems with a world-class team.
Join OpenRouter as a Forward Deployed Engineer to help customers implement and scale AI solutions effectively.
Join Together AI as a Staff ML Engineer to optimize voice model serving for real-time applications on a high-impact team.
Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.
Join Baseten as a Cloud Platform Engineer to build scalable infrastructure for deploying machine learning models.
Join Omnifold's Infrastructure Team to build robust systems for AI model training and deployment in a fast-paced environment.
Join Harvey AI as a Research Engineer to drive post-training experiments and enhance legal AI models.
Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.