"llm inference optimization" Jobs
83 open tech roles matching “llm inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.
Showing 20 of 83 results
Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.
Join Fireworks AI as a Software Engineer to design and build scalable infrastructure for generative AI systems.
Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.
Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.
Lead a multidisciplinary research team to advance large-scale machine learning efficiency at Databricks.
Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.