"llm inference optimization" Jobs

83 open tech roles matching “llm inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 83 results

Databricks

Join Databricks as a Staff Software Engineer to build LLM infrastructure for large-scale AI inference workloads.

Databricks San Francisco, California $190k–$265k/yr Published 1 month ago
Fireworks AI

Join Fireworks AI as a Software Engineer to design and build scalable infrastructure for generative AI systems.

Fireworks AI San Mateo Published 9 months ago
Flexible on stack
Fireworks AI

Join Fireworks AI as a Software Engineer focused on Performance Optimization to enhance AI infrastructure efficiency and speed.

Fireworks AI San Mateo Published 1 year ago
Flexible on stack
Anthropic

Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 3 months ago
Flexible on stack
Databricks

Lead a multidisciplinary research team to advance large-scale machine learning efficiency at Databricks.

Databricks Mountain View, California; San Francisco, California $270k–$340k/yr Published 3 months ago
Flexible on stack 60% coding
Together AI

Join Together AI as a Forward Deployed Engineer to optimize inference systems for strategic customers in a hands-on role.

Together AI Singapore Published 2 weeks ago
Flexible on stack 70% coding