"llm inference optimization" Jobs

8 open tech roles matching “llm inference optimization”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, Python, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 8 of 8 results

Anthropic

Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to enhance deployment infrastructure for AI systems in a collaborative environment.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 1 month ago
Flexible on stack
Anthropic

Join Anthropic's Inference team to design and maintain distributed systems that serve AI models to millions of users worldwide.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 2 months ago
Flexible on stack
LangChain

Join LangChain as a Research Engineer to enhance the capabilities of the LangSmith Engine for AI agents.

LangChain New York, NY Published 1 week ago
Flexible on stack
Harvey AI

Join Harvey AI as a Senior Software Engineer to build and operate core infrastructure for leading law firms and enterprises.

Harvey AI New York $161.3k–$241.9k/yr Published 3 weeks ago
Flexible on stack
Harvey AI

Join Harvey AI as a Staff Software Engineer to build and operate core infrastructure for AI workloads in a fast-growing company.

Harvey AI New York $231k–$340k/yr Published 3 weeks ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$485k/yr Published 4 months ago
Gusto
Gusto Denver, CO - Hybrid; New York, New York, United States; San Francisco, CA - Hybrid $217k–$265k/yr Published 3 months ago