Direct from source · No middlemen

Xla Jobs

8 open positions · Updated 1 day ago

Average salary (USD/year): 218.4k–473.6k/yr
8 roles Staff 4, Senior 2, Intern 1, Mid 1 3–10 yrs experience San Francisco +5 more Full-time, Internship newest 1 day ago

Showing 8 of 8 positions

Search with filters →
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco, California, United States Published 2 days ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working directly with hardware vendors.

Inferact Singapore S$200k–S$400k/yr Published 2 months ago
Flexible on stack
Applied Intuition

Join Applied Intuition as an ML Runtime Optimization Engineer to optimize ML models for embedded environments in a collaborative team.

Applied Intuition Sunnyvale Published 1 year ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA $315k–$560k/yr Published 10 months ago
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $280k–$850k/yr Published 11 months ago