"inference systems" Jobs

396 open tech roles matching “inference systems”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, AI/ML, Kubernetes. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 396 results

Faire

Own the measurement and optimization of long-term value in a hybrid role at Faire, a tech-driven wholesale platform.

Faire New York City, NY; San Francisco, CA $246.5k–$339k/yr Published 4 days ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Senior Applied AI Engineer to shape agent capabilities and enhance user experiences with cutting-edge AI technologies.

Perplexity AI San Francisco Published 5 days ago
baseten

Join Baseten as a Sr. Analyst in Revenue Strategy & Operations to shape GTM strategies for AI infrastructure.

baseten San Francisco Published 1 week ago
baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $405k–$485k/yr Published 5 months ago
Anthropic

Join Anthropic as a Staff Software Engineer to build next-generation observability systems for large-scale AI infrastructure.

Anthropic London, UK £325k–£390k/yr Published 1 week ago
Flexible on stack
Decagon

Design and operate data systems that power Decagon's AI products, ensuring high reliability and performance.

Decagon San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Figma
Figma San Francisco, CA • New York, NY • United States $153k–$376k/yr Published 1 year ago
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Perplexity

Join Perplexity as a staff Applied AI Engineer to shape agent capabilities and enhance user experiences with cutting-edge AI technologies.

Perplexity San Francisco Published 5 days ago
Anthropic

Join Anthropic as a Demand Planning expert to optimize AI infrastructure capacity and ensure timely delivery across multiple platforms.

Anthropic San Francisco, CA | New York City, NY $320k–$405k/yr Published 1 month ago
Flexible on stack
Harvey AI

Lead a high-performing team to develop and manage the model infrastructure platform at Harvey AI.

Harvey AI San Francisco $272k–$355k/yr Published 1 month ago
Heavy meetings
MaintainX

Lead the technical direction for predictive maintenance and asset intelligence initiatives at MaintainX, leveraging deep ML expertise.

MaintainX San Francisco Published 1 month ago
Flexible on stack
Meter

Join Meter as a Backend Engineer to design and implement a unified data interface for model development in a cutting-edge networking company.

Meter San Francisco $160k–$230k/yr Published 1 year ago
Flexible on stack
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
baseten

Join Baseten to lead capacity planning and operations in a fast-paced AI environment.

baseten San Francisco Published 5 days ago