"caching" Jobs

135 open tech roles matching “caching”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Python, Kubernetes, AWS. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 135 results

Anthropic

Join Anthropic as a Staff+ Software Engineer to build a managed caching service that enhances performance across AI systems.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $320k–$485k/yr Published 2 months ago
Flexible on stack
Cursor

Join Cursor as a Software Engineer on the Storage team to design and manage the data layer for millions of developers.

Cursor San Francisco Published 2 months ago
Flexible on stack
Vercel

Join Vercel as a Software Engineer to enhance CDN performance and work with a diverse tech stack in a collaborative environment.

Vercel Hybrid - San Francisco $172k–$258k/yr Published 2 months ago
Flexible on stack
Vercel
Vercel Hybrid - San Francisco $196k–$336k/yr Published 10 months ago
Together AI

Design and deliver multi-petabyte storage systems for AI workloads at Together AI, optimizing performance and cost.

Together AI San Francisco $250k–$300k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.

Anthropic San Francisco, CA | New York City, NY $350k–$850k/yr Published 3 days ago
Flexible on stack
Inferact

Join Inferact as a co-op student to work on cutting-edge AI inference systems in a hands-on engineering role.

Inferact San Francisco Published 2 days ago
Flexible on stack
Sesame

Join Sesame as an ML Model Serving Engineer to enhance our serving layer for voice agents with cutting-edge techniques.

Sesame San Francisco Published 1 year ago
Flexible on stack
Anthropic

Join Anthropic as a Staff Software Engineer to optimize and scale AI inference across major cloud platforms.

Anthropic San Francisco, CA $320k–$485k/yr Published 3 months ago
Flexible on stack
Anthropic

Join Anthropic's Inference team to design and maintain distributed systems serving AI models to millions globally.

Anthropic New York City, NY; San Francisco, CA | Seattle, WA $320k–$485k/yr Published 2 weeks ago
Flexible on stack
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Coreweave

Join CoreWeave as an Applied AI Engineer to enhance the performance of our inference platform through benchmarking and optimization.

Coreweave Bellevue, WA/ San Francisco, CA/ Sunnyvale, CA $188k–$275k/yr Published 6 months ago
Flexible on stack
baseten

Join Baseten as a Post-Training Research Scientist to advance AI research and collaborate on impactful projects.

baseten San Francisco Published 6 months ago
baseten

Join Baseten as a Post-Training Research Engineer to build in-house tooling for efficient and high-quality machine learning models.

baseten San Francisco Published 5 months ago
Inferact

Join Inferact as a TPU performance engineer to optimize vLLM for Google TPUs, enhancing AI inference performance.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Gamma

Build and scale backend systems for millions of users at Gamma, focusing on performance and reliability in a collaborative environment.

Gamma San Francisco $180k–$310k/yr Published 3 months ago
Flexible on stack
Perplexity AI

Join Perplexity AI as a Technical Staff member to enhance our AI inference engine with cutting-edge technologies.

Perplexity AI San Francisco Published 5 months ago
Flexible on stack