Direct from source · No middlemen

Gpu Jobs in San Francisco

61 open positions · Updated 2 hours ago

Average salary (USD/year): 241.5k–374.8k/yr
61 roles Senior 22, Staff 14, Mid 12, Lead 5, Director 3, Principal 2, C-level 1, Junior 1 1–10 yrs experience Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA / San Francisco, CA +20 more Full-time newest 2 hours ago

61 Gpu roles across 19 companies, most in AI/ML; 3% fully remote; typical advertised salary $207k

Work arrangement: 2 fully remote · 24 hybrid · 7 on-site · 28 not stated

Advertised salaries: p25 $190k · median $207k · p75 $290k (from 41 disclosed annual salaries, USD)

Counts are open roles Joblaze currently tracks on company career pages for this exact skill/location; salaries are advertised minimums, annual, converted to USD.

Showing 20 of 61 positions

Search with filters →
Fireworks AI

Drive adoption of Fireworks AI within ambitious GenAI startups, engaging with technical teams to create value.

Fireworks AI San Francisco, United States Published 16 hours ago
baseten

Join Baseten to lead capacity planning and operations in a fast-paced AI environment.

baseten San Francisco, United States Published 1 day ago
baseten

Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.

baseten San Francisco, United States Published 1 day ago
Anthropic

Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.

Anthropic San Francisco, CA | New York City, NY $350k–$850k/yr Published 2 days ago
Flexible on stack
Descript

Join Descript as a Senior Software Engineer to own and enhance our infrastructure platform, impacting AI model training and deployment.

Descript San Francisco, CA or Remote, US $220k–$292k/yr Published 1 week ago
Flexible on stack
baseten

Join Baseten as a hands-on Operations Manager to optimize the health and utilization of our GPU fleet in a fast-growing AI company.

baseten San Francisco Published 1 week ago
baseten

Lead the delivery of on-premises data center builds and GPU cloud programs in a high-growth AI infrastructure company.

baseten San Francisco Published 1 week ago
Together AI

Join Together AI as a Sales Development Engineer to engage with clients and drive AI-driven business opportunities.

Together AI San Francisco $90k–$150k/yr Published 10 months ago
Mithril

Join Mithril as a Site Reliability Engineer to enhance the stability and performance of our global GPU orchestration platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack 70% coding
Perplexity AI

Join Perplexity AI as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions.

Perplexity AI San Francisco Published 1 week ago
Reflection AI

Lead the Compute Platform team at Reflection AI, focusing on multi-cloud scheduling and GPU deployments while mentoring a team of systems engineers.

Reflection AI San Francisco, CA Published 4 weeks ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
Chai Discovery

Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.

Chai Discovery San Francisco office Published 9 months ago
Postman

Lead AI reliability engineering efforts to ensure the performance and scalability of Postman's AI-powered API services.

Postman San Francisco, California, United States $256k–$276k/yr Published 11 months ago
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 1 month ago
Page 1 of 4 Next →