Direct from source · No middlemen

Gpu Jobs in San Francisco

61 open positions · Updated 11 hours ago

Average salary (USD/year): 241.5k–374.8k/yr
61 roles Senior 22, Staff 14, Mid 12, Lead 5, Director 3, Principal 2, C-level 1, Junior 1 1–10 yrs experience Livingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA / San Francisco, CA +20 more Full-time newest 11 hours ago

61 Gpu roles across 19 companies, most in AI/ML; 3% fully remote; typical advertised salary $207k

Work arrangement: 2 fully remote · 24 hybrid · 7 on-site · 28 not stated

Advertised salaries: p25 $190k · median $207k · p75 $290k (from 41 disclosed annual salaries, USD)

Counts are open roles Joblaze currently tracks on company career pages for this exact skill/location; salaries are advertised minimums, annual, converted to USD.

Showing 20 of 61 positions

Search with filters →
Fireworks AI

Drive adoption of Fireworks AI within ambitious GenAI startups, engaging with technical teams to create value.

Fireworks AI San Francisco, United States Published 1 day ago
baseten

Join Baseten to lead capacity planning and operations in a fast-paced AI environment.

baseten San Francisco, United States Published 2 days ago
baseten

Lead relationships and market intelligence across hyperscalers and strategic neoclouds in a senior role at Baseten.

baseten San Francisco, United States Published 2 days ago
Anthropic

Join Anthropic as a Performance Engineer to optimize the inference engine for AI systems at scale.

Anthropic San Francisco, CA | New York City, NY $350k–$850k/yr Published 2 days ago
Flexible on stack
Descript

Join Descript as a Senior Software Engineer to own and enhance our infrastructure platform, impacting AI model training and deployment.

Descript San Francisco, CA or Remote, US $220k–$292k/yr Published 1 week ago
Flexible on stack
baseten

Join Baseten as a hands-on Operations Manager to optimize the health and utilization of our GPU fleet in a fast-growing AI company.

baseten San Francisco Published 1 week ago
baseten

Lead the delivery of on-premises data center builds and GPU cloud programs in a high-growth AI infrastructure company.

baseten San Francisco Published 1 week ago
Together AI

Join Together AI as a Sales Development Engineer to engage with clients and drive AI-driven business opportunities.

Together AI San Francisco $90k–$150k/yr Published 10 months ago
Mithril

Join Mithril as a Site Reliability Engineer to enhance the stability and performance of our global GPU orchestration platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack 70% coding
Perplexity AI

Join Perplexity AI as a Strategic Finance Lead to optimize GPU compute investments and drive capacity decisions.

Perplexity AI San Francisco Published 1 week ago
Reflection AI

Lead the Compute Platform team at Reflection AI, focusing on multi-cloud scheduling and GPU deployments while mentoring a team of systems engineers.

Reflection AI San Francisco, CA Published 4 weeks ago
Flexible on stack
Reflection AI

Design and operate large-scale GPU infrastructure for model inference and mid-training workloads at Reflection AI.

Reflection AI San Francisco, CA Published 5 months ago
Flexible on stack
Chai Discovery

Join Chai Discovery as a Software Engineer to optimize AI models for drug discovery in a fast-paced, innovative environment.

Chai Discovery San Francisco office Published 9 months ago
Postman

Lead AI reliability engineering efforts to ensure the performance and scalability of Postman's AI-powered API services.

Postman San Francisco, California, United States $256k–$276k/yr Published 11 months ago
baseten

Join Baseten as a Software Engineer to drive model performance systems at the intersection of HPC and LLM engineering.

baseten San Francisco Published 8 months ago
Flexible on stack
baseten

Join Baseten as a GPU Kernel Engineer to optimize high-performance GPU kernels for cutting-edge AI applications.

baseten San Francisco Published 1 year ago
Flexible on stack 70% coding
Inferact

Join Inferact as an AMD GPU performance engineer to optimize vLLM for the AMD accelerator ecosystem.

Inferact San Francisco $200k–$400k/yr Published 2 months ago
Flexible on stack
Inferact

Join Inferact as a performance engineer to optimize vLLM, the fastest AI inference engine, working with cutting-edge hardware.

Inferact San Francisco $200k–$400k/yr Published 7 months ago
Flexible on stack
Inferact

Join Inferact as a cluster administration engineer to manage high-performance GPU compute infrastructure for AI inference.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack
Inferact

Lead the engineering organization at Inferact to develop systems for vLLM, focusing on GPU performance and ML systems optimization.

Inferact San Francisco Published 1 month ago
Page 1 of 4 Next →