"gpu capacity" Jobs

87 open tech roles matching “gpu capacity”, taken straight from company career pages — not reposted from other job boards. Most in demand right now: Kubernetes, AI/ML, Python. Every listing is re-checked daily and closed roles are removed.

Showing 20 of 87 results

Twelve Labs

Lead and build a new team focused on developing Jockey Core, a reasoning LLM for video understanding at Twelve Labs.

Twelve Labs Seoul, South Korea Published 1 month ago
Flexible on stack
Mithril

Join Mithril as a Software Engineer to build and maintain backend systems for a cutting-edge AI infrastructure platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack
Atoms

Join Atoms as a Staff Machine Learning Infrastructure Engineer to design and build large-scale ML training infrastructure for autonomous transport models.

Atoms San Francisco, CA $224k–$280k/yr Published 2 months ago
Flexible on stack
Anthropic
Anthropic San Francisco, CA | New York City, NY | Seattle, WA $290k–$365k/yr Published 6 months ago
Fal

Lead data center operations at fal, overseeing strategy, capacity, and build-out for a high-performance GPU fleet.

Fal San Francisco Published 1 month ago
baseten

Join Baseten as a Product Manager to shape the future of AI infrastructure and enhance production inference capabilities.

baseten San Francisco Published 5 months ago
Perplexity

Join Perplexity as a technical program manager to drive the core inference platform and coordinate between model providers and engineering teams.

Perplexity San Francisco Published 1 week ago
Perplexity AI

Join Perplexity AI as a technical program manager to drive the core inference platform and coordinate across teams and model providers.

Perplexity AI San Francisco Published 1 week ago
baseten

Lead a team of cloud platform engineers to build scalable and reliable infrastructure for AI products at Baseten.

baseten San Francisco Published 3 months ago
Flexible on stack Heavy meetings
Together AI

Join Together AI as a Manager of Infrastructure Strategy & Operations to drive analytical decisions in scaling compute infrastructure.

Together AI San Francisco $220k–$260k/yr Published 3 months ago
Coreweave

Drive new business opportunities in AI infrastructure as a Principal Solution Specialist at CoreWeave.

Coreweave San Francisco, CA / Seattle, WA $198k–$264k/yr Published 3 weeks ago
Flexible on stack
baseten

Lead the emerging clouds and international coverage efforts at Baseten, a rapidly growing AI company.

baseten San Francisco Published 3 days ago
Coreweave
Coreweave Livingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA $165k–$242k/yr Published 4 months ago
baseten

Lead and mentor a team of Forward Deployed Engineers to optimize LLM inference workloads for Baseten customers.

baseten San Francisco Published 4 months ago
Flexible on stack Heavy meetings
Harvey AI
Harvey AI San Francisco $231k–$340k/yr Published 4 months ago
Mithril

Join Mithril as a Site Reliability Engineer to enhance the stability and performance of our global GPU orchestration platform.

Mithril Palo Alto / San Francisco Bay Area $170k–$230k/yr Published 4 months ago
Flexible on stack 70% coding
Anthropic

Join Anthropic as a Performance Engineer to optimize AI inference systems for throughput, latency, reliability, and correctness.

Anthropic San Francisco, CA | New York City, NY | Seattle, WA $350k–$850k/yr Published 3 months ago
Flexible on stack
Inferact

Join Inferact as a Site Reliability Engineer to enhance the reliability and performance of AI inference systems at scale.

Inferact San Francisco $200k–$400k/yr Published 3 weeks ago
Flexible on stack AI-first team
Harvey AI
Harvey AI San Francisco $193.4k–$290k/yr Published 4 months ago
Fal

Lead the technical accounting agenda at fal, focusing on complex compute cost accounting and revenue recognition in a growing AI company.

Fal San Francisco Published 2 months ago
AI-first team