← Back to results

Head of Data Center Operations

Lead data center operations at fal, overseeing strategy, capacity, and build-out for a high-performance GPU fleet.

Location
San Francisco
Compensation
Not disclosed
Level
director
Type
full time

Posted by employer 1 month ago

First seen on Joblaze 1 week ago

Last verified on the company career page 3 days ago

Apply at Fal → Save job Scanned from fal.ai

Requirements

Experience
10+ years

Not disclosed in this posting: compensation, work arrangement, visa sponsorship.

Joblaze summary

The Head of Data Center Operations at fal is responsible for overseeing the entire lifecycle of data center operations, from planning and design to deployment and steady-state management of GPU infrastructure. This role requires extensive experience in large-scale data center operations, particularly in high-density environments, and a strategic mindset to collaborate with engineering and finance teams. Ideal candidates will have a strong background in leading multi-site operations and managing vendor relationships. As fal scales its infrastructure to support the growing generative media market, this position plays a critical role in ensuring operational efficiency and reliability.

Joblaze insights

Quick facts

How much experience is required?
At least 10 years of relevant experience for this Head of Data Center Operations role.
What's the tech stack?
Joblaze extracted these technologies from the posting: AI/ML, Data Center, GPU, infrastructure.
What seniority level is this role?
Fal targets director candidates for this position.
Is this full-time or contract?
Full-time for this Head of Data Center Operations role at Fal.

From the original posting

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.

Own how fal runs its data centers. You'll lead data-center operations end to end, strategy, capacity, build-out, and steady-state, for the GPU fleet that powers every fal inference. You'll operate through our on-site lead and technicians while staying close to engineering and leadership, and help shape the compute organization as we scale.

WHAT YOU'LL OWN

  • Own the full DC operations lifecycle — planning, design, build-out, deployment, and steady-state operations across our sites.

  • Set infrastructure operations strategy — capacity, cost, reliability, and scale, with clear OKRs/KPIs, in lockstep with engineering, finance, and leadership.

  • Run operations through the on-site team — direct our site lead and technicians, own incident and escalation management, and drive uptime and MTTR across sites.

  • Partner with Engineering, Network, and Capacity Planning to execute infrastructure expansion while optimizing power, cooling, rack density, and deployment schedules.

  • Develop operational processes, documentation, and vendor governance as fal continues to scale its global infrastructure footprint.

WHAT YOU BRING

  • 10+ years in data center / infrastructure operations including senior leadership (Director level) owning multi-site or global operations.

  • 10MW+ facility experience — you've owned operations for a large-scale, mission-critical data center.

  • Full-lifecycle experience — planning and build-out through steady-state — ideally for high-density GPU / accelerator (AI inference) environments.

  • Strategic and cross-functional range — you partner naturally with engineering, finance, and business leadership on capacity, cost, and scale.

  • Command of commercial / vendor relationships and colocation / leased-space operations.

BONUS POINTS

  • You've stood up new data-center capacity from the ground up.

  • Experience operating high-density LPU/GPU or liquid-cooled environments.

  • You've built and scaled DC operations teams and the pipeline that feeds them.

U.S. EQUAL EMPLOYMENT OPPORTUNITY INFORMATION:

fal provides equal employment opportunities to applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or any other classification protected by applicable law.

Similar positions

Fal
Fal
Software Engineer, Platform
Fal · San Francisco