Fal logo

Head of Data Center Operations

Fal

San Francisco, CAFull-time$250–350K/yrPosted 2 days agoVerified open today

Get personalized insights for this job

Sign in to access AI-powered tools that help you stand out:

Resume evaluation
Tailored coaching
Mock interviews
Sign In to Get Started

At a glance

Compensation
$250–350K/yr
Location
San Francisco, CA
Schedule
Full-time
Work Authorization
Not specified

Job overview

Fal is hiring a Head of Data Center Operations. The Head of Data Center Operations will lead end-to-end data center operations, including strategy, capacity, build-out, and steady-state for the GPU fleet powering fal's AI inference. This role involves operating through on-site leads and technicians, collaborating with engineering and leadership, and shaping the compute organization as it scales. The position focuses on ensuring high-performance inference, orchestration, and observability for generative media products.

Key focus areas include Own the full DC operations lifecycle, Set infrastructure operations strategy, and Run operations through the on-site team.

Successful candidates bring 10+ Years In Data Center / Infrastructure Operations, Senior Leadership (Director Level) Owning Multi-site Or Global Operations, and 10MW+ Facility Experience. Important skills include Data Center Operations, Infrastructure Operations Strategy, Incident Management, Escalation Management, Uptime, and MTTR. Preferred (not required): New Data-center Capacity Stand-up, High-density LPU/GPU Operations, Liquid-cooled Environments Operations, and DC Operations Teams Building And Scaling.

Skills & qualifications

RequiredNice to have

Skills

Data Center OperationsInfrastructure Operations StrategyIncident ManagementEscalation ManagementUptimeMTTROperational Processes DevelopmentDocumentationVendor GovernanceStrategic ThinkingCollaborationCommercial RelationshipsVendor RelationshipsColocation OperationsLeased-Space OperationsNew Data-Center Capacity Stand-UpHigh-Density LPU/GPU OperationsLiquid-Cooled Environments OperationsDC Operations Teams Building and Scaling

Qualifications

10+ Years in Data Center / Infrastructure OperationsSenior Leadership (Director Level) Owning Multi-Site or Global Operations10MW+ Facility ExperienceFull-Lifecycle Experience

Full job description

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.

Own how fal runs its data centers. You'll lead data-center operations end to end, strategy, capacity, build-out, and steady-state, for the GPU fleet that powers every fal inference. You'll operate through our on-site lead and technicians while staying close to engineering and leadership, and help shape the compute organization as we scale.

WHAT YOU'LL OWN

  • Own the full DC operations lifecycle — planning, design, build-out, deployment, and steady-state operations across our sites.

  • Set infrastructure operations strategy — capacity, cost, reliability, and scale, with clear OKRs/KPIs, in lockstep with engineering, finance, and leadership.

  • Run operations through the on-site team — direct our site lead and technicians, own incident and escalation management, and drive uptime and MTTR across sites.

  • Partner with Engineering, Network, and Capacity Planning to execute infrastructure expansion while optimizing power, cooling, rack density, and deployment schedules.

  • Develop operational processes , documentation, and vendor governance as fal continues to scale its global infrastructure footprint.

WHAT YOU BRING

  • 10+ years in data center / infrastructure operations including senior leadership (Director level) owning multi-site or global operations.

  • 10MW+ facility experience — you've owned operations for a large-scale, mission-critical data center.

  • Full-lifecycle experience — planning and build-out through steady-state — ideally for high-density GPU / accelerator (AI inference) environments.

  • Strategic and cross-functional range — you partner naturally with engineering, finance, and business leadership on capacity, cost, and scale.

  • Command of commercial / vendor relationships and colocation / leased-space operations.

BONUS POINTS

  • You've stood up new data-center capacity from the ground up.

  • Experience operating high-density LPU/GPU or liquid-cooled environments.

  • You've built and scaled DC operations teams and the pipeline that feeds them.

Ready to take the next step in your career?