Fluidstack Logo

Fluidstack

Head of Infrastructure

Reposted 9 Days Ago
Be an Early Applicant
Remote
33 Locations
Senior level
Remote
33 Locations
Senior level
Lead the deployment of GPU supercomputers globally, managing supply chain, sourcing, and building a deployment team while ensuring timely delivery to customers.
The summary above was generated by AI
About FluidStack

Fluidstack is the AI Cloud Platform. We build GPU supercomputers for top AI labs, governments, and enterprises. Our customers include Mistral, Poolside, Black Forest Labs, Meta, and more.

Our team is small, highly motivated, and focused on providing a world class supercomputing experience. We put our customers first in everything we do, working hard to not just win the sale, but to win repeated business and customer referrals.

We hold ourselves and each other to high standards. We expect you to care deeply about the work you do, the products you build, and the experience our customers have in every interaction with us.

You must work hard, take ownership from inception to delivery, and approach every problem with an open mind and a positive attitude. We value effectiveness, competence, and a growth mindset.

About the Role

FluidStack is hiring a Head of Infrastructure to lead deployments of 10,000+ GPU supercomputers globally. Reporting directly to the co-founder/president, you will lead our engagements with OEMs, data centers, ISPs, and all relevant infrastructure partners. You will own sourcing, procurement, and be responsible for the timely deployment of some of the largest GPU supercomputers in the world.

You will be in charge of building a world-class deployment team to deliver multi-thousand GPU clusters in a matter of days. This is a unique opportunity to build the infrastructure function from the ground up in an extremely fast-paced environment, as well as a chance to shape the future of AI.

You are expected to have exceptional technical and interpersonal communication skills. You should be able to concisely and accurately share knowledge, in both written and verbal form, with teammates, customers, and suppliers.

Focus
  • You are responsible for the entire supply chain, from the original sourcing of each individual component up to the handover of the burned-in cluster to our customer.

  • You own relationships with OEMs, with the responsibility of continuously improving delivery timelines and costs across the entire supply chain.

  • You will design and build AI clusters, combining your deep knowledge in the area with high-level customer requirements and our past deployment learnings.

  • You are responsible for sourcing additional data center capacity to support our rapidly scaling supercomputer business.

  • You will hire and manage a small but effective “swat team” of deployment engineers, responsible for the world’s fastest setup, burn-in, and delivery of reliable GPU clusters.

  • You will partner with engineering, sales, finance, and legal to always have infrastructure ready one step ahead of our customer needs.

  • You will be required to travel significantly, be it to conferences/trade shows, data centers, customer sites, OEM factories, etc.

About You

An ideal candidate meets at least the following requirements:

  • 3+ years of related experience deploying GPU clusters; 5+ years deploying infrastructure at global scale.

  • On-site experience physically setting up hardware in data centers.

  • Strong relationships with, and history of procuring from, compute and storage OEMs, data centers, ISPs, and others.

  • Experience with InfiniBand or RoCE networking deployments.

  • An understanding of the software that runs on these clusters: Kubernetes/SLURM, PyTorch/Jax, etc.

  • Extreme attention to detail and ability to prioritize and deliver in a fast-paced environment.

  • Highly proactive with an extreme sense of urgency and ownership.

  • Strong engineering background, preferably in Computer Engineering, Electrical Engineering, Computer Science, Software Engineering, Math, Operations Research, Logistics, or similar fields.

Exceptional candidates have one or more of the following experiences:

  • You have designed, built, and operated a 4000+ GPU cluster.

  • You have build tooling to manage bare metal hardware via MaaS, Netbox, or similar tooling.

  • You have deployed and managed petabyte scale all-flash storage systems, including DDN, VAST, and/or Weka; or Ceph, LUSTRE, or similar open source tools.

Benefits
  • Competitive total compensation package (salary + equity).

  • Retirement or pension plan, in line with local norms.

  • Health, dental, and vision insurance.

  • Generous PTO policy, in line with local norms.

  • All-paid business travel to hardware and data center shows around the world.

Top Skills

Gpu Clusters
Infiniband
Jax
Kubernetes
PyTorch
Roce
Slurm

Similar Jobs

13 Days Ago
Remote
30 Locations
Senior level
Senior level
Big Data • Cloud • Digital Media • Machine Learning • Mobile • Software • Industrial
The Senior Implementation Consultant focuses on maximizing Autodesk software value for enterprise customers through process development, training, and technical guidance in the AEC sector.
Top Skills: Autodesk
13 Days Ago
Remote
Vari, GRC
Entry level
Entry level
Fintech • Payments • Financial Services
Red Cell Partners invites individuals with a strong entrepreneurial spirit and a belief in innovation to submit their information for future job openings.
20 Days Ago
Easy Apply
Remote
28 Locations
Easy Apply
Junior
Junior
Cloud • Security • Software • Cybersecurity • Automation
The Executive Business Administrator will support Legal and Corporate Affairs, manage executives' calendars, organize travel, plan events, and handle recruitment tasks.
Top Skills: Google WorkspaceNavanSlackZoom

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account