Fluidstack Logo

Fluidstack

Forward Deployed Engineer

Reposted 8 Days Ago
Be an Early Applicant
Remote
37 Locations
Mid level
Remote
37 Locations
Mid level
As a Forward Deployed Engineer, you will assist customers by deploying GPU clusters, optimizing infrastructure, migrating data, and debugging issues while supporting customer needs.
The summary above was generated by AI
About Fluidstack

Fluidstack is the AI Cloud Platform. We build GPU supercomputers for top AI labs, governments, and enterprises. Our customers include Mistral, Poolside, Black Forest Labs, Meta, and more.

Our team is small, highly motivated, and focused on providing a world class supercomputing experience. We put our customers first in everything we do, working hard to not just win the sale, but to win repeated business and customer referrals.

We hold ourselves and each other to high standards. We expect you to care deeply about the work you do, the products you build, and the experience our customers have in every interaction with us.

You must work hard, take ownership from inception to delivery, and approach every problem with an open mind and a positive attitude. We value effectiveness, competence, and a growth mindset.

About the Role

Forward Deployed Engineers are a customer’s trusted advisor and technical counterpart throughout the lifecycle of their AI workloads.

FDEs work across multiple areas of the organization, including software engineering, SRE, infrastructure engineering, networking, solutions architecture, and technical support.

FDEs are expected to have strong technical and interpersonal communication skills. You should be able to concisely and accurately share knowledge, in both written and verbal form, with teammates, customers, and partners.

As an FDE, your responsibilities are aligned with the success of our customers, and you’ll work side-by-side with them to deeply understand their workloads and solve some of their most pressing challenges. A day’s work may include:

  • Deploying clusters of 1,000+ GPUs using custom written playbooks; modifying these tools as necessary to provide the perfect solution for a customer.

  • Validating correctness and performance of underlying compute, storage, and networking infrastructure, and working with providers to optimize these subsystems.

  • Migrating petabytes of data from public cloud platforms to local storage, as quickly and cost effectively as possible.

  • Debugging issues anywhere in the stack, from “this server’s fan is blocked by a plastic bag” to “optimizing S3 dataloaders from buckets in different regions”.

  • Building internal tooling to decrease deployment time and increase cluster reliability, including automation where the customer benefits clearly outweigh the implementation overhead.

  • Supporting customers as part of an on-call rotation, up to two weeks per month.

Focus
  • A customer-centric attitude, an accountability mindset, and a bias to action.

  • A track record of shipping clean, well-documented code in complex environments.

  • An ability to create structure from chaos, navigate ambiguity, and adapt to the dynamic nature of the AI ecosystem.

  • Strong technical and interpersonal communication skills, a low ego, and a positive mental attitude.

An ideal candidate meets at least the following requirements:

  • 2+ years of SWE, SRE, DevOps, Sysadmin, and/or HPC engineering experience.

  • Great verbal and written communication skills in English.

  • Experience deploying and operating Kubernetes and/or SLURM clusters.

  • Experience in writing Go, Python, Bash.

  • Experience using Ansible, Terraform, and other automation or IAC tools.

  • Strong engineering background, preferably in Computer Science, Software Engineering, Math, Computer Engineering, or similar fields.

Exceptional candidates have one or more of the following experiences:

  • You have built and operated an AI workload at 1000+ GPU scale.

  • You have built multi-tenant, hyperscale Kubernetes based services.

  • You have physically deployed infrastructure in a datacenter, managed bare metal hardware via MaaS or Netbox, etc.

  • You have deployed and managed multi-tenant InfiniBand or RoCE networks.

  • You have deployed and managed petabyte scale all-flash storage systems, including DDN, VAST, and/or Weka; or Ceph, LUSTRE, or similar open source tools.

Interview Process

If your application passes the screening stage, you will be invited to a 15 minute hiring manager call. If you clear the initial phone interview, you will enter the main process, which consists of three 45 minute interviews: a technical deep dive, customer communications and debugging session, and culture fit interview.

Our goal is to finish the main process within one week. All interviews will be conducted virtually.

Benefits
  • Competitive total compensation package (cash + equity).

  • Retirement or pension plan, in line with local norms.

  • Health, dental, and vision insurance.

  • Generous PTO policy, in line with local norms.

  • Fluidstack is remote first, but has offices in key hubs. For all other locations, we provide access to WeWork.

Top Skills

Ansible
Bash
Go
Kubernetes
Python
Slurm
Terraform

Similar Jobs

15 Days Ago
Easy Apply
Remote
Hybrid
28 Locations
Easy Apply
Senior level
Senior level
Big Data • Cloud • Software • Database
The Senior Staff Forward Deployed AI Engineer will facilitate application modernization projects using MongoDB technologies, provide feedback for product development, and collaborate with teams to enhance customer success in complex software implementations.
Top Skills: C#Generative AiJavaJavaScriptMongoDBMs Sql ServerOraclePostgresSQLSybase
15 Days Ago
Easy Apply
Remote
Hybrid
28 Locations
Easy Apply
50K-130K
Senior level
50K-130K
Senior level
Big Data • Cloud • Software • Database
The Senior Forward Deployed AI Engineer will work on application modernization projects, leveraging generative AI technologies and collaborating across teams to enhance MongoDB's product capabilities.
Top Skills: C#JavaJavaScriptJSONMongoDBMs Sql ServerOraclePostgresSybase
15 Days Ago
Remote
Moschato, GRC
Junior
Junior
Information Technology
The Associate Software Engineer is responsible for maintaining product architecture, translating business needs into technical specifications, and participating in design and troubleshooting activities.
Top Skills: .Net 6.Net Core 3.1Angular 12C#JavaScriptMs Sql Server 2012SQL

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account