Vast.ai Logo

Vast.ai

Developer Relations Engineer

Posted 13 Days Ago
Be an Early Applicant
In-Office
Los Angeles, CA, USA
160K-200K Annually
Entry level
In-Office
Los Angeles, CA, USA
160K-200K Annually
Entry level
Deploys, fine-tunes, and serves open-source AI models on rented GPUs using tools such as vLLM and SGLang. Builds public example repositories, Docker templates, benchmarks, and live endpoints, then teaches through technical writing, videos, talks, livestreams, and community support. Reports product friction to engineering and product teams, represents Vast at hackathons and conferences, and serves as a highly credible technical voice across developer communities.
The summary above was generated by AI
About Vast.ai

Vast.ai runs one of the largest GPU marketplaces in the world. Over 20,000 GPUs, from RTX 4090s up to B300s, rented by more than 25,000 customers a month who train, fine-tune, and serve AI models on them. We are profitable, the team is flat, and we ship quickly. Your application goes straight to the hiring team.

The role

This is an engineering job that happens in public. You will be our heaviest user.

Most of your week goes to renting GPUs on Vast and building things on them. Serving open source models with vLLM and SGLang. Fine-tuning. Running ComfyUI pipelines. Keeping live inference endpoints up, including token endpoints on markets like OpenRouter. Writing the example repos, Docker templates, and benchmarks that other developers copy.

The rest of your week is spent showing your work. Short videos, technical write-ups, answers in Discord and on Reddit, hackathons. You are not building the Vast product. You are using it harder than any of our customers, in the open, and reporting back everything that breaks.

If you want a content calendar and a campaign plan, this is the wrong job. If you already spin up GPUs on weekends because you want to see whether the new model actually runs, keep reading.

What you'll do
  • Run real workloads on Vast every week. Deploy, fine-tune, and serve open source models, and keep live endpoints running.

  • Build public example repos, Docker templates, and benchmarks that developers can clone and use.

  • Turn what you build into teaching: guides, short videos, talks, livestreams.

  • Be the most technically credible voice in our Discord, on Reddit, and on X.

  • Represent Vast at hackathons and conferences, roughly one trip a month.

  • Report the friction, the breakage, and the missing docs straight to engineering and product.

What we need to see
  • You can take an open source model from Hugging Face to a running endpoint on rented GPUs with nobody helping you. Linux, Docker, Python, SSH.

  • You can debug the GPU stack. Driver and CUDA mismatches, out of memory errors, and multi-GPU configuration including NCCL and topology issues do not scare you.

  • Public proof of work you have shipped. A GitHub profile, technical writing, or live projects. We will ask you to walk us through code you wrote and explain the parts that went wrong.

  • You work without a roadmap and prefer it that way. Nobody here will hand you a task list.

  • You can explain a hard thing clearly, in writing and out loud.

Nice to have
  • Content or side projects that found a real audience.

  • Previous developer relations, community, or customer-facing work. Useful here, but it is a bonus and not the job.

  • vLLM or SGLang internals, distributed training with FSDP or DeepSpeed, CUDA.

  • Time at a cloud, GPU, AI infrastructure, or developer tools company.

First 90 days
  • Ship several public example projects on Vast, including at least one live serving endpoint.

  • Publish benchmarks or guides that developers actually use and pass around.

  • Become the recognizable technical voice in our community channels.

  • Represent Vast at a sponsored hackathon.

Compensation & benefits

$160,000 to $200,000 base, plus equity and bonus. Health, dental, vision, and life insurance. 401(k) with company match. Meals on site. Travel and conference budget. This role is on site in San Francisco or Los Angeles.

HQ

Vast.ai Los Angeles, California, USA Office

Los Angeles, CA, United States

Similar Jobs

Yesterday
In-Office
300K-350K Annually
Senior level
300K-350K Annually
Senior level
Artificial Intelligence • Information Technology
Build developer-facing tools, integrations, documentation, demos, and machine learning recipes for Inkling and Tinker. Collaborate directly with open source developers and researchers to debug issues, gather feedback, improve products, and shape the developer experience roadmap. Represent the company through technical writing, talks, tutorials, livestreams, conferences, and meetups. Requires strong Python engineering skills and hands-on experience fine-tuning, evaluating, or deploying large language models.
Top Skills: InklingLarge Language ModelsMultimodal ModelsOpen Source SoftwarePythonReinforcement Learning Fine-TuningSupervised Fine-TuningTinker
6 Days Ago
Remote or Hybrid
3 Locations
Entry level
Entry level
Artificial Intelligence • Software • Industrial • Generative AI
Build scientific advisory and technical enablement programs for enterprise and research partners adopting an AI materials platform. Engage computational chemists and materials scientists, provide technical guidance, lead workshops and demonstrations, create documentation, notebooks, APIs, and reference materials, and relay partner feedback to AI Research, Chemistry, Product, Marketing, and PR teams.
Top Skills: APIsAseComputational Chemistry ToolsMachine LearningMolecular Simulation SoftwarePythonPyTorchRdkit
7 Days Ago
Hybrid
188K-250K Annually
Entry level
188K-250K Annually
Entry level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Developers Relations technical practitioner who helps enterprise teams deploy production AI workloads on Lambda. Creates technical guides, demonstrations, open-source examples, talks, workshops, and videos; grows reach through professional networks, communities, customers, and partners; represents Lambda at conferences and podcasts; gathers field feedback; and collaborates with engineering, MLE, marketing, and go-to-market teams.
Top Skills: 1-Click ClustersAILambda CloudMlopsObservability

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account