Luma AI Logo

Luma AI

Research Scientist / Engineer – Performance Optimization

Reposted 11 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in United States
180K-250K
Senior level
Remote
Hiring Remotely in United States
180K-250K
Senior level
As a Research Engineer, you'll optimize and implement models for data processing and inference, and work on performance enhancements in distributed systems using PyTorch and CUDA.
The summary above was generated by AI
About the Role

The Performance Optimization team at Luma is dedicated to maximizing the efficiency and performance of our AI models. Working closely with both research and engineering teams, this group ensures that our cutting-edge multimodal models can be trained efficiently and deployed at scale while maintaining the highest quality standards.

Responsibilities
  • Profile and optimize GPU/CPU/Accelerator code for maximum utilization and minimal latency

  • Write high-performance PyTorch, Triton, CUDA, deferring to custom PyTorch operations if necessary

  • Develop fused kernels and leverage tensor cores and modern hardware features for optimal hardware utilization on different hardware platforms

  • Optimize model architectures and implementations for distributed multi-node production deployment

  • Build performance monitoring and analysis tools and automation

  • Research and implement cutting-edge optimization techniques for transformer model

Experience
  • Expert-level proficiency in Triton/CUDA programming and GPU optimization

  • Strong PyTorch skills

  • Experience with PyTorch kernel development and custom operations

  • Proficiency with profiling tools (NVIDIA Nsight, torch profiler, custom tooling)

  • Deep understanding of transformer architectures and attention mechanisms

  • (Preferred) Experience with compilers/exporters such as torch.compile, TensorRT, ONNX, XLA

  • (Preferred) Experience optimizing inference workloads for latency and throughput

  • (Preferred) Experience with Triton compiler and kernel fusion techniques

  • (Preferred) Knowledge of warp-level intrinsics and advanced CUDA optimization

  • (Preferred) Background in compiler optimization or hardware-software co-design

Compensation

  • The pay range for this position in California is $180,000 - $250,000yr; however, base pay offered may vary depending on job-related knowledge, skills, candidate location, and experience. We also offer competitive equity packages in the form of stock options and a comprehensive benefits plan. 

Your applications are reviewed by real people.

Top Skills

C++
Cuda
PyTorch
Triton

Similar Jobs

An Hour Ago
Easy Apply
Remote
United States
Easy Apply
135K-155K
Expert/Leader
135K-155K
Expert/Leader
AdTech • Cloud • Digital Media • Marketing Tech • Analytics • Consulting
Lead business development for enterprise clients focusing on Google-based analytics solutions. Shape growth strategies, enhance relationships, and drive revenue across services while collaborating with various stakeholders.
Top Skills: Ga4Google Cloud PlatformGoogle Marketing PlatformSalesforce
An Hour Ago
In-Office or Remote
2 Locations
140K-192K Annually
Senior level
140K-192K Annually
Senior level
Computer Vision • Healthtech • Information Technology • Logistics • Machine Learning • Software • Manufacturing
The Senior Manager of Account Management will lead a team, driving retention and growth for enterprise clients while establishing consultative relationships and operational excellence.
Top Skills: SalesforceVitallyZendesk
An Hour Ago
Remote or Hybrid
New York, NY, USA
70K-85K Annually
Junior
70K-85K Annually
Junior
Productivity • Sales • Software
Responsible for generating qualified leads for monday Dev by prospecting and qualifying potential customers while managing CRM records and performance metrics.
Top Skills: Crm SoftwareSalesforce

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account