Together AI Logo

Together AI

Systems Research Engineer Intern - GPU Programming (Winter 2027)

Posted 12 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA
58-63 Hourly
Internship
In-Office
San Francisco, CA
58-63 Hourly
Internship
Develop and optimize GPU-accelerated kernels and algorithms for ML/AI applications. Collaborate with modeling, algorithm, hardware, and software teams on GPU and programming-model co-design, integrate optimized solutions, and apply performance profiling techniques to improve scalability and efficiency.
The summary above was generated by AI
Role Overview

As a Systems Research Engineer Intern specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. Working closely with the modeling and algorithm team, you will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.

This internship is based on-site at our San Francisco HQ, running through the Winter term from January to April.

Responsibilities
  • Optimize and fine-tune GPU code to achieve better performance and scalability
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems
  • Stay up-to-date with the latest advancements in GPU programming techniques and technologies
Requirements
  • Strong background in GPU programming and parallel computing, such as CUDA and/or Triton.
  • Knowledge of ML/AI applications and models
  • Knowledge of performance profiling and optimization tools for GPU programming
  • Excellent problem-solving and analytical skills
About Together AI

Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month.

Internship Program Details: 

Our internship program runs 12 to 14 weeks, giving you the opportunity to work alongside industry-leading engineers and researchers across multiple teams. This cohort's internship dates span January 4th to April 9th.

Compensation

We offer competitive compensation, housing stipends, and other competitive benefits. The estimated US hourly rate for this role is $58 to $70 an hour. Our hourly rates are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge.

Equal Opportunity

Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.

Please see our privacy policy at https://www.together.ai/privacy

Similar Jobs

12 Minutes Ago
Remote or Hybrid
159K-231K Annually
Senior level
159K-231K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Develop data-centric machine learning solutions for autonomous driving. Design dataset mixtures, curation and mining methods, model-training recipes, scaling studies, and offline evaluations using real and synthetic driving data. Apply self-supervised learning, imitation learning, reinforcement learning, and foundation-model fine-tuning. Train models across large multi-GPU and multi-node datasets, diagnose data-related failures, and collaborate on deploying models into onboard driving systems.
Top Skills: SparkNumpyPandasPythonPyTorchSQL
12 Minutes Ago
Hybrid
153K-234K Annually
Senior level
153K-234K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Build and operate production software, backend services, data pipelines, infrastructure, and workflow automation for autonomous vehicle simulation testing. Automate test creation, execution, validation, monitoring, and maintenance while integrating evaluation systems and partner tools. Lead technical design, code reviews, architecture, and cross-functional execution across autonomy, safety, systems, data, and infrastructure teams. Improve reliability, observability, quality, and scalability of simulation testing through cloud-native engineering and analytics.
Top Skills: APIsCi/CdCloud ComputingContainerizationData PipelinesDistributed SystemsMicroservicesNoSQLObservabilityPythonSQLTime-Series AnalysisWorkflow Orchestration
12 Minutes Ago
Remote or Hybrid
129K-198K Annually
Senior level
129K-198K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Designs and leads scalable cloud applications, virtualization platforms, and automation for embedded software development and co-simulation. Responsibilities include architecting multi-region cloud infrastructure, developing deployment and maintenance automation, integrating regression testing workflows into CI/CD, and supporting cybersecurity, scalability, resiliency, and cost optimization. The role serves as a subject matter expert, leads cross-functional integrations, collaborates with suppliers, and presents technical solutions to leadership and external organizations.
Top Skills: AutovalAWSAzureAzure Event GridAzure Event HubsAzure Service BusBicepCC++CaplCi/CdClaudeCursorDockerDspace SystemdeskEcsGCPGerritGitGithub ActionsGithub CopilotGleanIntrepid Vehicle SpyJavaJenkinsKubernetesMicronautMicrosoft 365 CopilotPythonQuarkusSingularitySpring BootSystemcTerraformVector CanapeVector CanoeVeos

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account