Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AITogether AISan Francisco, California, United States
Join the waitlist to applySave this job

Invite-only right now: save jobs, track applications, build tailored resumes

Already have an account? Log in

Posted

9/18/2026

Employment

Intern

Range

$58 - $63/hr

Work style

On-site

AI documents

Powered by AI

Jigup writes these against this posting once you are in, using the profile you build once.

Career path

See where this role leads

Jigup maps the next three moves from a job like this one, with the titles and the skills each step asks for.

Join the waitlist

AI summary

Core responsibilities

You will develop and optimize GPU-accelerated kernels and algorithms while collaborating with cross-functional teams to integrate these solutions into AI systems. Additionally, you will contribute to the co-design of efficient GPU architectures and stay current with industry advancements.

Requirements overview

Candidates must have a strong background in GPU programming and parallel computing, specifically with CUDA or Triton. Proficiency in ML/AI applications and experience with performance profiling and optimization tools are also required.

Key skills

GPU ProgrammingParallel ComputingCUDATritonMachine LearningArtificial IntelligencePerformance ProfilingOptimizationAlgorithm DesignModel Architecture

Resume keywordsJigup Pro

This job lists resume keywords

The terms this posting uses, pulled out so you can mirror them in your resume. Jigup Pro members see the list on the job board.

Join the waitlist

About Together AI

Industry

Software Development

Employees

431

Type

Privately Held

Size

201-500 employees

Together AI is the AI Native Cloud, purpose-built for AI engineers and researchers with a full suite of tooling across inference, model shaping, and pre-training. AI natives can use Together AI as a full-stack AI platform — from a high- performance inference engine built for reliable and fast scaling to on-demand GPU clusters and massive-scale AI factories. Together AI continuously pushes the frontier forward by productizing cutting-edge research from our world-leading AI systems research team. By combining research velocity with production-grade infrastructure, we enable companies to reliably scale AI-native applications as fast as the field evolves. Trusted by leading AI natives like Cursor, Decagon, Eleven Labs, AI21, Hedra, and Cartesia, as well as SaaS innovators such as Salesforce, Zoom, and Zomato, Together AI powers the next generation of AI-native applications.

View company page

Job categories

TechnologyEngineeringSoftwareScience & ResearchData & Analytics

Description

Role Overview As a Systems Research Engineer Intern specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. Working closely with the modeling and algorithm team, you will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation. This internship is based on-site at our San Francisco HQ, running through the Winter term from January to April. Responsibilities Optimize and fine-tune GPU code to achieve better performance and scalability Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems Stay up-to-date with the latest advancements in GPU programming techniques and technologies Requirements Strong background in GPU programming and parallel computing, such as CUDA and/or Triton. Knowledge of ML/AI applications and models Knowledge of performance profiling and optimization tools for GPU programming Excellent problem-solving and analytical skills About Together AI Together AI, the AI Native Cloud, is purpose-built for AI engineers. AI application developers get high-performance inference that scales reliably, fine-tuning and reinforcement learning for creating frontier-level specialized models, and pre-training at massive scale for fully custom intelligence, all around a marketplace of leading open models that teams can run, adapt, and own. Trusted by Cursor, Decagon, ElevenLabs, Salesforce, and Zoom, Together serves 400+ trillion tokens a month. Internship Program Details: Our internship program runs 12 to 14 weeks, giving you the opportunity to work alongside industry-leading engineers and researchers across multiple teams. This cohort's internship dates span January 4th to April 9th. Compensation We offer competitive compensation, housing stipends, and other competitive benefits. The estimated US hourly rate for this role is $58 to $63 an hour. Our hourly rates are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. Equal Opportunity Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. Please see our privacy policy at https://www.together.ai/privacy

Requirements

  • GPU Programming
  • Parallel Computing
  • CUDA
  • Triton
  • Machine Learning
  • Artificial Intelligence
  • Performance Profiling
  • Optimization
  • Algorithm Design
  • Model Architecture

Benefits

  • Housing stipends
  • Competitive compensation

More entry level jobs in San Francisco, CA

All jobs in San Francisco, CA

More jobs at Together AI

Ready to apply?

Jigup is invite-only right now. Join the waitlist to save jobs, track applications, and build tailored resumes.

Join the waitlist