Together AI

Systems Research Engineer Intern - GPU Programming (Winter 2027)

Together AI
Apply
2 days ago
San Francisco, CA, USAIntern
H1B sponsor

Responsibilities

  • Optimize and fine-tune GPU code for improved performance and scalability.
  • Collaborate with cross-functional teams to integrate GPU-accelerated solutions into existing software systems.
  • Co-design GPU kernels and model architectures with modeling and algorithm teams.
  • Contribute to the co-design of efficient GPU architectures and programming models.
  • Stay current with advancements in GPU programming techniques and technologies.

Requirements

  • Strong background in GPU programming and parallel computing, such as CUDA and/or Triton.
  • Knowledge of ML/AI applications and models.
  • Knowledge of GPU performance profiling and optimization tools.
  • Excellent problem-solving and analytical skills.

Benefits

  • On-site role at the San Francisco headquarters.
  • Winter internship running from January through April, with cohort dates of January 4 through April 9.
  • Internship duration of 12 to 14 weeks.
  • Competitive compensation and housing stipend.
  • Other competitive benefits.

Categories

Together AI

About Together AI

201-500 employees

Together AI builds an AI-native cloud platform for developers, offering high-performance inference, fine-tuning/model shaping, and large-scale pre-training on on-demand GPU clusters with APIs and managed services. It emphasizes open-source models that teams can run and adapt, and also provides infrastructure for decentralized and scalable workloads. Founded in 2022 and headquartered in San Francisco, it is privately held and reports notable customers including Cursor, ElevenLabs, Salesforce, and Zoom.

Contact me