Nvidia

Deep Learning Compiler Intern - 2027

Nvidia
Apply
20 hours ago
Shanghai, China or Beijing, ChinaIntern

Responsibilities

  • Design and implement the DSL and core compiler for a tile-aware GPU programming model.
  • Iterate on compiler architecture to optimize performance.
  • Investigate next-generation GPU architectures and develop solutions across the DSL and compiler stack.
  • Explore compiler and DSL designs for agentic coding workflows.
  • Analyze emerging AI/LLM workloads and integrate with AI/ML frameworks.

Requirements

  • Master's or PhD in computer engineering, computer science and engineering, computer science, AI, or a related discipline, or equivalent experience.
  • Excellent C/C++ programming and software engineering skills.
  • Strong understanding of computer architecture and problem abstraction and resolution methodologies.
  • Compiler experience with MLIR, TVM, Triton, or LLVM is desired.
  • GPU architecture knowledge and fast kernel programming skills are a plus.
  • Knowledge of LLM algorithms, an HPC domain, multi-GPU distributed communication, or agentic coding workflows is a plus.
  • ACM background and excellent oral English communication are pluses.

Benefits

  • NVIDIA offers competitive salaries and a comprehensive benefits package for employees and their families.

Tech Stack

Categories

Nvidia

About Nvidia

10,000+ employees

Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.

Contact me