20 hours ago
Shanghai, China +2 moreIntern
Responsibilities
- Integrate new communication-library features into AI frameworks from proof of concept through performance analysis and production.
- Analyze AI workloads and frameworks to identify multi-GPU communication requirements and opportunities.
- Collaborate with teams developing and using current AI models.
- Author custom communication and fused compute-communication kernels for high performance on NVIDIA platforms.
- Research techniques to achieve GPU performance goals.
- Build fault-tolerant and elastic solutions for large-scale or dynamic AI workloads.
- Collaborate with a distributed team across multiple time zones.
Requirements
- Pursuing an M.S. or Ph.D. in computer engineering, computer science, or electrical engineering.
- Strong background in communication, kernel authoring, and/or AI training or inference.
- Rapid prototyping and development experience with Python, C++, CUDA, or related DSLs such as Triton and cuTe.
- Solid understanding of large language models and parallelism.
- Adaptability, willingness to learn, and effective communication skills.
Benefits
- NVIDIA offers competitive salaries and a comprehensive benefits package for employees and their families.
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
