Nvidia

Deep Learning Performance Architect

Nvidia
Apply
7 days ago
Shanghai, China or Beijing, ChinaSenior
H1B Sponsor

Responsibilities

  • Develop highly optimized deep learning kernels for inference.
  • Perform performance optimization, analysis, tuning, modeling, profiling, debugging, and code optimization.
  • Collaborate with cross-functional teams in automotive, image understanding, and speech understanding.
  • Occasionally travel to conferences and customers for technical consultation and training.

Requirements

  • Master's degree, PhD, or equivalent experience in computer engineering, computer science and engineering, computer science, artificial intelligence, or a related discipline.
  • At least five years of relevant work experience.
  • Excellent C/C++ programming and software design skills.
  • Knowledge of CPU and GPU architecture, performance modeling, profiling, debugging, and code optimization.
  • GPU programming experience with CUDA or OpenCL is desired.
  • Python experience is a plus.

Tech Stack

Categories

Nvidia

About Nvidia

10,000+ employees

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Contact me