16 hours ago
Shanghai, ChinaMid Level
Responsibilities
- Develop robust inference software that scales across multiple platforms.
- Perform performance analysis, optimization, and tuning.
- Follow artificial intelligence research developments and update TensorRT and TensorRT Edge LLM features.
- Collaborate with software, research, and product teams to guide machine learning inference direction.
- Provide highly responsive customer support and communication as needed.
Requirements
- Master's degree or higher in Computer Engineering, Computer Science, Applied Mathematics, or a related computing-focused field, or equivalent experience.
- 4+ years of relevant software development experience.
- Excellent C/C++ programming and software design skills, including debugging, performance analysis, and test design.
- Curiosity about artificial intelligence and awareness of developments in deep learning, LLMs, and generative models.
- Experience with deep learning frameworks such as PyTorch.
- Ability to work proactively and independently.
- Excellent written and oral English communication skills.
- Strong customer communication skills.
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
