Nvidia

Software Engineer, LLM Inference

Nvidia
Apply
16 hours ago
Shanghai, ChinaMid Level

Responsibilities

  • Develop robust inference software that scales across multiple platforms.
  • Perform performance analysis, optimization, and tuning.
  • Follow artificial intelligence research developments and update TensorRT and TensorRT Edge LLM features.
  • Collaborate with software, research, and product teams to guide machine learning inference direction.
  • Provide highly responsive customer support and communication as needed.

Requirements

  • Master's degree or higher in Computer Engineering, Computer Science, Applied Mathematics, or a related computing-focused field, or equivalent experience.
  • 4+ years of relevant software development experience.
  • Excellent C/C++ programming and software design skills, including debugging, performance analysis, and test design.
  • Curiosity about artificial intelligence and awareness of developments in deep learning, LLMs, and generative models.
  • Experience with deep learning frameworks such as PyTorch.
  • Ability to work proactively and independently.
  • Excellent written and oral English communication skills.
  • Strong customer communication skills.

Tech Stack

Categories

Nvidia

About Nvidia

10,000+ employees

Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.

Contact me