2 days ago
Austin, TX, USA or Santa Clara, CA, USAMid Level
H1B sponsor
Base Salary
$124k - $196k/yr
Responsibilities
- Research and prototype systems optimizations for deep learning models across high-level frameworks and low-level CUDA.
- Architect and optimize distributed computing systems from single-node execution through cluster-scale supercomputing.
- Design, implement, and optimize custom CUDA kernels for neural network architectures and workloads.
- Identify and resolve hardware-software performance bottlenecks in training and inference pipelines.
- Collaborate with AI researchers, hardware and software architects, compiler authors, kernel developers, and CUDA driver experts on systems and algorithm co-design.
- Develop profiling tools and runtime systems for new deep learning paradigms.
- Produce maintainable code that can transition into open-source releases, framework integrations, internal tools, or commercial products.
Requirements
- Bachelor's, master's, or PhD in computer science, computer engineering, electrical engineering, or a related field, or equivalent experience.
- At least 2 years of relevant industry experience or equivalent academic experience after degree achievement.
- Strong proficiency in C++ and Python.
- Strong fundamentals in deep learning, particularly transformers, and experience profiling and optimizing generative AI, vision, or diffusion models.
- Strong understanding of distributed computing, multi-node scaling, systems programming, computer architecture, and low-level performance optimization.
- Hands-on experience with GPU architectures, CUDA programming, kernel optimization, and workload profiling.
- Research background in machine learning systems or adjacent fields and demonstrated initiative across the technology stack.
- Preferred experience includes deep learning framework internals, communication libraries, distributed machine learning, low-precision arithmetic, deep learning compilers, ML systems, parallel simulation environments, and agentic AI systems.
Benefits
- Eligible for equity and benefits.
- Base salary range is 124,000 USD to 195,500 USD, determined by location, experience, and comparable employee pay.
- Applications accepted at least until October 3, 2026; the posting is for an existing vacancy.
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
