2 days ago
Santa Clara, CA, USAMid Level
H1B Sponsor
Base Salary
$124k - $196k/yr
Responsibilities
- Design, implement, validate, and ship production C/C++ features and programming-model capabilities in the CUDA driver and runtime.
- Optimize kernel launch, synchronization, memory management and movement, CPU–GPU coordination, and system interconnects for latency, throughput, bandwidth, efficiency, and scalability.
- Investigate complex performance problems end to end through workload analysis, measurement, modeling, root-cause isolation, production implementation, and quantitative validation.
- Establish performance expectations, characterize new silicon, support platform bring-up, close software and hardware gaps, and drive release readiness.
- Translate workload and platform evidence into CUDA API, programming-model, systems-software, and future hardware improvements.
- Lead cross-layer feature development and investigations, define technical direction, mentor engineers, and influence hardware/software decisions.
Requirements
- Bachelor’s, master’s, or doctoral degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent practical experience.
- At least 2 years of relevant systems-software development experience.
- Strong production C/C++ systems-programming experience delivering substantial features, optimizations, or fixes in complex codebases.
- Strong operating-system and concurrency knowledge, including threads, synchronization, processes, virtual memory, and user/kernel interactions.
- Strong computer-architecture knowledge, including processors, memory hierarchy, caching and coherence, data movement, and system interconnects.
- Demonstrated experience improving software performance through measurement, bottleneck identification, effective implementation, and quantitative validation.
- Sound technical judgment, ownership of ambiguous problems, and clear communication across teams and disciplines.
- Direct CUDA or GPU experience is valuable but not required with deep systems-software, operating-system, computer-architecture, and performance-engineering foundations.
- Preferred experience includes GPU or accelerator drivers, runtimes, kernel software, firmware, compilers, performance-critical systems, pre-silicon analysis, platform bring-up, performance modeling, hardware/software co-design, or demanding AI/DL, HPC, graphics, automotive, and robotics workloads.
- Evidence of technical invention, such as software-performance patents, novel production designs, or measurement-backed recommendations influencing hardware or architecture.
- Python or another scripting language used for experimentation, data analysis, or visualization.
Benefits
- Eligible for equity and benefits.
- Base salary depends on location, experience, and comparable employee pay, with a stated range of 124,000 USD to 195,500 USD.
- Applications will be accepted at least until September 5, 2026.
- This posting is for an existing vacancy.
About Nvidia
Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.
