7 days ago
Responsibilities
- Collaborate with developers, researchers, and framework maintainers to identify and resolve CPU performance challenges.
- Profile, analyze, benchmark, and optimize applications from algorithm-level implementation through CPU microarchitecture.
- Contribute to open-source frameworks, software stacks, reference implementations, and performance libraries.
- Work with architecture, research, libraries, tools, and system software teams to improve platform performance.
- Provide technical insights for future CPU designs, compiler toolchains, and development workflows.
Requirements
- Bachelor’s, master’s, or PhD in Computer Science, Computer Engineering, or a related field.
- At least 5 years of relevant experience in performance engineering or CPU optimization.
- Strong proficiency in C/C++ and/or Python, with deep knowledge of algorithms and software architecture.
- Understanding of CPU microarchitecture, performance analysis tools, and optimization methodologies.
- Demonstrated experience with CPU benchmarking and bottleneck-driven performance tuning.
- Strong communication and organizational skills for cross-functional collaboration and managing multiple priorities.
- Preferred experience includes CPU optimization of AI or data preprocessing pipelines, HPC applications, parallel computing, distributed runtimes, SIMD instruction sets, low-level intrinsics, vectorization, and open-source performance or HPC tools.
Benefits
- Competitive salaries.
- Comprehensive benefits package.
- Company culture and a diverse work environment.
About Nvidia
Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.
