1 day ago
Santa Clara, CA, USAMid Level / Senior
H1B sponsor
Base Salary
$124k - $242k/yr
Responsibilities
- Evaluate cloud-native, full-stack microservices applications that support AI use cases using NVIDIA frameworks, SDKs, and microservices.
- Design and implement agentic workflows using Retrieval-Augmented Generation and advanced AI models.
- Evaluate user experiences and technical performance of AI solutions, document findings, and recommend product improvements to senior executives and engineering management.
- Collaborate with product, marketing, hardware, software engineering, and QA teams to improve NVIDIA products.
- Create developer-focused tutorials and code samples demonstrating NVIDIA tools and libraries.
- Write technical whitepapers and product briefs and conduct technical demonstrations at industry conferences.
Requirements
- Bachelor’s or master’s degree in Software Engineering, Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent experience.
- At least 3 years of experience.
- Proficiency in Python and JavaScript for programming and debugging, with a foundation in data structures, algorithms, and software design principles.
- Basic familiarity with C++ and its use in high-performance computing environments.
- Experience building cloud-native systems optimized for Kubernetes deployment with vLLM and NVIDIA Triton Inference Server.
- Understanding of API design principles for scalable, production-ready inference systems.
- Advanced LLM, modern AI software architecture, cloud API, LLM inference framework deployment, open-source, or public-facing technical content experience is preferred.
Benefits
- Base salary ranges from $124,000 to $195,500 for Level 2 and from $152,000 to $241,500 for Level 3, depending on location, experience, and comparable employee pay.
- Eligible for equity and benefits.
- Located in Silicon Valley.
- Applications are accepted at least until October 5, 2026.
Tech Stack
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
