2 days ago
Base Salary
$184k - $357k/yr
Responsibilities
- Develop and optimize Rust-based data-processing frameworks for large datasets in GPU-accelerated environments supporting LLM training.
- Lead development and optimization of GPU-accelerated Agentic Retrieval pipeline components and improve model performance and total cost of ownership.
- Collaborate with LLM and ML researchers to build full-stack data-preparation pipelines for multimodal models.
- Implement benchmarking, profiling, and Python-based algorithm optimization for LLM applications across system architectures.
- Build and evaluate proofs of concept, define production roadmaps, and develop tools and library features within the LLM ecosystem.
- Build employee productivity products and integrated generative-AI and copilot experiences.
- Help maintain the Continuous Delivery pipeline and operational standards for safe, rapid production releases.
- Conduct peer reviews and contribute to frameworks, standards, scalability, performance, correctness, and new technology adoption.
Requirements
- Bachelor’s or master’s degree in Computer Science, Computer Engineering, or a related field, or equivalent experience.
- At least 6 years of experience in a similar or related role.
- Experience delivering software in a cloud context and familiarity with managing cloud infrastructure.
- Knowledge of MLOps technologies including Docker-Compose, containers, Kubernetes, and data center deployments.
- In-depth hands-on understanding of NLP, LLM, VLM, generative AI, and Agentic Retrieval workflows.
- Ability to work effectively with multifunctional teams, principals, and architects across organizational boundaries and geographies.
- Strong communication skills for explaining sophisticated technical topics clearly and impactful conclusions.
Benefits
- Base salary ranges from 184,000 USD to 287,500 USD for Level 4 and from 224,000 USD to 356,500 USD for Level 5, with eligibility for equity and benefits.
- Applications will be accepted at least until October 6, 2026.
- NVIDIA uses AI tools in its recruiting processes and is an equal opportunity employer.
Tech Stack
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
