2 days ago
Tel Aviv-Yafo, IsraelSenior
Responsibilities
- Develop scalable services, APIs, and libraries for managing and loading AI models and data.
- Integrate cloud infrastructure, storage, and AI frameworks in Kubernetes environments.
- Guide an open-source project’s technical direction, review contributions, and coordinate community releases.
- Deliver features from design through production and improve performance and reliability through testing, monitoring, and benchmarking.
- Collaborate with engineers and product managers, contribute to technical decisions, and share knowledge.
Requirements
- A relevant degree or equivalent practical experience.
- At least 5 years of experience developing backend services, infrastructure, or developer tools.
- Proficiency in Python, Go, or another relevant programming language, with willingness to learn new technologies.
- Hands-on experience with Kubernetes, containers, and Linux.
- Understanding of distributed systems and infrastructure fundamentals, including compute, networking, and storage.
- Experience delivering and supporting reliable production software.
- Regular use of AI tools for design, coding, testing, and debugging, with careful evaluation of generated outputs for correctness, security, and maintainability.
- Ability to investigate problems, explain technical tradeoffs, and collaborate across teams.
- Preferred experience with AI model loading, inference, PyTorch, vLLM, SGLang, cloud storage, caching, performance optimization, open-source projects, libraries, SDKs, C++, or GPU computing.
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
