1 day ago
Remote, India or Bengaluru, IndiaSenior
Responsibilities
- Design and ship distributed systems services in Java, Go, and Rust for NVIDIA Cloud Functions.
- Improve the performance, reliability, and scaling behavior of systems routing AI workloads across distributed GPU fleets.
- Automate and optimize cloud-native build, test, integration, and release processes.
- Integrate the platform with NVIDIA technologies including KAI Scheduler, NVIDIA NIM, Grove, and Dynamo.
- Steward the open-source project by triaging community issues, reviewing pull requests, and writing contributor documentation.
Requirements
- Bachelor’s or Master’s degree in Computer Science or equivalent experience.
- At least 3 years of hands-on software engineering experience.
- Expert-level knowledge of Go, C, or Rust and strong understanding of data structures, algorithms, and distributed software architecture.
- Strong understanding of Kubernetes, container technologies, and hands-on automation with continuous integration frameworks such as GitLab and ArgoCD.
- Expertise in Bash or Python and experience with Unix/Linux kernel internals.
- Understanding of performance, security, and reliability in complex distributed systems.
- Preferred experience with pub-sub models, message queues, high-throughput network paths, HTTP/2, gRPC, Kubernetes Custom Resources, Kubernetes Operators, and cloud service providers.
Benefits
- Competitive salary and a generous benefits package.
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
