1 day ago
Munich, Germany or Warsaw, PolandSenior
Responsibilities
- Advise on and help maintain large-scale computational and AI infrastructure, including monitoring, logging, and workload orchestration.
- Provide consultative guidance and hands-on troubleshooting across bare metal, operating systems, software stacks, container platforms, networking, and storage.
- Assess customer environments and recommend production-ready Kubernetes container platforms integrated with enterprise networking and storage.
- Develop, refine, and document standard methodologies, operational guidelines, runbooks, onboarding materials, and best-practice guides.
- Support development activities and lead proofs of concept and proofs of value for new features, architectures, and upgrade approaches.
- Act as the technical leader for assigned customer accounts and influence long-term DevOps, platform architecture, infrastructure, and operations decisions.
Requirements
- Bachelor's, master's, or PhD in computer science, electrical or computer engineering, physics, mathematics, or a related field.
- At least 5 years of professional experience managing scalable cloud environments and working in automation engineering roles.
- Hands-on experience managing HPC and AI clusters, including deployment, optimization, and troubleshooting.
- Hands-on experience deploying, configuring, and optimizing NVIDIA GPU-accelerated infrastructure, including driver management, CUDA integration, and GPU workload profiling.
- Extensive Kubernetes experience covering container orchestration, resource scheduling, scaling, and integration with GPU-accelerated and HPC environments.
- Strong knowledge of data center architectures, networking fundamentals, HPC and AI technologies, Linux, operating-system security, and storage systems.
- Proficiency in Python, Bash, configuration management, infrastructure-as-code tools, and observability stacks.
- Experience with Ansible, Terraform, Grafana, Loki, and Prometheus is required or expected within the stated technical profile.
- Strong solution-architecture, customer-consulting, architectural-review, presentation, and executive-stakeholder engagement skills.
- Preferred experience includes CI/CD pipelines, Kubernetes operators, SLURM, MPI, enroot, cluster change management, NVIDIA Base Command Manager, RDMA, InfiniBand, and RoCE.
Benefits
- Base salary in Poland is listed as 221,250 PLN–383,500 PLN for Level 3 and 292,500 PLN–507,000 PLN for Level 4.
Tech Stack
Categories
Solutions Engineering
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
