2 days ago
Base Salary
$224k - $357k/yr
Responsibilities
- Define and evolve multi-tenant control-plane and data-plane architecture using Open vSwitch, OVN, OpenFlow, and overlay networks.
- Provide technical leadership through architecture, design and code reviews, mentoring, implementation guidance, and production-readiness decisions.
- Develop production software for Kubernetes networking, network plugins, distributed control planes, Linux host networking, virtual-machine networking, and network automation.
- Translate infrastructure requirements into secure orchestration services, APIs, component boundaries, state models, compatibility strategies, and delivery plans.
- Drive Open vSwitch and OVN integration across configuration, lifecycle management, upgrades, interoperability, performance, failure recovery, and large-scale production operation.
- Build automated unit, integration, system, performance, scale, and upgrade tests for continuous integration, deployment, and release qualification.
- Participate in on-call rotations, lead networking escalations, investigate incidents, and prevent recurrence using packet captures, telemetry, profiling, experiments, and source debugging.
- Define and improve reliability, performance, security, and resource-efficiency objectives with monitoring, telemetry, tracing, and service-level visibility.
Requirements
- BS or MS in Computer Science, Computer Engineering, or a related field, or equivalent experience.
- 12+ years of experience designing, implementing, testing, and maintaining production software in both C and Go.
- Experience using Bash and Python for testing, diagnostics, builds, or operational automation.
- Hands-on experience developing, integrating, and troubleshooting Open vSwitch and OVN, including OpenFlow and control-plane-to-data-plane behavior.
- Production Kubernetes networking experience covering container network interfaces, network policy, node and pod traffic paths, network plugins, upgrades, and failure modes.
- Strong Linux networking fundamentals, including IP, TCP, UDP, routing, switching, overlay networks, tunneling, network namespaces, and network policy.
- Experience architecting distributed systems with secure service APIs, state management, consistency, scalability, compatibility, and failure handling.
- Experience with automated testing, CI/CD, deployments, upgrades, observability, and performance analysis throughout the production lifecycle.
- Experience supporting production services through on-call rotations and coordinating evidence-based incident response.
- Demonstrated technical leadership through written designs, consequential decisions, cross-team collaboration, critical reviews, mentoring, and delivery of significant systems.
- Preferred experience contributing to Open vSwitch, OVN, OVN-Kubernetes, Kubernetes networking, or related open-source projects.
- Preferred experience with large-scale cloud and accelerated virtualization systems, including SR-IOV, RDMA, SmartNICs, data processing units, network function virtualization, KVM/QEMU, or container runtimes.
- Preferred experience building secure, high-performance gRPC or REST services with transport security and strong authentication.
- Preferred experience using agentic AI and AI-assisted software-development tooling, including coding agents, reusable skills, or Model Context Protocol integrations.
Benefits
- Competitive salary, equity, and a generous benefits package.
- Applications are accepted at least until September 17, 2026.
- This posting is for an existing vacancy.
Tech Stack
Categories
About Nvidia
Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.
