Nvidia

Senior AI infrastructure engineer - EDA Infrastructure

Nvidia
Apply
3 days ago
Remote, Worldwide +2 moreSenior
H1B Sponsor

Base Salary

$184k - $357k/yr

Responsibilities

  • Build and operate scalable telemetry pipelines for metrics, logs, traces, and events across on-premise, cloud-service-provider, and NVIDIA Cloud Platform clusters.
  • Establish standardized instrumentation, collection, storage, and access patterns for telemetry.
  • Deliver dashboards, alerting, and analysis capabilities for service visibility, detection, and troubleshooting.
  • Standardize and automate incident, maintenance, service on-call, and support on-call workflows.
  • Integrate operational data and lifecycle signals to improve ownership, escalation, communication, and post-incident learning.
  • Build reporting and AI-assisted tooling to reduce manual work and improve operational responsiveness.
  • Build and maintain physical hardware and software catalogs covering infrastructure inventory, service ownership, dependencies, and documentation.
  • Create data models and integration pipelines connecting clusters, hardware, services, teams, and operational workflows.
  • Provide self-service infrastructure discovery capabilities for engineers.
  • Lead cross-functional initiatives involving engineering, product, finance, security, and external partners.

Requirements

  • Bachelor’s degree in Computer Science, Computer Engineering, or a related technical field, or equivalent experience.
  • At least 8 years of experience in infrastructure security, platform engineering, or security tooling.
  • Proficiency in one or more programming languages such as Python, Go, TypeScript, or Java.
  • Strong understanding of software and infrastructure principles and experience applying them in production environments.
  • Ability to lead cross-functional initiatives across internal teams and external partners.
  • Preferred: experience building and operating incident-management processes with internally developed or external SaaS tools.
  • Preferred: experience building, deploying, and maintaining ML models in production and familiarity with AI agent frameworks or orchestration tools.
  • Preferred: experience building and operating modern observability platforms for metrics, logs, traces, and profiling.
  • Preferred: experience with service catalogs and configuration management databases (CMDBs).

Benefits

  • Eligible for equity and benefits.
  • Base salary range is $184,000–$287,500 USD for Level 4 or $224,000–$356,500 USD for Level 5.
  • Applications will be accepted at least until September 4, 2026.

Categories

Nvidia

About Nvidia

10,000+ employees

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Contact me