Nvidia

Senior Solutions Architect, Generative AI

Nvidia
Apply
15 hours ago
Bengaluru, IndiaSenior

Responsibilities

  • Architect end-to-end generative AI solutions centered on Speech AI, LLMs, agentic workflows, and RAG.
  • Work directly with customers to understand business challenges, define requirements, and design tailored solutions.
  • Support pre-sales activities through technical presentations, demonstrations, workshops, and design sessions.
  • Train, fine-tune, deploy, and optimize LLM and Speech AI models for production inference.
  • Design and integrate RAG workflows into customer applications and systems.
  • Collaborate with NVIDIA engineering teams and provide technical leadership on generative AI best practices.

Requirements

  • B.Tech, master's, Ph.D. in Computer Science or Artificial Intelligence, or equivalent experience.
  • At least 8 years of hands-on technical experience focused on generative AI, LLM training, and Speech AI.
  • Proven experience deploying and optimizing LLM and Speech AI models for production inference.
  • Strong understanding of language model architectures including GPT-3, BERT, or similar models.
  • Expertise training and fine-tuning LLMs with TensorFlow, PyTorch, or Hugging Face Transformers.
  • Experience with model deployment, inference optimization, GPUs, GPU cluster architecture, and parallel processing.
  • Strong communication skills and experience presenting technical solutions or leading workshops and training sessions.
  • Preferred experience with Docker, Kubernetes, distributed computing, NVIDIA GPU technologies, and GPU cluster management.

Benefits

  • Competitive salary and generous benefits package.
  • NVIDIA is an equal opportunity employer committed to a diverse work environment.

Tech Stack

DockerHugging Face TransformersKubernetesPyTorchTensorFlow

Categories

Solutions Engineering
Nvidia

About Nvidia

10,000+ employees

Nvidia designs and sells GPUs and accelerated computing platforms for data centers, AI/ML, graphics, gaming, and automotive, monetizing through hardware, software platforms (CUDA, AI frameworks), and systems like DGX and networking. Customers include cloud providers, enterprises, researchers, and OEMs. Founded in 1993 and headquartered in Santa Clara, it is a public company traded on NASDAQ under NVDA.

Contact me