Senior Product Architect

NVIDIA NVIDIA · Semiconductors · Santa Clara, CA +1 · Remote

This role focuses on architecting and defining product architectures for AI infrastructure, with a strong emphasis on systems, networking, compute, and storage. The role involves leading prototyping, testing, and collaborating with various teams and customers to develop AI solutions, particularly those involving agentic and RAG-based workflows, and inference at scale. The primary output is the architecture and design for AI infrastructure products.

What you'd actually do

  1. Stay at the forefront of technology trends in AI Infrastructure & Agentic AI, proposing new product concepts
  2. Define detailed product architectures, including performance, scalability, interoperability, and datacenter deployment requirements
  3. Lead prototyping and testing processes to validate product design and functionality, with a focus on scale, diagnosing bottlenecks and driving innovations in performance
  4. Iterate and refine designs based on feedback, testing results, and evolving requirements
  5. Collaborate with internal hardware, software, security and systems teams to ensure flawless integration of networking, compute and storage ecosystems

Skills

Required

  • 12+ years of experience architecting datacenter-scale HPC or AI infrastructure
  • Bachelors degree in computer science or related field (or equivalent experience)
  • Strong background in technologies that enable large scale infrastructure management of complex systems, networks, and storage
  • Extensive experience with DevOps solutions, including Python, Ansible, Container Runtimes, Kubernetes, and data center deployments
  • Familiarity with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning, and model evaluation
  • Exceptional communication skills

Nice to have

  • Prior first-hand experience on having built large scale AI infrastructure
  • Advanced certifications or publications in AI, deep learning, or related fields
  • Experience leading high-impact projects or initiatives in cutting-edge tech domains

What the JD emphasized

  • architecting datacenter-scale HPC or AI infrastructure
  • extensive experience with DevOps solutions
  • familiarity with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning, and model evaluation
  • Prior first-hand experience on having built large scale AI infrastructure

Other signals

  • architecting datacenter-scale HPC or AI infrastructure
  • defining detailed product architectures
  • lead prototyping and testing processes
  • collaborate with internal hardware, software, security and systems teams
  • work closely with partners and customers to conceptualize ideal solutions
  • familiarity with AI workloads, including agentic & RAG-based workflows, inference at scale, large scale training & fine-tuning, and model evaluation