Distinguished AI Researcher, Advanced Development Engineer

NVIDIA NVIDIA · Semiconductors · Raanana, Israel

This role focuses on researching and defining co-design solutions for AI infrastructure, spanning GPU computing, networking, storage, and large-scale distributed systems. The goal is to optimize and operate massive-scale AI environments for both training and inference workloads, influencing product direction and system design.

What you'd actually do

  1. Research, define, and drive co-design solutions across AI software stacks involving GPU computing, RDMA networking, NVLink interconnects, storage, and large-scale distributed infrastructure
  2. Lead deep technical investigations and proof-of-concept development in areas related to distributed AI, deep learning systems, high-performance computing, virtualization, memory management, and large-scale systems design
  3. Work closely with multiple teams across NVIDIA to advance next-generation AI infrastructure technologies and influence cross-stack co-design decisions
  4. Analyze end-to-end system bottlenecks and opportunities across compute, networking, memory, and storage layers to improve performance, scalability, and efficiency
  5. Guide the design of environments that support massive-scale AI training and inference workloads in production

Skills

Required

  • M.S. or Ph.D. in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience
  • 20+ years of industry experience in systems design, distributed systems, system software, or related fields
  • Deep background in distributed systems, parallel computation, networking, storage, virtualization, and memory management
  • Experience building, scaling, and supporting environments for massive AI workloads
  • Strong understanding of large-scale systems operational aspects, including performance, reliability, scalability, and supportability
  • Excellent communication, collaboration, and technical leadership skills

Nice to have

  • Chief Scientist, CTO, or equivalent senior technical leadership experience in smaller companies is highly desirable
  • Proven research track record in AI systems, large-scale infrastructure, or advanced systems co-design
  • Deep expertise in GPU systems, RDMA, NVLink, and storage technologies for AI and HPC workloads
  • Strong record of cross-functional influence spanning hardware, system software, networking, and infrastructure

What the JD emphasized

  • 20+ years of industry experience
  • Chief Scientist, CTO, or equivalent senior technical leadership experience in smaller companies is highly desirable
  • Proven research track record in AI systems, large-scale infrastructure, or advanced systems co-design

Other signals

  • AI infrastructure
  • distributed systems
  • GPU computing
  • large-scale AI training and inference