NVIDIA currently has 496 active AI-related job listings. The majority of these roles, 52%, are focused on serving infrastructure, with agents representing another significant segment at 23%. Engineering is the dominant function, with 441 positions. The United States leads hiring geographies with 287 roles, followed by China with 64. Frequent tech tags include model_serving, inference_infra, and agent_orchestration, suggesting a focus on deployment and management of AI models. Over the last 30 days, NVIDIA posted 214 new AI roles, a 27% decrease compared to the previous 30-day period.
Currently tracking 440 active AI roles, down 53% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $100k–$575k (avg $262k).
NVIDIA currently has 487 active AI-related roles in our index. The most common open titles are: Deep Learning Performance Architect (4), Senior Deep Learning Performance Architect (4), AI Research Scientist (3), Developer Technology Engineer - AI (3), Manager, Deep Learning Algorithms (3). Most positions are in Engineering and Research.
NVIDIA's active AI hiring is concentrated in: serving infrastructure (54%), agents (21%), application (8%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
NVIDIA is hiring AI talent in: United States (286 roles), China (59 roles), Israel (50 roles), Germany (21 roles).
Job postings at NVIDIA most frequently reference: model serving, inference infra, agent orchestration, llm observability, multimodal.
In the past 30 days, NVIDIA has posted 110 new AI-related roles. That is a -50% change versus the prior 30 days (218 → 110).
| Title | Stage | AI score |
|---|---|---|
| Senior Software Engineer, CUTLASS Performance Senior Software Engineer role focused on optimizing the performance of CUTLASS, a high-performance linear algebra and Tensor Core primitive ecosystem for NVIDIA GPUs. The role involves benchmarking deep learning models, identifying performance gaps, developing tooling for optimization, and acting as a performance representative across NVIDIA teams. | Serve | 7 |
| Deep Learning Performance Architect NVIDIA is seeking a Deep Learning Performance Architect to optimize deep learning hardware and software architecture, analyze performance of deep learning algorithms on different architectures, identify bottlenecks, and explore new features and hardware capabilities. Requires a strong background in computer architecture and experience with deep learning platforms and frameworks. | Serve |
| 7 |
| Principal Architect, System Software - Orbital Data Center NVIDIA is seeking a Principal Architect to lead the system software architecture for their Orbital Data Center (ODC) modules, specifically Space-1. This role involves designing and implementing a resilient, production-ready inference platform for the harsh environment of low-Earth orbit, covering the full stack from firmware to AI workloads. The architect will collaborate with hardware teams, drive customer use cases, and ensure the platform operates reliably for 5-year missions, enabling AI adoption in space. | Serve | 7 |
| Software Engineer, TensorRT Specialized Platforms - New College Grad 2025 Software Engineer role focused on developing and optimizing high-performance deep learning inference software (TensorRT) for specialized platforms. Requires strong C++ skills, familiarity with deep learning frameworks, and interest in performance optimization and systems programming. | Serve | 7 |
| Deep Learning Compiler Engineer - CUDA NVIDIA is seeking a Deep Learning Compiler Engineer to design and implement DSLs and compiler cores for emerging GPU architectures, focusing on optimizing performance for AI/LLM workloads and integrating with AI/ML frameworks. | Serve | 7 |
| Developer Technology Engineer, AI NVIDIA Developer Technology Engineer focused on optimizing AI and deep learning applications on GPU architectures, working with customers to provide AI solutions, and collaborating with internal teams to influence future hardware and software design. | Serve | 7 |
| Senior HPC AI Cluster Engineer NVIDIA is seeking an experienced HPC-AI Engineer to join their Networking Clusters Solutions Infrastructure team. The role involves designing, implementing, and maintaining large-scale HPC/AI clusters, managing job schedulers, developing CI/CD pipelines, and automating infrastructure deployment and monitoring. The engineer will work with cutting-edge hardware and software, support R&D, and engage in POCs for future improvements. | Serve | 7 |
| Senior Power Analysis and Optimization Engineer This role focuses on applying AI, ML, and LLMs to optimize power efficiency in NVIDIA's GPUs and SoCs. The engineer will develop and productionize ML/RL-based models for power analysis and optimization, design and train custom LLMs for interpreting power data and recommending improvements, and apply AI to tune power-efficient configurations. The role involves analyzing power data, partnering with cross-functional teams, and automating flows. | ServeData | 7 |
| Senior Software Engineer — cuEquivariance Senior Software Engineer to join the cuEquivariance team, which builds and ships production GPU kernels and software interfaces for equivariant deep learning. The role involves CUDA kernel engineering, Python library development (PyTorch/JAX), and collaboration with research teams and external framework developers to accelerate geometric neural networks on NVIDIA GPUs. | Serve | 7 |
| Senior System Software Engineer - AI Performance and Efficiency Tools NVIDIA is seeking a Senior System Software Engineer to develop tools for AI researchers and SW/HW teams running AI workloads on GPU clusters. The role involves building internal profiling, analysis, debugging, benchmarking, and simulation tools to improve the performance and efficiency of AI workloads and systems. This includes partnering with HW architects and understanding deep learning frameworks, distributed training/inference, and GPU cluster technologies. | ServeData | 7 |
| Senior Systems Software Engineer - GPU Performance at Scale Senior Systems Software Engineer focused on GPU performance at scale for AI workloads. This role involves leading performance practices, aligning AI workloads with hardware, developing insights into AI workload performance, debugging complex issues, and collaborating with various software and firmware teams to optimize AI workload performance on NVIDIA GPUs. | Serve | 7 |
| Senior Software Engineer, AI Resiliency Senior Software Engineer to lead the development of AI software resiliency for large-scale AI supercomputers (100,000+ GPUs), focusing on features like fast checkpoint-recovery, error detection/isolation, and straggler/hang detection to minimize cluster downtime. The role involves hands-on C++ and Python coding, debugging, fault tolerance, and collaboration with AI researchers and hardware/software teams, integrating resiliency into AI frameworks like PyTorch and JAX/XLA. Experience with distributed systems, fault tolerance, AI frameworks, and debugging tools is required, with a preference for experience in training models, CUDA/NCCL/MPI, checkpointing strategies, and large-scale AI clusters/HPC. | Serve | 7 |
| Senior Networking Performances Architect NVIDIA is seeking a Senior Networking Performances Architect to shape the future of high-performance and ML/AI computing. This role will analyze network feature performance for AI workloads on large-scale HPC clusters, develop network behavior models, and generate insights for next-generation products. The ideal candidate will have a strong background in system engineering/architecture, performance research, Python, and a good understanding of AI models and large-scale networks. | Serve | 7 |
| Systems Software Engineer - New College Grad 2026 Systems Software Engineer role focused on applying AI and computational methods to accelerate semiconductor manufacturing and design using GPUs. The role involves developing and optimizing complex software solutions, with a strong emphasis on performance and parallel programming. | Serve | 7 |
| Senior Deep Learning Systems Engineer, Datacenters Senior Deep Learning Systems Engineer focused on analyzing and optimizing the performance and power consumption of deep learning applications on datacenter hardware, influencing the design of future AI systems and software stacks. This role involves developing software infrastructure, analysis tools, and profiling methodologies for DL workloads, with a strong emphasis on system architecture and performance analysis. | Serve | 7 |
| Senior HPC and AI Operation Engineer NVIDIA is seeking a Senior HPC and AI Operation Engineer to manage and maintain large-scale HPC/AI clusters, including job scheduling, CI/CD pipelines, and troubleshooting from bare metal to application level. The role involves supporting R&D activities and engaging in POCs, requiring strong Linux administration, scripting, and knowledge of HPC/AI technologies, storage, and networking. | Serve | 7 |
| Senior System Software Engineer - AI Performance and Efficiency Tools Develops internal profiling, analysis, debugging, benchmarking, and simulation tools for AI workloads running on GPU clusters, supporting AI researchers and SW/HW teams to improve performance and efficiency. | ServeData | 7 |
| Senior Developer Technology Engineer - Windows AI Platform Senior Developer Technology Engineer focused on optimizing and deploying AI/GenAI applications on NVIDIA RTX platforms, particularly LLMs on Windows. This role involves working with internal teams and external developers, analyzing performance, conducting training, and improving user experience with OSS software like Llama.cpp and Ollama. Collaboration with driver and architecture teams is key to influencing future GPU features. | ServeAgent | 7 |
| Senior Deep Learning Tools Engineer – CUDA Tile Senior Deep Learning Tools Engineer at NVIDIA focused on performance validation, analysis, and tracking for AI workloads accelerated by CUDA Tile compiler technologies and GPU systems. The role involves designing and developing performance testing frameworks, building automated CI/CD pipelines, implementing benchmarking systems, analyzing performance trends, and collaborating with compiler and architecture teams to resolve performance issues. Requires strong programming skills in Python, experience with CI/CD, deep learning frameworks, and hardware-aware performance analysis. | Serve | 7 |
| Senior Systems Software Engineer - GPU Performance at Scale Senior Systems Software Engineer focused on GPU performance at scale for AI workloads, involving collaboration with various hardware and software teams to optimize large-scale computing platforms and deliver insights into AI workload performance. | Serve | 7 |
| Senior Compiler Engineer - AI NVIDIA is seeking a Senior Compiler Engineer with expertise in machine learning and compiler technologies to focus on applied AI and ML within compilers and development tools. The role involves working with Python, C/C++, Julia, and Lisp/Scheme, with a strong foundation in compilers, code generation, and GPU architecture. Experience with LLVM is a plus. | Serve | 7 |
| Distinguished Software Architect - Deep Learning and HPC Communications Distinguished Software Architect role focused on designing and researching next-generation communication libraries and platforms for Deep Learning and High Performance Computing at NVIDIA. The role involves co-designing HW/SW solutions with GPU, Networking, and SW architects, driving adoption of new communication technologies, and keeping up with DL research. Requires deep expertise in HPC, parallel programming, communication runtimes, system/GPU architecture, and networking, with strong programming skills in C/C++. | Serve | 7 |
| Manager, Next-Gen AI Cluster Validation Manager to lead a team developing and validating next-generation NVIDIA AI supercomputing systems, integrating new compute, networking, storage, and software. Focus on building a platform for software development, automation, and performance engineering, and supporting large-scale deployments for AI and HPC. | Serve | 7 |
| GPU Power Architect - New College Grad 2026 NVIDIA is seeking a New College Grad Datacenter GPU Power Architect to contribute to the research and development of energy-efficient GPU and SOC architectures. The role involves developing power estimation models and tools, exploring energy efficiency at GPU and Datacenter levels, and deploying machine learning techniques to model GPU, CPU, Switch, and platform performance and power. The candidate will understand GenAI/HPC workload characteristics to drive HW/SW features for Perf@Watt improvements. | Serve | 7 |
| Senior Software Engineer, Data Center Workloads – Infrastructure Senior Software Engineer focused on developing and executing software-driven characterization workflows for AI workloads on NVIDIA rack-scale systems. The role involves analyzing, characterizing, and optimizing power, performance, and drive behavior across the full stack, including GPUs, CPUs, networking, and system software. Key responsibilities include building automated frameworks for data collection and analysis, investigating system behavior, and supporting new platform bring-up. | Serve | 7 |
| Senior Deep Learning Compiler Engineer NVIDIA is seeking a Senior Deep Learning Compiler Engineer to develop compiler optimization algorithms for deep learning networks. This role involves collaborating with deep learning software framework and hardware architecture teams to accelerate next-generation deep learning software, focusing on public APIs, performance, and compiler infrastructure for neural networks. | Serve | 7 |
| Senior Software Engineer - Deep Learning Compiler CI Infrastructure Senior Software Engineer to own and evolve CI/CD infrastructure for NVIDIA's deep learning compiler stacks. Responsibilities include designing and operating scalable CI systems for ML workloads, delivering performance signals, and applying AI/agent-based workflows to improve developer efficiency and triage. | Serve | 7 |
| Senior Software Developer, AI Networking Senior Software Developer focused on AI Networking at NVIDIA, developing communication frameworks, production tools, and benchmarks for large-scale AI training and inference systems. The role involves enabling new AI models, analyzing workloads, designing automation, and collaborating with hardware teams. | ServeData | 7 |
| SoC Product Architect, Telecom AI RAN NVIDIA is seeking a Lead SoC Product Architect for their Telecom AI RAN platform, focusing on defining the architecture and roadmap for radio and distributed unit products. The role involves analyzing workloads, driving competitive analysis, synthesizing customer requirements, and collaborating with engineering teams to ensure efficient implementation of AI-native RAN applications. The ideal candidate will have extensive experience in wireless RAN/baseband architecture or SoC product definition, with a strong understanding of 3GPP RAN standards and L1/PHY algorithms. | Serve | 7 |
| Senior System Software Engineer - Neural Graphics Performance Senior System Software Engineer focused on optimizing neural graphics performance, specifically Gaussian Splatting and neural reconstruction algorithms, for applications in robotics, healthcare, and AV development. The role involves implementing and optimizing reconstruction/rendering algorithms using CUDA and Slang, optimizing data processing pipelines, and influencing software architecture for performance. | ServeData | 7 |
| Senior System Software Engineer - Dynamo-Triton Inference Server Senior System Software Engineer to work on Dynamo-Triton Inference Server, a GPU-accelerated AI inference serving platform. The role involves developing high-performance inference software, contributing to feature development, driving customer adoption, and optimizing throughput and latency for both LLM and non-LLM workloads. | Serve | 7 |
| Senior Software Performance Engineer - AV Platform Senior Software Performance Engineer for Autonomous Vehicles platform, focusing on optimizing latency and throughput of L2/L3/L4 autonomous driving solutions on NVIDIA's heterogeneous hardware architectures. Requires strong C++ skills, parallel programming, performance analysis, and experience with GPGPU/CUDA. | ServeAgent | 7 |
| Senior AI and ML HPC Cluster Engineer This role focuses on designing, implementing, and managing large-scale GPU compute clusters for AI/ML and HPC workloads. It involves infrastructure engineering, automation, and supporting researchers with performance analysis and optimization. The role requires expertise in cluster management, Linux administration, container technologies, scripting, and MPI workflows. | Serve | 7 |
| Developer Technology Engineer – AI NVIDIA Developer Technology Engineer focused on optimizing deep learning and machine learning workloads on NVIDIA's accelerated computing platform (GPU, CPU, DPU) for key customers. Requires strong C/C++ and CUDA experience, with an MS/PhD in CS or related field. | Serve | 7 |
| Manager, Software Architecture Manager for a systems and networking engineering team focused on building distributed AI communication systems (libraries, frameworks, system integrations) for GPUs, nodes, and storage. The role involves setting technical direction, leading execution, and fostering technical excellence within the team, with a focus on AI infrastructure problems. | Serve | 7 |
| Senior Performance Engineer Senior Performance Engineer at NVIDIA focusing on optimizing AI and HPC workloads on GPU/CPU clusters. Responsibilities include profiling, benchmarking, identifying bottlenecks, and developing performance analysis tools, with a strong emphasis on high-performance networking and telemetry. | Serve | 7 |
| Senior Software Engineer - Verification AI Infrastructure Senior Software Engineer focused on building and optimizing scalable software automation systems with AI/ML integration for NVIDIA's Data Center environments. The role involves developing automation and validation tools, improving system performance, and troubleshooting complex issues in distributed systems. | Serve | 7 |
| Senior Software Architect, AI Systems and Networking This role focuses on building and optimizing systems-level software for high-performance communication and memory management libraries essential for distributed AI workloads. It involves hardware-software co-optimization, profiling data movement, and integrating networking capabilities into AI serving stacks, bridging applied research and production engineering. | Serve | 7 |
| Deep Learning Kernel Software Performance Architect - New College Grad 2026 NVIDIA is seeking a Deep Learning Kernel Software Performance Architect to develop and analyze processor and system architectures that accelerate machine learning and data analytics applications. The role involves debugging deep learning software, developing analysis tools, and collaborating with various NVIDIA teams to optimize performance. | Serve | 7 |
| Senior Computer Vision and Deep Learning Hardware Architect NVIDIA is seeking an Autonomous Vehicle Performance Architecture Engineer to design, model, and verify state-of-the-art programmable vision accelerators (PVA) for automotive and robotics. The role involves optimizing software for autonomous driving solutions, analyzing and prototyping applications, building performance models for future architectures, and collaborating with teams to enhance PVA architecture. Requires a Masters/PhD, 3+ years of relevant experience, strong C/C++ and computer architecture skills, and performance modeling/optimization expertise. Experience in DSP programming, autonomous vehicle software, deep learning, computer vision, and self-driving cars is a plus. | ServePost-train | 7 |
| Senior Software Engineer, NCCL Senior Software Engineer role focused on designing, implementing, and maintaining highly-optimized communication runtimes for Deep Learning frameworks and HPC programming interfaces on GPU clusters. This involves system software development, parallel programming interface contributions, and proof-of-concept creation for new designs and hardware features. | Serve | 7 |
| Senior Manager, Site Reliability Engineering Senior Manager of Site Reliability Engineering to lead and reshape IT operations at scale, building AI-powered systems for reliability, speed, and employee experience. Focuses on transforming Incident, Problem, and Change Management using observability, AI insights, and orchestration to move towards predictive and autonomous operations. | Serve | 7 |
| Manager, AI Networking Performance Research and Analysis Manager for AI Networking Performance Research and Analysis at NVIDIA, focusing on optimizing networking technologies (NIC, Switch) for AI workloads like LLM training and inference. The role involves end-to-end performance strategy, from pre-silicon modeling to GA, and building telemetry frameworks and dashboards for performance tracking and root cause analysis. Requires strong experience in high-performance networking, cluster performance, and managing engineering teams, with a focus on Python, Bash, and C/C++. | ServeAgent | 7 |
| Senior Software Engineer, AI Inference Senior Software Engineer focused on optimizing and scaling AI inference for large language models, working with customers and contributing to open-source projects like vLLM. | Serve | 7 |
| Senior Software Engineer, Machine Learning Inference Senior Software Engineer role focused on designing and implementing inference software optimizations for NVIDIA TensorRT and TensorRT-LLM to accelerate AI applications on NVIDIA GPUs. Involves C++, Python, and CUDA development, collaboration with AI experts, and optimization of deep learning frameworks and compilers. | Serve | 7 |
| Senior Math Libraries Engineer - Sparsity in AI Software engineer to design and develop C++ libraries and tools for unstructured sparsity in Deep Learning (DL) and High-Performance Computing (HPC) on NVIDIA GPUs. This involves DSL specifications, on-demand code generation, and enabling the system in Python/PyTorch. The role focuses on performance evaluation, library quality, and collaboration with product management. | Serve | 7 |
| Senior Software Engineer, JAX Senior Software Engineer focused on performance optimizations for JAX, a deep learning framework, to build a scalable platform for data, training, and analysis. The role involves developing core JAX components, working with AI researchers, and building tools to improve AI system development efficiency. | Serve | 7 |
| Senior AI and FSI Developer Technology Engineer Senior AI and FSI Developer Technology Engineer at NVIDIA focused on optimizing AI and HPC workloads on NVIDIA CPUs and GPUs for the financial services industry. The role involves researching, designing, and developing techniques to accelerate these workloads, profiling and eliminating performance bottlenecks, and collaborating with internal and external experts to influence future hardware and software designs. The engineer will also publish and present their work. | Serve | 7 |
| Senior Software Engineer - NIM Factory Container and Cloud Infrastructure Senior Software Engineer role focused on container and cloud infrastructure for NVIDIA Inference Microservices (NIMs) and hosted services. The role involves designing and implementing container strategies, building enterprise-grade software for container build, packaging, and deployment, and improving reliability, performance, and scale across thousands of GPUs, with a focus on disaggregated LLM inference. | Serve | 7 |
| Developer Technology Engineer, HPC and AI NVIDIA is seeking a Developer Technology Engineer to research, develop, and optimize deep learning, machine learning, and HPC workloads on NVIDIA's accelerated computing platform. The role involves working with customers and internal teams to address real-world use cases and performance challenges. | Serve | 7 |