NVIDIA currently has 496 active AI-related job listings. The majority of these roles, 52%, are focused on serving infrastructure, with agents representing another significant segment at 23%. Engineering is the dominant function, with 441 positions. The United States leads hiring geographies with 287 roles, followed by China with 64. Frequent tech tags include model_serving, inference_infra, and agent_orchestration, suggesting a focus on deployment and management of AI models. Over the last 30 days, NVIDIA posted 214 new AI roles, a 27% decrease compared to the previous 30-day period.
Currently tracking 440 active AI roles, down 53% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $100k–$575k (avg $262k).
NVIDIA currently has 487 active AI-related roles in our index. The most common open titles are: Deep Learning Performance Architect (4), Senior Deep Learning Performance Architect (4), AI Research Scientist (3), Developer Technology Engineer - AI (3), Manager, Deep Learning Algorithms (3). Most positions are in Engineering and Research.
NVIDIA's active AI hiring is concentrated in: serving infrastructure (54%), agents (21%), application (8%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
NVIDIA is hiring AI talent in: United States (286 roles), China (59 roles), Israel (50 roles), Germany (21 roles).
Job postings at NVIDIA most frequently reference: model serving, inference infra, agent orchestration, llm observability, multimodal.
In the past 30 days, NVIDIA has posted 110 new AI-related roles. That is a -50% change versus the prior 30 days (218 → 110).
| Title | Stage | AI score |
|---|---|---|
| Engineering Manager, Inference Benchmarking — AI Perf Engineering Manager for NVIDIA's AIPerf platform, a standard for assessing LLM serving performance. The role involves leading a team to build and advance the platform, focusing on core infrastructure, accuracy of benchmark results, and advising on upstream engine integrations for various AI workloads (LLM, multimodal, diffusion, computer vision). Requires strong systems engineering, inference infrastructure, and open-source community experience. | Serve | 8 |
| Senior Software Engineer, Generative AI Systems Senior Software Engineer role focused on building and scaling Generative AI systems, including LLMs, agentic AI, and RAG pipelines. Responsibilities include designing and developing infrastructure for ML training and inference, creating evaluation frameworks, optimizing RAG pipelines, and building backend services and APIs. Requires strong software engineering fundamentals, experience with ML systems, distributed infrastructure, and GenAI workflows. |
| AgentServe |
| 8 |
| GPU Performance Engineer - Neural Reconstruction GPU Performance Engineer focused on optimizing neural reconstruction and Gaussian Splatting workloads. This role involves profiling, identifying bottlenecks, and improving performance in CUDA, PyTorch, and C++ for training and rendering, while ensuring reconstruction quality is maintained. It requires strong programming, GPU optimization, and performance analysis skills, with collaboration across research and engineering teams. | ServeData | 8 |
| Senior AI Tools Engineer, SRE Operations - GeForce NOW This role focuses on building and deploying AI/ML tools, specifically LLM- and Agent-based systems, to analyze production data for a global service (GeForce Now). The goal is to automate root cause analysis for incidents and predict future service trends, requiring strong data pipeline management and expertise in AI frameworks. | AgentData | 8 |
| Principal Machine Learning Engineer, Accelerated Apache Spark This role focuses on applying ML/AI to optimize and accelerate Apache Spark workloads on NVIDIA GPUs, involving performance prediction, adaptive systems, and developing AI agents for system issue resolution and optimization. The role requires significant experience in ML/DL solution design, productionization, and large-scale data processing platforms like Spark, with a focus on LLM/GenAI, reinforcement learning, and adaptive ML systems. | AgentServe | 8 |
| Director, AI Enablement Director of AI Enablement at NVIDIA, responsible for developing and implementing an AI enablement roadmap to accelerate AI adoption and agentic developments across various internal workflows. The role focuses on building tools, services, and blueprints for NVIDIA AI developers, optimizing AI development, and transforming NVIDIA into an AI-native company. | Agent | 8 |
| AI Software Engineer, Kernel Libraries - New College Grad 2026 AI Software Engineer focused on developing inference systems software stack, including libraries, code generators, and GPU kernels for NVIDIA's hardware. The role involves innovating for efficient AI inference, optimizing kernels, designing abstractions for LLM serving engines, and building JIT compilers and runtimes. Collaboration with internal teams and contributions to open-source projects like FlashInfer, vLLM, and SGLang are expected. | Serve | 8 |
| Senior AI Infrastructure Software Engineer - DGX Cloud NVIDIA is seeking a Senior AI Infrastructure Software Engineer to design, build, and maintain AI platforms for large-scale AI training, inferencing, fine-tuning, and Agentic AI in production. The role involves developing platform and tools for AI/ML workload efficiency, resiliency, and observability, with a focus on distributed systems and Kubernetes. | Serve | 8 |
| Software Engineer - AI Research Clusters Software Engineer to build and maintain GPU clusters for internal AI researchers, focusing on reliability, performance, and self-service. The role involves applying AIOps and Agentic AI to reduce operational toil and support the training, fine-tuning, and deployment of advanced ML models. | Serve | 8 |
| Senior Engineer - AI Agents and Systems Senior Engineer role focused on deploying advanced AI agent frameworks and local runtimes to Windows and NVIDIA GeForce RTX GPUs, ensuring open-source AI agents run locally, safely, and efficiently on consumer PCs, and creating the foundation of the desktop AI operating system. | Agent | 8 |
| Senior Engineer - AI Agents and Systems Senior Engineer role focused on deploying advanced AI agent frameworks and local runtimes to Windows and NVIDIA GeForce RTX GPUs, ensuring open-source AI agents run locally, safely, and efficiently on consumer PCs. The role involves leading development for the foundation of the desktop AI operating system by combining local inference with robust privacy routers and sandboxed execution. | Agent | 8 |
| Senior Performance Compiler Engineer - Triton Senior Performance Compiler Engineer to work on the open-source Triton compiler project, focusing on using compilers to improve AI performance on NVIDIA GPUs for large language models, agents, and other AI applications. The role involves investigating GPU hardware, designing and implementing compiler technology using MLIR to optimize kernel descriptions for efficient GPU code generation, and collaborating with internal teams. | Serve | 8 |
| Senior Systems Engineer, Neural Graphics Senior Systems Engineer role focused on integrating AI and traditional rendering techniques for real-time visual experiences. The role involves taking innovative techniques like AlpaDreams and driving them into production-ready, real-time pipelines, owning the end-to-end path from prototype to shipping product, and solving complex systems challenges related to latency, memory, and throughput. Requires deep expertise in graphics and AI, with a strong track record of shipping impactful products and experience with systems-level thinking. | ShipAgent | 8 |
| Senior GPU System Architect Seeking a Senior GPU System Architect to design multi-GPU scale-up and scale-out systems for AI and HPC datacenters. The role involves defining system architectures that integrate GPU compute, memory, and interconnects for optimal AI performance and scalability. Requires deep experience in system-level fabric/networking architecture and hardware-software co-design. | Serve | 8 |
| AI Research Engineer - Applied Scientist Compilers AI Research Engineer/Applied Scientist focused on Compilers/Low-level optimization to develop AI compiler solutions for NVIDIA's software stack and GPU acceleration. Responsibilities include applying AI to compilation, implementing AI-based solutions for GPU programming, building training pipelines (fine-tuning, RL), defining model I/O, developing evaluation frameworks, prompt engineering, integrating learned policies, prototyping models, creating datasets, and applying RL for optimization. | Post-trainServe | 8 |
| Senior Research Engineer, Robotics Systems Senior/Principal Engineer in robotics systems, focusing on foundation models and full-stack technology for humanoid robots. Responsibilities include designing teleoperation software, optimizing control stacks, deploying neural network models on hardware, and collaborating on the MLOps lifecycle. Requires strong robotics and software engineering background, with experience in real-time control and deploying ML models on robotic hardware. | ShipData | 8 |
| Senior Perception Engineer - Autonomous Vehicles Senior Perception Engineer at NVIDIA focused on developing and productizing autonomous driving solutions using deep learning and multi-sensor fusion. The role involves applied research, algorithm development, and ensuring solutions meet production requirements for safety, latency, and robustness. | ShipPost-train | 8 |
| Senior Software Engineer, Agentic AI Senior Software Engineer to develop core libraries for Agentic Applications, focusing on building foundational technology, scalable capabilities, reusable blocks, and high-quality libraries to accelerate developer productivity and ensure agent quality. The role involves benchmarking, identifying bottlenecks, and optimizing performance, cost, and latency for agents. Collaboration with teams on data pipelines, RAG, vector databases, and GPU-optimized workflows is expected. | Agent | 8 |
| Software Engineering Intern, AI Tools - Fall 2026 NVIDIA is seeking a Software Engineering Intern to join its AI Tools and Infrastructure team, focusing on Agentic AI. The intern will work with LLMs and orchestration frameworks to design, build, and deploy intelligent agents and AI tools, gaining exposure to the full lifecycle of agent development from conceptualization to deployment. | Agent | 8 |
| Senior Deep Learning Performance Architect Senior Deep Learning Performance Architect at NVIDIA to design and evaluate hardware architectures for AI/HPC applications, focusing on LLM inference and training performance, and optimizing system bottlenecks. | ServePost-train | 8 |
| Senior Data Center Performance Engineer - Benchmarking and Optimization Senior Data Center Performance Engineer at NVIDIA focused on benchmarking and optimizing data center platforms for AI training, inference, and HPC workloads. Responsibilities include designing benchmarks, characterizing workloads, identifying bottlenecks, and driving performance improvements through system tuning and architectural recommendations. | Serve | 8 |
| Senior GenAI Engagement Lead, Partner Platforms This role focuses on driving the technical integration of Generative AI software with enterprise partners, involving hands-on design and deployment of RAG, LLM inference, and Multi-Agent workflows. The position requires deep technical expertise in AI/ML, partner engagement, and understanding of the GenAI lifecycle, with a focus on production deployments and influencing product roadmaps. | AgentServe | 8 |
| NCX Engineer, AI Accelerator This role focuses on engineering and deploying AI infrastructure and solutions for strategic customers, optimizing large-scale training and inference workloads on NVIDIA's AI platform. It involves MLOps, Kubernetes, GPU scheduling, and performance tuning, with a strong emphasis on customer-facing technical support and collaboration. | ServePost-train | 8 |
| Senior AI Solutions Architect NVIDIA is seeking an AI Solutions Architect with deep expertise in AI solutions and scalable data center infrastructure. The role involves embedding NVIDIA software into customer architectures, improving application performance, and establishing technical foundations for next-generation AI systems. Responsibilities include supporting business development, working directly with developers and customers, analyzing architectures for acceleration opportunities, and delivering trainings. | ServeAgent | 8 |
| Senior Deep Learning Framework Communications Engineer Senior Deep Learning Framework Communications Engineer at NVIDIA, focusing on integrating and optimizing communication libraries (NCCL, NVSHMEM) within AI frameworks (PyTorch, TRT-LLM, vLLM, JAX) to enhance performance for large-scale AI training and inference. The role involves deep analysis of AI workloads, compiler improvements, and kernel authoring for multi-GPU systems. | Serve | 8 |
| Senior Scientific Machine Learning Engineer – Earth-2 Develops and enhances machine learning frameworks (NVIDIA PhysicsNeMo, NVIDIA Earth2Studio) for scientific ML technology in weather, climate, and earth system modeling. Focuses on implementing new deep learning techniques and enhancing Earth-2 technologies. | Post-train | 8 |
| Senior Solutions Architect, Generative AI Data Processing Senior Solutions Architect role focused on assisting customers in deploying Generative AI solutions, particularly for data processing and agentic workflows, using NVIDIA's AI technology stack. The role involves technical advisory, system design, and implementation at scale, with a strong emphasis on Deep Learning, LLMs, and GPU technologies. | AgentServe | 8 |
| Director, System Software Engineering - Metropolis Accelerated and Inferencing Software NVIDIA is seeking a Director of System Software Engineering to lead teams responsible for the full lifecycle of Vision AI strategy, from model onboarding to production deployment. The role focuses on transforming foundation models into real-time, GPU-accelerated video intelligence systems, scaling multimodal reasoning, and enabling agentic development workflows. Key responsibilities include architecting and operationalizing inference acceleration, driving implementations of frameworks like TensorRT and VLLM, collaborating with partners on custom models, and ensuring performance benchmarking. The ideal candidate has extensive experience in deep learning, GPU optimization, and leading engineering teams in embedded and enterprise platforms. | ServeAgent | 8 |
| Director, Isaac for Healthcare Engineering Director of Engineering for NVIDIA's Isaac for Healthcare initiative, focusing on building a platform for healthcare robotics companies to develop, simulate, train, and deploy physical AI systems. The role involves platform leadership, team building, partner enablement, technical strategy, and cross-functional collaboration, with a strong emphasis on shipping sophisticated software platforms at scale. | ShipData | 8 |
| Manager, Solutions Architecture - Global Partner Team Manager of Solutions Architecture for NVIDIA's Global Partner Team, focusing on leading technical engagements with GSIs and AI consulting firms. The role involves building and scaling Agentic AI services, providing architectural oversight for complex AI workflows, and collaborating with product and engineering teams. Requires deep technical expertise in Generative/Agentic AI, RAG, LLM orchestration, and AI infrastructure. | Agent | 8 |
| Senior Software Architect - Deep Learning and HPC Communications Senior Software Architect role at NVIDIA focused on designing and implementing next-generation data center platforms and scalable communication software for AI and HPC workloads. The role involves investigating performance bottlenecks, developing new communication technologies, exploring hardware/software co-design, and building proofs-of-concept to drive innovation in large-scale GPU clusters. | Serve | 8 |
| Director, Product Platform Retail and CPG Industries NVIDIA is seeking a Director to define and build the Retail & CPG Industries product platform. This role involves architecting and developing a platform leveraging NVIDIA's full stack, including Agentic AI and accelerated computing, to reshape digital commerce, supply chains, and intelligent stores. The platform will utilize NVIDIA Nemo microservices and Nemotron models, with a focus on Agentic AI for various business functions. The ideal candidate will have hands-on development experience with AI agents, LLMs, RAG, and distributed systems, and will collaborate with engineering teams to deliver a production-ready, scalable platform. | AgentServe | 8 |
| Senior Solutions Architect - AI Factory Deployment Senior Solutions Architect focused on deploying and validating AI factories, specifically running and debugging AI/LLM workloads on GPU clusters. Responsibilities include setting up environments, executing benchmarks, resolving performance issues, building observability, and recommending optimizations. | Serve | 8 |
| Senior Software Engineer - VLM Microservices for Neural Reconstruction Senior Software Engineer to design, build, and optimize containerized inference execution for 3D Vision Language Models (VLMs) for neural reconstruction, turning research into production-grade software (NIMs). The role involves developing benchmarks, releasing and maintaining models, contributing to open-source projects like vLLM, and collaborating with research and product teams. Requires experience with AI distributed systems, inference platforms, Python/C++, and software engineering fundamentals. | ServePost-train | 8 |
| Senior Applied Machine Learning Engineer - VLSI Design NVIDIA is seeking a Senior Applied Machine Learning Engineer to build AI-driven software systems for circuit design, combining automation algorithms, DL models, and agentic workflows. The role involves working on pre-silicon and post-silicon hardware design data, circuit optimization, and AI systems for EDA/design automation, translating requirements into AI/ML and agentic system problems, and testing/releasing models and AI systems. | AgentData | 8 |
| Applied Machine Learning Engineer, Circuit Design - New College Grad 2026 NVIDIA is seeking an Applied Machine Learning Engineer for their Circuit Design team, focusing on building AI-driven software systems that combine automation algorithms, DL models, and agentic workflows to accelerate end-to-end circuit design. The role involves working with hardware design data, circuit optimization, and developing AI/ML solutions for EDA, with a focus on agent-driven design exploration and optimization. | Agent | 8 |
| Senior Solutions Architect, Generative AI Senior Solutions Architect role focused on customer engagements for NVIDIA's generative AI technologies, involving AI model training and deployment optimization, particularly for LLMs and recommenders in the consumer internet industry. Requires strong coding, GPU optimization, and communication skills. | ServeData | 8 |
| Principal AI and ML Infra Software Engineer, GPU Clusters This role focuses on enhancing the efficiency of AI and ML research on GPU clusters by collaborating with researchers to identify and address infrastructure deficiencies. The engineer will optimize performance, monitor resource utilization, and contribute to the AI/ML infrastructure ecosystem, keeping up-to-date with the latest AI/ML technologies. | Serve | 8 |
| Principal Cloud Services Software Engineer NVIDIA DGX Cloud Team is seeking a Principal Cloud Services Software Engineer to develop and optimize AI infrastructure services for large-scale AI training workflows. The role involves designing and implementing resilient, efficient services orchestrated by Kubernetes, with a focus on backend development, distributed systems, and high-performance computing. | ServeAgent | 8 |
| Senior Deep Learning Software Engineer - Autonomous Vehicles Senior Deep Learning Software Engineer focused on developing and productizing deep learning solutions for autonomous vehicles. The role involves training, fine-tuning, optimizing perception DNNs, applying quantization, improving DNN architectures, and enhancing inference speed and power consumption. It requires strong programming skills, experience with deep learning frameworks, computer vision tasks, and familiarity with CNNs and Transformer architectures. Experience with low precision inference, quantization, and NVIDIA software libraries is a plus. | ServePost-train | 8 |
| Compiler Engineer - AI Inference NVIDIA is seeking an AI Compiler Engineer to optimize kernel generation and computational graph optimizations for AI inference and training workloads on next-generation GPUs. The role involves hands-on development, collaboration on hardware/software co-design, and scaling AI deployments in datacenters. | ServePost-train | 8 |
| Senior Software Engineer, Metropolis Vision AI Senior Software Engineer to develop and optimize high-performance Vision AI pipelines and large-scale distributed services for processing video, image, and 3D data. The role involves crafting real-time systems, developing multi-modal perception, using simulation/synthetic data, and profiling/tuning GPU-accelerated inference pipelines. Collaboration with research and platform teams is key, with an emphasis on bringing research into production at scale. | ServePost-train | 8 |
| Senior Software Engineer, AI Networking Senior Software Engineer role focused on building and productizing ML tools for optimizing AI workloads (LLM training/inference) across GPU/CPU clusters, with a focus on networking and system resource utilization. Involves distributed deep learning, ML-based optimization techniques, and performance analysis. | ServeAgent | 8 |
| Deep Learning Architect, LLM Inference - New College Grad 2026 The role focuses on optimizing LLM inference server performance, workload characterization, and benchmarking for NVIDIA's GPUs. It involves collaborating with AI startups, developing performance tools, contributing to deep learning software projects, and guiding inference serving direction. | Serve | 8 |
| Senior Software Engineer - Robotics Senior Software Engineer role at NVIDIA focused on building Physical AI systems for humanoid robots. The role involves defining technical direction for generative AI workflows in robotics, spanning simulation, real-world deployment, and continuous learning. Responsibilities include building the humanoid reference platform, integrating NVIDIA products, and providing technical mentorship. Requires significant robotics software engineering experience, with a focus on AI-powered robots and robot learning. | ShipData | 8 |
| Senior AI-Native Systems Software Engineer, TensorRT Senior engineer to architect and build an AI-native framework using AI agents for software development, focusing on scaling, performance optimization, and integrating SOTA models for inference. | AgentServe | 8 |
| OEM Solutions Architect - AI Full Stack Public Sector NVIDIA is seeking a Solutions Architect to be the lead technical authority for Federal partnerships, focusing on deploying Generative AI at scale for U.S. Government agencies. The role involves architecting and optimizing the 'AI Factory,' leading POCs for NVIDIA's AI software stack, and navigating complex Federal security frameworks. The ideal candidate has extensive experience in full-stack data center architecture, the AI lifecycle (data curation, fine-tuning, inference orchestration), and strategic communication with both technical and leadership audiences within the public sector. | ServePost-train | 8 |
| Solutions Architect – OEM AI Solutions Architect role focused on integrating NVIDIA's software and tools with OEM partners' offerings, specifically for AI security and accelerated compute solutions. The role involves designing GPU-accelerated pipelines, developing proof-of-concept technologies, guiding Agentic AI workflows (RAG, GNN), and translating AI/cybersecurity research into deployable architectures for enterprise and government customers. | AgentData | 8 |
| Senior Staff Software Engineer - Agentic Automation Senior Staff Software Engineer to own engineering efforts for NVIDIA enterprise systems, transforming support into AI-infused automated resolution systems using LLM-based agents, tool calling, RAG, and orchestration frameworks. Requires full-stack experience, strong systems thinking, and incident management skills. | Agent | 8 |
| Solutions Architect, Inference Deployments This role focuses on building and deploying AI inference solutions at scale using NVIDIA's GPU technology and Kubernetes. The Solutions Architect will collaborate with engineering, DevOps, and customers to optimize and serve generative AI models, ensuring low-latency inference in enterprise environments. | Serve | 8 |