NVIDIA currently has 496 active AI-related job listings. The majority of these roles, 52%, are focused on serving infrastructure, with agents representing another significant segment at 23%. Engineering is the dominant function, with 441 positions. The United States leads hiring geographies with 287 roles, followed by China with 64. Frequent tech tags include model_serving, inference_infra, and agent_orchestration, suggesting a focus on deployment and management of AI models. Over the last 30 days, NVIDIA posted 214 new AI roles, a 27% decrease compared to the previous 30-day period.
Currently tracking 440 active AI roles, down 50% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $100k–$575k (avg $262k).
NVIDIA currently has 487 active AI-related roles in our index. The most common open titles are: Deep Learning Performance Architect (4), Senior Deep Learning Performance Architect (4), AI Research Scientist (3), Developer Technology Engineer - AI (3), Manager, Deep Learning Algorithms (3). Most positions are in Engineering and Research.
NVIDIA's active AI hiring is concentrated in: serving infrastructure (54%), agents (21%), application (8%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.
NVIDIA is hiring AI talent in: United States (286 roles), China (59 roles), Israel (50 roles), Germany (21 roles).
Job postings at NVIDIA most frequently reference: model serving, inference infra, agent orchestration, llm observability, multimodal.
In the past 30 days, NVIDIA has posted 110 new AI-related roles. That is a -50% change versus the prior 30 days (218 → 110).
| Title | Stage | AI score |
|---|---|---|
| Solutions Architect, Agentic AI NVIDIA is seeking Solutions Architects to build and deploy agentic AI applications at scale for enterprises, focusing on integrating enterprise data, developing multi-modal dialogue systems, and task-specific agents. The role involves working with agentic frameworks, providing feedback to improve software products, and educating vertical teams. | Agent | 8 |
| Senior Solutions Architect, Generative AI Senior Solutions Architect role focused on customer engagements, improving AI workload performance, and developing proof-of-concepts for Generative AI solutions (LLMs, recommenders) using NVIDIA software and technologies. Requires strong coding, GPU optimization, and communication skills. | ServeAgent |
| 8 |
| Principal Deep Learning Communication Architect NVIDIA is seeking a Principal Deep Learning Communication Architect to lead the technical roadmap for communication libraries across next-generation platforms, ensuring seamless scaling of models to massive clusters. The role involves designing and optimizing communication primitives for heterogeneous interconnects, co-designing with application developers and silicon architects, and developing analytical models for system behavior. Expertise in parallel computing, HPC/distributed deep learning, inference engines, and GPU architecture is required. | ServeAgent | 8 |
| Technical Lead, GenAI - Autonomous Vehicles This role is a Technical Lead focused on Generative AI within Autonomous Vehicles, engaging with developer ecosystems and partners to promote NVIDIA's AI platforms. The candidate will act as a technical advisor, develop expertise in NVIDIA's platforms, create enablement resources, and represent partner needs internally. Requires a strong technical background in AI, AV systems, and GenAI model development, with experience in production code, DevOps, and DL/RL frameworks. | Agent | 8 |
| Senior Software Engineer, Computer Vision - Autonomous Vehicles Senior Software Engineer at NVIDIA for Autonomous Vehicles, focusing on Computer Vision and Machine Learning for offline perception tasks. Responsibilities include advancing DL components for training and inference, developing tools for large datasets, and integrating DL algorithms into large-scale pipelines. | DataServe | 8 |
| Senior Architect NVIDIA is seeking a Senior Architect to lead the development of software infrastructure for AI-driven scientific discovery in chemistry and materials science. The role involves shaping NVIDIA ALCHEMI and its ecosystem, translating AI research (ML interatomic potentials, generative modeling) into product direction, and engaging with internal/external stakeholders. The ideal candidate has a PhD or equivalent experience, 8+ years of AI/ML software development for chemistry/materials, strong GPU computing and ML framework experience, and expertise in scientific software architecture. | Ship | 8 |
| AI for Design Engineer Develop and deploy AI agents and frameworks for hardware verification tasks, processing codebases and optimizing retrieval/generation algorithms for enterprise data. | Agent | 8 |
| Engineering Manager, Prediction and Planning - Autonomous Vehicles Engineering Manager for NVIDIA's Autonomous Vehicles division, leading teams to build and scale AI-native autonomous driving systems, integrating classical safety stacks with foundation models and large-scale AI systems from research to production. | ShipAgent | 8 |
| Senior Integration Engineer - Autonomous Vehicles NVIDIA is seeking a Senior Integration Engineer to work on their end-to-end autonomous driving application, focusing on integrating modular software components and optimizing performance on heterogeneous hardware architectures. The role involves defining software architecture for L2/L3/L4 autonomous driving solutions, performing in-vehicle and simulation testing, and developing efficient C++ code using CUDA. | Agent | 8 |
| Senior Integration Engineer - Autonomous Vehicles Senior Integration Engineer for NVIDIA's end-to-end autonomous driving application, focusing on integrating software components, optimizing performance, and developing efficient C++ code on heterogeneous hardware architectures (including GPUs) for L2/L3/L4 autonomous driving solutions. | AgentServe | 8 |
| Senior Product Manager, AI Frameworks Product Manager for AI Frameworks at NVIDIA, focusing on Recommender Systems and Generative Recommendation Models. The role involves building products for frontier RecSys and Generative Recommendation Models on Nvidia systems, enabling researchers and operators, and pushing the boundaries of what is possible in research-to-production. Responsibilities include creating and optimizing pre-training/inference and post-training frameworks, developing product strategy, roadmaps, and go-to-market plans, and collaborating with internal and external customers. Requires experience with training/inference post-training and optimization software, GenAI/ML concepts, large-scale distributed systems, and technical product management. | Post-trainServe | 8 |
| Senior Product Manager, AI Inference - Dynamo Product Manager for NVIDIA Dynamo, a distributed inference framework for LLMs and Generative AI. Focuses on defining the roadmap for high-scale serving, optimizing hardware-software co-design, and developing agentic inference capabilities. Collaborates with engineering, open-source communities, and customers to integrate model evaluation into workflows. | ServeAgent | 8 |
| AI and FSI Developer Technology Engineer - New College Grad 2026 NVIDIA is seeking an AI and FSI Developer Technology Engineer to optimize AI and HPC workloads on NVIDIA GPUs and CPUs, focusing on performance tuning and eliminating bottlenecks for financial markets. The role involves research, development, analysis, and collaboration with experts to improve performance across the stack, from algorithms to kernels. The engineer will also publish and present their work and influence future hardware/software designs. | Serve | 8 |
| Senior Software Engineer, Platform Engineering Senior Software Engineer to build next-generation AI platforms and products, focusing on agentic AI systems, RAG, and scalable infrastructure for enterprise workflows. | Agent | 8 |
| Solutions Architect, Physical AI and Robotics NVIDIA is looking for a Solutions Architect to guide partners in building enterprise Physical AI systems using Omniverse, Cosmos, synthetic data, and coding-agent-assisted digital twins workflows. The role involves technical advising on simulation, digital twins, robotics, industrial autonomy, and auto, focusing on architecture, compute, testing, and rollout strategies. Key responsibilities include guiding partners on synthetic data generation, evaluation methods, using coding agents for development acceleration, defining benchmarks, advising on compute infrastructure for simulation and inference, and building reference architectures. | AgentData | 8 |
| Senior Systems Software Engineer, E-commerce AI Platform - GeForce NOW Senior Systems Software Engineer to architect and deploy production-grade AI agents for NVIDIA's e-commerce platform, focusing on personalization, logistics, and customer experience. Requires expertise in Python, Java, GoLang, distributed systems, and AI frameworks like LangChain/LangGraph. | Agent | 8 |
| Senior SOC Product Architect Physical AI Platforms This role focuses on architecting physical AI platforms for automotive and robotics, specifically defining the SoC architecture for embedded computer vision and AI systems. The individual will analyze use cases, map requirements to hardware/software features, define system requirements, and drive recommendations into product roadmaps. The role involves deep benchmarking, customer interaction, technical leadership, and mentorship, with a strong emphasis on functional safety (ISO 26262, SOTIF). | Serve | 8 |
| Senior Technical Program Manager - Agentic System Senior Technical Program Manager to drive and coordinate cross-functional teams for large-scale technical projects in agentic AI, connecting foundation models with real-world applications for edge deployment and AI workflows. | Agent | 8 |
| Principal Software Engineer - Enterprise AI Platform Principal Software Engineer to lead security foundations for autonomous, self-evolving agents in an enterprise setting. This role involves defining security requirements, designing scalable architectures with guardrails, implementing isolation and access controls, building secure data access pathways, establishing observability and auditing, and operating a continuous evaluation framework for agent behavior. The goal is to enable developer velocity while ensuring robust safety and security for agents that generate and execute code and access data. | Agent | 8 |
| Senior Power Analysis and Optimization Engineer Senior Engineer to apply AI/ML and LLMs to power analysis and optimization for NVIDIA's GPUs and SoCs. Focus on developing and productionizing ML/RL models and custom LLMs to improve energy efficiency, interpret power data, and recommend optimizations. Involves RTL analysis, Verilog prototyping, and automation. | ServeData | 8 |
| Senior Machine Learning Applications and Compiler Engineer, LPX Develops algorithms and optimizations for NVIDIA's LPX inference and compiler stack, focusing on mapping neural network workloads onto future NVIDIA platforms and optimizing end-to-end inference performance. Requires strong software engineering, compiler/runtime development, and deep learning framework experience. | Serve | 8 |
| Senior Software Engineer, TensorRT-LLM NVIDIA is seeking a Senior Software Engineer for its TensorRT-LLM team to develop and scale inferencing software for LLMs and Generative AI. The role involves crafting robust inferencing software, performing benchmarking and profiling for GPU applications, writing high-quality Python code for LLM inference, and improving the TensorRT-LLM library. Collaboration with software, research, and product teams is key. | Serve | 8 |
| Senior Software Engineer – TensorRT Edge-LLM Senior Software Engineer to develop and optimize a state-of-the-art inference framework for Large Language, Vision-Language, and Multimodal models on edge and embedded platforms, focusing on real-time performance and constrained environments. | Serve | 8 |
| Senior Performance Engineer - Deep Learning Senior Performance Engineer at NVIDIA focused on optimizing Deep Learning models and frameworks (PyTorch, JAX) for NVIDIA GPUs. The role involves building and supporting Transformer Engine, collaborating on systems research for performance improvements, implementing and benchmarking new DL models, contributing to MLPerf, and engaging with the open-source community and enterprise customers. It also involves influencing future hardware and software design. | ServePost-train | 8 |
| Senior System Software Engineer, 3D Computer Vision Senior System Software Engineer focused on 3D Computer Vision at NVIDIA, involving the development and deployment of advanced neural reconstruction models for generating 3D scenes. The role requires strong programming skills in Python and C/C++, a background in computer vision and deep learning, and experience with production-grade software development. | Post-trainServe | 8 |
| Senior AI Research Scientist, Robotics Digital Twins Senior AI Research Scientist role focused on developing digital twins for chemical, biological, and physical laboratories, integrating AI agents with science experiments, and collaborating with robotics and software engineers. Requires a Ph.D. and 5+ years of AI research experience in robotics. | Agent | 8 |
| Senior Software Engineer, Quantized Inference Senior Software Engineer focused on optimizing quantized inference for LLMs by implementing recipes, developing kernels, and collaborating on inference engines like vLLM and TRT-LLM. The role involves model export pipelines, benchmarking, and data analysis tooling. | Serve | 8 |
| Senior Compiler Engineer, AI Inference Performance NVIDIA is seeking a Senior Compiler Engineer to optimize AI inference performance for their Deep Learning & AI Compiler (DLC) team. The role involves analyzing deep learning networks, developing compiler optimization algorithms, and collaborating with framework and architecture teams to accelerate next-generation deep learning software for various AI applications. | Serve | 8 |
| Senior Compiler Engineer, AI Inference Platforms NVIDIA is seeking a Senior Compiler Engineer to join its Deep Learning & AI Compiler (DLC) team. The role involves analyzing deep learning networks, developing compiler optimization algorithms, and collaborating with framework and architecture teams to accelerate AI inference performance on NVIDIA GPUs. The compiler is critical for data centers, personal devices, automotive, and robotics, focusing on inference performance, build time, memory footprints, and ease of use. | Serve | 8 |
| Research Scientist, Security and Privacy - PhD New College Grad 2026 Research Scientist focused on security and privacy for AI systems, aiming to develop hardware, software, and algorithms for trustworthy AI with verifiable protection. Requires a PhD and expertise in areas like computer architecture, programming languages, applied cryptography, or AI/ML algorithms, with a strong publication record. | Post-train | 8 |
| Principal GenAI Engagement Lead, Partner Platforms This role focuses on driving the technical integration of NVIDIA's Generative AI software with enterprise partners, including ISVs and CSPs. The Principal GenAI Engagement Lead will build trusted relationships, accelerate adoption, and influence product direction by designing and shipping methodologies, code, and reference architectures for RAG, LLM inference, and Multi-Agent workflows. The role requires a strong background in AI/ML, deep learning, and enterprise-grade GenAI systems, with experience in various LLM application stages and MLOps. The individual will act as the key technical lead, ensuring the deployment of robust, scalable GenAI solutions. | AgentServe | 8 |
| AI Chip Design Engineer - New College Grad 2026 NVIDIA is seeking an AI Chip Design Engineer to develop and integrate AI capabilities into verification tasks. The role involves creating AI agents to enhance productivity, building production infrastructure for these agents, and optimizing algorithms for enterprise data. Requires strong proficiency in LLM libraries, GPU/CPU architectures, and HW verification methodologies. | Agent | 8 |
| Senior Solutions Architect – Simulation Solutions 3D Reconstruction This role focuses on developing and scaling AI platforms for simulation and 3D reconstruction, particularly within the Omniverse ecosystem. The Senior Solutions Architect will act as a technical advisor, prototype solutions, implement intricate technical systems, provide technical enablement, and advocate for partner needs. The role requires expertise in AI, systems knowledge, autonomous systems, simulation, generative AI, Python, C++, DL/RL frameworks, computer vision, and 3D reconstruction. | AgentServe | 8 |
| AI Chip Design Engineer - New College Grad 2026 NVIDIA is seeking an AI Chip Design Engineer to develop and integrate AI capabilities into verification tasks, focusing on building and maintaining infrastructure for AI agents that process large codebases and optimize verification flows. The role involves developing retrieval and generation algorithms, integrating AI optimizations, and working with HW engineering teams. | Agent | 8 |
| Developer Relations Manager – AI Natives NVIDIA is seeking a Developer Relations Manager to engage with AI-native companies, helping them design, optimize, and scale their AI platforms on NVIDIA technologies. The role involves advising founders and engineering teams on building agentic systems, AI copilots, and multimodal applications, with a focus on accelerating training, optimizing inference, and delivering AI experiences. The ideal candidate has deep technical expertise in AI systems, developer platforms, and large-scale inference infrastructure. | ServeAgent | 8 |
| Senior AI Performance and Efficiency Engineer Senior AI/ML Performance and Efficiency Engineer focused on optimizing GPU cluster performance for AI/ML researchers by addressing infrastructure and application bottlenecks. This role involves building tools, analyzing efficiency, and collaborating across teams to improve hardware, software, and infrastructure usage for various ML workloads like Robotics, Autonomous vehicles, LLMs, and Videos. | Serve | 8 |
| Engineering Manager, AI Developer Technology Engineering Manager for NVIDIA's AI Developer Technology team, focused on leading a team to optimize and develop algorithms for Deep Learning and Machine Learning applications, influencing next-generation hardware/software, and collaborating with customers and internal teams. The role involves optimizing training and inference performance on NVIDIA hardware. | ServePost-train | 8 |
| Senior Developer Technology Engineer - AI Senior Developer Technology Engineer focused on researching and optimizing AI/ML workloads for GPU acceleration, involving deep analysis, performance tuning, and collaboration with the developer community and internal teams to influence next-generation hardware and software design. | Serve | 8 |
| Senior Design Automation Engineer, Applied AI NVIDIA is seeking an Applied AI Engineer to lead end-to-end solution development for timing and constraint analysis workflows in VLSI/ASIC design. The role involves data generation, model training, orchestration, and building autonomous agents that interact with timing tools. The engineer will develop AI-driven solutions, integrate data sources, implement scalable orchestration, and build interpretable AI pipelines using GNNs, LLMs, and reasoning engines. Experience with Python, PyTorch/TensorFlow, graph/agentic AI frameworks, and EDA tools is required. | AgentData | 8 |
| Senior Product Architect, Storage NVIDIA is seeking a Senior Product Architect to design and validate AI storage infrastructure, focusing on optimizing systems for large-scale foundation model training, disaggregated inference, and agentic AI pipelines. The role involves architecting end-to-end reference architectures, defining system-level architectures, and collaborating with partners and customers to deliver proof-of-concepts. | AgentServe | 8 |
| Technical Marketing Engineer This role leads complex, cross-functional programs for next-generation generative AI systems, focusing on media content creation. It involves translating research into execution roadmaps, defining program plans, and managing model release readiness across various stages from research to product integration. The role requires strong program management skills in AI/ML and a solid understanding of generative AI systems. | ShipPretrain | 8 |
| Senior Software Engineer, Video Analytics Senior Software Engineer role focused on building large-scale distributed Vision AI platforms for video analytics using NVIDIA Metropolis. The role involves designing and developing functionalities for video processing, integrating VLMs, CV models, and LLMs, and optimizing performance on NVIDIA hardware. Requires strong software development experience with ML systems, C++, Python, and GPU acceleration. | ShipServe | 8 |
| Senior Software Engineer, Robotics - Isaac Lab Senior Software Engineer to join the Isaac Lab team, focusing on developing a platform for robot learning, including perception-in-the-loop reinforcement learning, multi-agent/multi-task learning, and VLA & RL integration. The role involves sim-to-real efforts, defining training workflows, and collaborating with research teams to advance humanoid robots. | ShipData | 8 |
| Senior HPC Performance Engineer - AI for Science at Scale Senior HPC Performance Engineer focused on optimizing large-scale, CUDA-backed ML training frameworks for AI in Science applications, particularly in digital biology and chemistry. The role involves kernel design, GPU porting, distributed learning, and algorithmic improvements within HPC software stacks. | ServePost-train | 8 |
| Senior AI Application Developer - GPU and SOC Architecture Modeling Senior AI Application Developer role focused on developing and deploying scalable GenAI applications to accelerate GPU/SOC architecture modeling. The role involves integrating LLMs into existing workflows, collaborating with hardware architects and infrastructure engineers, and researching emerging AI technologies. Requires proficiency in C++, Python, ML frameworks, and hands-on experience with LLMs and multimodal models. | Agent | 8 |
| Architect, AI Solutions Engineering NVIDIA is looking for an AI Solutions Architect to scale internal AI platforms and solutions for thousands of developers. The role involves identifying AI opportunities, setting system outcomes, optimizing performance and cost, and collaborating with AI product vendors. Requires strong experience in building large-scale distributed systems and hands-on experience with LLMs, RAG, fine-tuning, and agentic orchestration. | AgentServe | 8 |
| Manager, Deep Learning Algorithms Manager for Deep Learning Algorithms at NVIDIA, focusing on productizing DL models, optimizing inference, and leading engineering teams. The role involves working with LLMs/VLMs, inference optimization, and collaborating across NVIDIA to develop state-of-the-art algorithms for GPU-accelerated platforms. | Serve | 8 |
| Distinguished Engineer, JAX Distinguished Engineer to develop NVIDIA's AI platform, focusing on performance optimizations in deep learning frameworks using JAX. The role involves designing and implementing core JAX components, driving peak performance on NVIDIA products, and building tools to increase the efficiency of AI-based system development teams. It bridges numerical computing, simulation, and deep learning research with real-world applications. | Serve | 8 |
| Distinguished Engineer - Dynamo Distinguished Engineer role focused on NVIDIA Dynamo, an AI inferencing platform. The role involves technical leadership, driving product direction, and contributing to open-source projects to achieve state-of-the-art performance and scalability for AI inference across modalities on NVIDIA hardware. | Serve | 8 |
| Principal Software Engineer - Dynamo Principal Software Engineer for NVIDIA Dynamo, an open-source platform for efficient, scalable inference of large language and reasoning models in distributed GPU environments. Focuses on Kubernetes serving, scalability, disaggregated serving, dynamic GPU scheduling, intelligent routing, and distributed KV cache management. | ServeAgent | 8 |