AI Hire Signal
JobsCompaniesTrendsInsightsWeekly
JobsStrategy timeline
AI Hire Signal

Tracking AI hiring across 200+ US tech companies. Stage, salary, and stack signals on every role — refreshed weekly.

Contact

Browse

JobsCompaniesTrendsInsightsWeekly

Resources

AboutSitemapRobots

Legal

PrivacyTerms
© 2026 AI Hire Signal·Not affiliated with companies shown

Currently tracking 440 active AI roles, down 53% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $100k–$575k (avg $262k).

Hiring
440 / 623
Momentum (4w)
↓-386 -53%
340 opens last 4w · 726 prior 4w
Salary range · avg $262k
$100k–$575k
USD · disclosed roles only
Tracked since
May '25
last role 4w ago
Hiring velocityscroll left for older weeks
1 new role
Dec 30
1 new role
Mar 10
1 new role
24
1 new role
Apr 28
4 new roles
May 12
5 new roles
19
3 new roles
26
3 new roles
Jun 2
2 new roles
9
1 new role
16
2 new roles
23
3 new roles
30
4 new roles
Jul 7
1 new role
14
2 new roles
28
4 new roles
Aug 11
6 new roles
18
2 new roles
25
3 new roles
Sep 1
8 new roles
15
3 new roles
22
6 new roles
29
2 new roles
Oct 6
2 new roles
13
3 new roles
20
6 new roles
27
9 new roles
Nov 3
8 new roles
10
8 new roles
17
4 new roles
24
11 new roles
Dec 1
9 new roles
8
14 new roles
15
10 new roles
22
8 new roles
29
107 new roles
Jan 5
22 new roles
12
45 new roles
19
32 new roles
26
59 new roles
Feb 2
64 new roles
9
63 new roles
16
83 new roles
23
83 new roles
Mar 2
88 new roles
9
97 new roles
16
72 new roles
23
215 new roles
30
158 new roles
Apr 6
250 new roles
13
199 new roles
20
332 new roles
27
304 new roles
May 4
189 new roles
11
131 new roles
18
102 new roles
25
129 new roles
Jun 1
122 new roles
8
49 new roles
15
40 new roles
22

NVIDIA currently has 496 active AI-related job listings. The majority of these roles, 52%, are focused on serving infrastructure, with agents representing another significant segment at 23%. Engineering is the dominant function, with 441 positions. The United States leads hiring geographies with 287 roles, followed by China with 64. Frequent tech tags include model_serving, inference_infra, and agent_orchestration, suggesting a focus on deployment and management of AI models. Over the last 30 days, NVIDIA posted 214 new AI roles, a 27% decrease compared to the previous 30-day period.

Auto-generated from active job postings · last refreshed 2026-05-24

Frequently asked questions

  • What AI roles is NVIDIA hiring for?

    NVIDIA currently has 487 active AI-related roles in our index. The most common open titles are: Deep Learning Performance Architect (4), Senior Deep Learning Performance Architect (4), AI Research Scientist (3), Developer Technology Engineer - AI (3), Manager, Deep Learning Algorithms (3). Most positions are in Engineering and Research.

  • What stage of AI development does NVIDIA focus on?

    NVIDIA's active AI hiring is concentrated in: serving infrastructure (54%), agents (21%), application (8%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.

  • Where is NVIDIA hiring AI talent?

    NVIDIA is hiring AI talent in: United States (286 roles), China (59 roles), Israel (50 roles), Germany (21 roles).

  • What technologies does NVIDIA's AI team work with?

    Job postings at NVIDIA most frequently reference: model serving, inference infra, agent orchestration, llm observability, multimodal.

  • How many AI roles has NVIDIA posted recently?

    In the past 30 days, NVIDIA has posted 110 new AI-related roles. That is a -50% change versus the prior 30 days (218 → 110).

Jobs (375)

434 AI · 1824 total active
FilteredFunctionEngineering×
Show
Active onlyAI only (≥ 7)
Stage
AllData · 17Pretrain · 20Post-train · 28Serve · 236Agent · 95Eval Gate · 5Ship · 33
Function
AllEngineering · 375Research · 57Product · 2
Country
AllUnited States · 259China · 55Israel · 43Germany · 21Switzerland · 18United Kingdom · 14India · 13Poland · 12Vietnam · 12Canada · 10Italy · 7Netherlands · 6Singapore · 6France · 5Taiwan · 4Finland · 2Spain · 2Armenia · 1Czech Republic · 1Hungary · 1Japan · 1Romania · 1South Korea · 1Sweden · 1
Sort
AI scoreRecentTitle
TitleStageFunctionLocationFirst seenAI score
Senior Applied AI Engineer, Product Simulation
Senior Applied AI Engineer at NVIDIA to lead the rebuild of a silicon productization toolchain around AI. The role involves building agentic systems to demystify chip feature interactions, integrating AI tools into an agent harness, and leading eval-driven development for applied AI in production.
AgentEngineeringSanta Clara, CA2w ago8
Software Engineer, AI Networking Architect
NVIDIA is seeking an AI Networking Architect to optimize AI workload performance by analyzing AI models, distributed training, and inference workloads, and translating research insights into software, hardware, and networking architecture requirements. The role involves building platforms and simulations to evaluate trade-offs and influence future NVIDIA product roadmaps.
ServeAgent
101–150 of 375← Prev1234…8Next →
Engineering
Tel Aviv, Israel +1
2w ago
8
Senior Software Engineer, Agentic Engineering
Senior Software Engineer to build agentic workflows for code generation, testing, and tuning within NVIDIA's frameworks and compilers. The role involves partnering with internal teams to develop and integrate AI agents into engineering processes, focusing on multi-agent orchestration and autonomous loops.
AgentEngineeringSanta Clara, CA +1 · Remote2w ago8
GPU Performance Engineer - Neural Reconstruction
GPU Performance Engineer focused on optimizing neural reconstruction and Gaussian Splatting workloads, involving PyTorch, CUDA, and GPU profiling to improve training and rendering performance.
ServePost-trainEngineeringCanada · Remote3w ago8
Developer Technology Engineer - AI
NVIDIA is seeking an AI Developer Technology Engineer to study and develop cutting-edge deep learning techniques, analyze and optimize performance on GPU architectures, and work with customers to provide AI solutions using GPUs. The role involves close collaboration with internal NVIDIA teams to influence future architectures and software platforms.
ServeEngineeringShanghai, China +23w ago8
Systems Performance Engineer, Agentic AI Workloads – New College Grad 2026
This role focuses on modeling, simulating, and analyzing the system-level performance of agentic AI workloads in datacenter environments. The engineer will develop simulators, characterize LLM serving traffic, identify performance bottlenecks, and provide architectural recommendations for next-generation AI systems. The role requires strong programming skills in C++ and Python, a solid understanding of queueing theory, traffic modeling, and statistics, as well as fundamentals of deep learning and LLM inference serving.
ServeAgentEngineeringSanta Clara, CA +23w ago8
Software Engineering Manager, Robotics Neural Reconstruction and Real2Sim Applications
NVIDIA is seeking an Engineering Manager to lead a team focused on robotics Neural Reconstruction & Real2Sim Applications, advancing technologies for creating digital twins and workflows at scale for physical AI.
ShipDataEngineeringSanta Clara, CA3w ago8
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Senior Engineer focused on Applied AI and AI Infrastructure for Chip Design and DFX at NVIDIA. The role involves building and managing deployment cycles for ML & Gen AI projects, establishing robust AI infrastructure, and applying AI methods to solve complex problems in Design For Test. Requires expertise in agents, multi-agentic ecosystems, SQL, ETL, data modeling, cloud platforms, and strong programming skills in Python/C++.
AgentServeEngineeringSanta Clara, CA3w ago8
Applied AI Engineer - VLSI Design
NVIDIA is seeking an Applied AI Engineer to develop and deploy AI agents leveraging LLMs to solve complex problems in VLSI design. The role involves designing and building infrastructure for LLM-powered engineering assistants and multi-turn dialogue systems, fine-tuning models, and integrating them with CAD flows.
AgentEngineeringSanta Clara, CA3w ago8
Senior ASIC AI Engineer
Develop AI powered methodologies and Agents to generate micro-architecture, RTL, and physical design starting with specification, using AI agents to process large data and existing codebase to generate skills that can be widely used. Evaluate latest Multi-agent collaboration frameworks and apply them to generate area/power/timing/functionally accurate designs for memory system units in the GPU.
AgentEngineeringSanta Clara, CA3w ago8
Senior System Software Engineer, Robotics
NVIDIA is seeking a Senior System Software Engineer for their Robotics Platform Team, focusing on humanoid robots and embodied intelligence. The role involves integrating robotics software stacks, enabling deployment of foundation models and RL policies, developing validation workflows, and optimizing system metrics. The engineer will work with AI, simulation, and hardware teams to bring up and harden robotic systems.
ShipAgentEngineeringShanghai, China3w ago8
Deep Learning Computer Architect - New College Grad 2026
NVIDIA is seeking a Deep Learning Computer Architect to design hardware accelerator and processor architectures for next-generation platforms, enabling state-of-the-art machine learning and data analytics. The role involves analyzing DL methods, proposing new features for acceleration, and studying their benefits, with a focus on LLM workloads and deep learning kernels.
ServeEngineeringSanta Clara, CA +13w ago8
Manager, Deep Learning Algorithms
Manager to lead engineering activities for productizing Deep Learning models, focusing on implementing and optimizing state-of-the-art algorithms for GPU-accelerated platforms. The role involves leading a team, collaborating with internal partners on roadmap development, and deploying training and inference workloads.
ServeDataEngineeringWarsaw, Poland +1 · Remote4w ago8
Engineering Manager, Inference Benchmarking — AI Perf
Engineering Manager for NVIDIA's AIPerf platform, a standard for assessing LLM serving performance. The role involves leading a team to build and advance the platform, focusing on core infrastructure, accuracy of benchmark results, and advising on upstream engine integrations for various AI workloads (LLM, multimodal, diffusion, computer vision). Requires strong systems engineering, inference infrastructure, and open-source community experience.
ServeEngineeringSanta Clara, CA +5 · Remote4w ago8
AI Computing Development Engineer, TensorRT and TensorRT-LLM
NVIDIA is seeking software engineers to develop and optimize AI inference software (TensorRT/TensorRT-LLM) for GPUs. The role involves performance analysis, tuning, integrating new advancements, and collaborating across teams to shape the future of machine learning inferencing.
ServeEngineeringShanghai, China4w ago8
GPU Performance Engineer - Neural Reconstruction
GPU Performance Engineer focused on optimizing neural reconstruction and Gaussian Splatting workloads. This role involves profiling, identifying bottlenecks, and improving performance in CUDA, PyTorch, and C++ for training and rendering, while ensuring reconstruction quality is maintained. It requires strong programming, GPU optimization, and performance analysis skills, with collaboration across research and engineering teams.
ServeDataEngineeringCA +5 · Remote4w ago8
Perception Engineer - Autonomous Driving
NVIDIA is hiring a Perception Engineer for their Autonomous Driving team in China. The role involves research, design, and implementation of software features for autonomous driving perception, including DNN improvement, evaluation, and deployment. Requires strong C++/PyTorch, ML/DL techniques for Computer Vision, and experience with perception stacks. Familiarity with DNN development, network acceleration, and GPU computing is a plus.
ShipPost-trainEngineeringShanghai, China +15w ago8
Principal Machine Learning Engineer, Accelerated Apache Spark
This role focuses on applying ML/AI to optimize and accelerate Apache Spark workloads on NVIDIA GPUs, involving performance prediction, adaptive systems, and developing AI agents for system issue resolution and optimization. The role requires significant experience in ML/DL solution design, productionization, and large-scale data processing platforms like Spark, with a focus on LLM/GenAI, reinforcement learning, and adaptive ML systems.
AgentServeEngineeringSanta Clara, CA5w ago8
Senior DGX Cloud AI Infrastructure Software Engineer
NVIDIA is seeking a Senior DGX Cloud AI Infrastructure Software Engineer to design, build, and maintain AI infrastructure for large-scale AI training and inferencing. The role involves optimizing efficiency and resiliency of AI workloads, developing scalable AI and Data infrastructure tools, and ensuring high availability of AI systems.
ServeDataEngineeringShanghai, China5w ago8
Senior AI Infrastructure Software Engineer - DGX Cloud
NVIDIA is seeking a Senior AI Infrastructure Software Engineer to design, build, and maintain AI platforms for large-scale AI training, inferencing, fine-tuning, and Agentic AI in production. The role involves developing platform and tools for AI/ML workload efficiency, resiliency, and observability, with a focus on distributed systems and Kubernetes.
ServeEngineeringSanta Clara, CA +3 · Remote6w ago8
Senior Engineer - AI Agents and Systems
Senior Engineer role focused on deploying advanced AI agent frameworks and local runtimes to Windows and NVIDIA GeForce RTX GPUs, ensuring open-source AI agents run locally, safely, and efficiently on consumer PCs, and creating the foundation of the desktop AI operating system.
AgentEngineeringSanta Clara, CA +16w ago8
Senior Engineer - AI Agents and Systems
Senior Engineer role focused on deploying advanced AI agent frameworks and local runtimes to Windows and NVIDIA GeForce RTX GPUs, ensuring open-source AI agents run locally, safely, and efficiently on consumer PCs. The role involves leading development for the foundation of the desktop AI operating system by combining local inference with robust privacy routers and sandboxed execution.
AgentEngineeringSanta Clara, CA +16w ago8
Manager, Test and Tools Development Engineering
Manager for a test and tools development engineering team focused on building autonomous systems and AI-powered quality infrastructure for Omniverse. The role involves leading a team to design agentic test pipelines, multi-agent orchestration for test generation, failure triage, and establishing evaluation frameworks for AI-generated outputs.
AgentEngineeringPune, India6w ago8
Senior AI Engineer, Agents and Developer Workflows
Senior AI Engineer role focused on developing and deploying AI agents and LLM-based solutions to automate software engineering workflows within NVIDIA. The role involves creating tools to improve developer efficiency, accelerate feedback loops, and enhance release reliability, with a focus on predictive modeling for risk identification and leveraging RAG and fine-tuning techniques.
AgentEngineeringBeijing, China +17w ago8
Senior Performance Compiler Engineer - Triton
Senior Performance Compiler Engineer to work on the open-source Triton compiler project, focusing on using compilers to improve AI performance on NVIDIA GPUs for large language models, agents, and other AI applications. The role involves investigating GPU hardware, designing and implementing compiler technology using MLIR to optimize kernel descriptions for efficient GPU code generation, and collaborating with internal teams.
ServeEngineeringRedmond, WA +5 · Remote7w ago8
Senior Systems Engineer, Neural Graphics
Senior Systems Engineer role focused on integrating AI and traditional rendering techniques for real-time visual experiences. The role involves taking innovative techniques like AlpaDreams and driving them into production-ready, real-time pipelines, owning the end-to-end path from prototype to shipping product, and solving complex systems challenges related to latency, memory, and throughput. Requires deep expertise in graphics and AI, with a strong track record of shipping impactful products and experience with systems-level thinking.
ShipAgentEngineeringSanta Clara, CA7w ago8
Senior Data and AI Solutions Engineer
Senior Data and AI Solutions Engineer at NVIDIA to partner with engineering teams, transform data, BI, automation, and agentic AI into measurable engineering productivity gains. Develop and deploy production-grade AI and data solutions, including agentic systems and RAG pipelines, from concept to deployment.
AgentServeEngineeringYokneam, Israel7w ago8
Senior GPU System Architect
Seeking a Senior GPU System Architect to design multi-GPU scale-up and scale-out systems for AI and HPC datacenters. The role involves defining system architectures that integrate GPU compute, memory, and interconnects for optimal AI performance and scalability. Requires deep experience in system-level fabric/networking architecture and hardware-software co-design.
ServeEngineeringSanta Clara, CA7w ago8
Applied AI Engineer - Silicon Co-Design Group
NVIDIA is seeking an Applied AI Engineer to design, develop, and integrate AI/LLM-powered systems into their chip design and automation infrastructure. The role involves architecting and implementing solutions to enhance workflow efficiency, scalability, and intelligence, driving initiatives from concept to deployment. Requires hands-on experience building and deploying ML/AI systems or data-intensive backend services, with a focus on owning AI agents or LLM-powered workflows end-to-end.
AgentEngineeringShanghai, China7w ago8
Senior Research Engineer, Robotics Systems
Senior/Principal Engineer in robotics systems, focusing on foundation models and full-stack technology for humanoid robots. Responsibilities include designing teleoperation software, optimizing control stacks, deploying neural network models on hardware, and collaborating on the MLOps lifecycle. Requires strong robotics and software engineering background, with experience in real-time control and deploying ML models on robotic hardware.
ShipDataEngineeringSanta Clara, CA7w ago8
Senior Perception Engineer - Autonomous Vehicles
Senior Perception Engineer at NVIDIA focused on developing and productizing autonomous driving solutions using deep learning and multi-sensor fusion. The role involves applied research, algorithm development, and ensuring solutions meet production requirements for safety, latency, and robustness.
ShipPost-trainEngineeringSanta Clara, CA7w ago8
Applied AI Engineer - DFT Methodology
NVIDIA is seeking an Applied AI Engineer to explore and architect generative AI solutions, including LLMs, RAGs, and Agentic AI workflows, for Design-for-Test (DFT) and VLSI problems. The role involves deploying predictive ML models for silicon lifecycle management and collaborating with VLSI/DFX teams to integrate AI solutions. Experience in applied ML for chip design and deploying generative AI for engineering use cases is required.
AgentEngineeringBangalore, India7w ago8
Senior Deep Learning Performance Architect
Senior Deep Learning Performance Architect at NVIDIA to design and evaluate hardware architectures for AI/HPC applications, focusing on LLM inference and training performance, and optimizing system bottlenecks.
ServePost-trainEngineeringSanta Clara, CA +17w ago8
Senior Data Center Performance Engineer - Benchmarking and Optimization
Senior Data Center Performance Engineer at NVIDIA focused on benchmarking and optimizing data center platforms for AI training, inference, and HPC workloads. Responsibilities include designing benchmarks, characterizing workloads, identifying bottlenecks, and driving performance improvements through system tuning and architectural recommendations.
ServeEngineeringSanta Clara, CA +1 · Remote7w ago8
NCX Engineer, AI Accelerator
This role focuses on engineering and deploying AI infrastructure and solutions for strategic customers, optimizing large-scale training and inference workloads on NVIDIA's AI platform. It involves MLOps, Kubernetes, GPU scheduling, and performance tuning, with a strong emphasis on customer-facing technical support and collaboration.
ServePost-trainEngineeringSanta Clara, CA +17w ago8
Machine Learning Applications and Compiler Engineer, LPX - New College Grad 2026
NVIDIA is seeking engineers to develop algorithms and optimizations for their LPX inference and compiler stack, working at the intersection of large-scale systems, compilers, and deep learning to optimize neural network workloads on future NVIDIA platforms. The role involves building and maintaining high-performance runtime and compiler components, defining workload mappings, integrating with the SW ecosystem, benchmarking, profiling, and collaborating with hardware teams. It also includes prototyping new compilation techniques and publishing technical work.
ServeEngineeringToronto, ON +1 · Remote7w ago8
Senior Deep Learning Framework Communications Engineer
Senior Deep Learning Framework Communications Engineer at NVIDIA, focusing on integrating and optimizing communication libraries (NCCL, NVSHMEM) within AI frameworks (PyTorch, TRT-LLM, vLLM, JAX) to enhance performance for large-scale AI training and inference. The role involves deep analysis of AI workloads, compiler improvements, and kernel authoring for multi-GPU systems.
ServeEngineeringSanta Clara, CA +4 · Remote7w ago8
Senior Scientific Machine Learning Engineer – Earth-2
Develops and enhances machine learning frameworks (NVIDIA PhysicsNeMo, NVIDIA Earth2Studio) for scientific ML technology in weather, climate, and earth system modeling. Focuses on implementing new deep learning techniques and enhancing Earth-2 technologies.
Post-trainEngineeringSanta Clara, CA +1 · Remote7w ago8
Director, System Software Engineering - Metropolis Accelerated and Inferencing Software
NVIDIA is seeking a Director of System Software Engineering to lead teams responsible for the full lifecycle of Vision AI strategy, from model onboarding to production deployment. The role focuses on transforming foundation models into real-time, GPU-accelerated video intelligence systems, scaling multimodal reasoning, and enabling agentic development workflows. Key responsibilities include architecting and operationalizing inference acceleration, driving implementations of frameworks like TensorRT and VLLM, collaborating with partners on custom models, and ensuring performance benchmarking. The ideal candidate has extensive experience in deep learning, GPU optimization, and leading engineering teams in embedded and enterprise platforms.
ServeAgentEngineeringSanta Clara, CA8w ago8
Director, Isaac for Healthcare Engineering
Director of Engineering for NVIDIA's Isaac for Healthcare initiative, focusing on building a platform for healthcare robotics companies to develop, simulate, train, and deploy physical AI systems. The role involves platform leadership, team building, partner enablement, technical strategy, and cross-functional collaboration, with a strong emphasis on shipping sophisticated software platforms at scale.
ShipDataEngineeringSanta Clara, CA8w ago8
Manager, Solutions Architecture - Global Partner Team
Manager of Solutions Architecture for NVIDIA's Global Partner Team, focusing on leading technical engagements with GSIs and AI consulting firms. The role involves building and scaling Agentic AI services, providing architectural oversight for complex AI workflows, and collaborating with product and engineering teams. Requires deep technical expertise in Generative/Agentic AI, RAG, LLM orchestration, and AI infrastructure.
AgentEngineeringSanta Clara, CA8w ago8
Senior Software Architect - Deep Learning and HPC Communications
Senior Software Architect role at NVIDIA focused on designing and implementing next-generation data center platforms and scalable communication software for AI and HPC workloads. The role involves investigating performance bottlenecks, developing new communication technologies, exploring hardware/software co-design, and building proofs-of-concept to drive innovation in large-scale GPU clusters.
ServeEngineeringSanta Clara, CA +4 · Remote8w ago8
Senior Software Engineer, Deep Learning Inference
Senior Software Engineer focused on optimizing deep learning inference for LLMs and omnimodal architectures on NVIDIA hardware, including GPU kernel tuning, distributed inference, and contributing to open-source libraries.
ServeEngineeringTel Aviv, Israel8w ago8
Senior Hardware Architect, Deep Learning GPU and System
Senior Hardware Architect role focused on designing next-generation GPUs and systems to advance the state of AI, analyzing deep learning workloads, and proposing new features for acceleration. Requires 8+ years of experience in performance, hardware architecture, and deep learning analysis.
ServeEngineeringYokneam, Israel8w ago8
Senior Software Engineer - VLM Microservices for Neural Reconstruction
Senior Software Engineer to design, build, and optimize containerized inference execution for 3D Vision Language Models (VLMs) for neural reconstruction, turning research into production-grade software (NIMs). The role involves developing benchmarks, releasing and maintaining models, contributing to open-source projects like vLLM, and collaborating with research and product teams. Requires experience with AI distributed systems, inference platforms, Python/C++, and software engineering fundamentals.
ServePost-trainEngineeringSanta Clara, CA +18w ago8
Senior Applied Machine Learning Engineer - VLSI Design
NVIDIA is seeking a Senior Applied Machine Learning Engineer to build AI-driven software systems for circuit design, combining automation algorithms, DL models, and agentic workflows. The role involves working on pre-silicon and post-silicon hardware design data, circuit optimization, and AI systems for EDA/design automation, translating requirements into AI/ML and agentic system problems, and testing/releasing models and AI systems.
AgentDataEngineeringSanta Clara, CA8w ago8
Applied Machine Learning Engineer, Circuit Design - New College Grad 2026
NVIDIA is seeking an Applied Machine Learning Engineer for their Circuit Design team, focusing on building AI-driven software systems that combine automation algorithms, DL models, and agentic workflows to accelerate end-to-end circuit design. The role involves working with hardware design data, circuit optimization, and developing AI/ML solutions for EDA, with a focus on agent-driven design exploration and optimization.
AgentEngineeringSanta Clara, CA +1 · Remote8w ago8
Principal AI and ML Infra Software Engineer, GPU Clusters
This role focuses on enhancing the efficiency of AI and ML research on GPU clusters by collaborating with researchers to identify and address infrastructure deficiencies. The engineer will optimize performance, monitor resource utilization, and contribute to the AI/ML infrastructure ecosystem, keeping up-to-date with the latest AI/ML technologies.
ServeEngineeringSanta Clara, CA +18w ago8
Senior Deep Learning Software Engineer - Autonomous Vehicles
Senior Deep Learning Software Engineer focused on developing and productizing deep learning solutions for autonomous vehicles. The role involves training, fine-tuning, optimizing perception DNNs, applying quantization, improving DNN architectures, and enhancing inference speed and power consumption. It requires strong programming skills, experience with deep learning frameworks, computer vision tasks, and familiarity with CNNs and Transformer architectures. Experience with low precision inference, quantization, and NVIDIA software libraries is a plus.
ServePost-trainEngineeringSanta Clara, CA +3 · RemoteApr 248
Compiler Engineer - AI Inference
NVIDIA is seeking an AI Compiler Engineer to optimize kernel generation and computational graph optimizations for AI inference and training workloads on next-generation GPUs. The role involves hands-on development, collaboration on hardware/software co-design, and scaling AI deployments in datacenters.
ServePost-trainEngineeringSanta Clara, CAApr 248