AI Hire Signal
JobsCompaniesTrendsInsightsWeekly
JobsStrategy timeline
AI Hire Signal

Tracking AI hiring across 200+ US tech companies. Stage, salary, and stack signals on every role — refreshed weekly.

Contact

Browse

JobsCompaniesTrendsInsightsWeekly

Resources

AboutSitemapRobots

Legal

PrivacyTerms
© 2026 AI Hire Signal·Not affiliated with companies shown

Currently tracking 440 active AI roles, down 53% versus the prior 4 weeks. Primary focus: Serve · Engineering. Salary range $100k–$575k (avg $262k).

Hiring
440 / 623
Momentum (4w)
↓-386 -53%
340 opens last 4w · 726 prior 4w
Salary range · avg $262k
$100k–$575k
USD · disclosed roles only
Tracked since
May '25
last role 4w ago
Hiring velocityscroll left for older weeks
1 new role
Dec 30
1 new role
Mar 10
1 new role
24
1 new role
Apr 28
4 new roles
May 12
5 new roles
19
3 new roles
26
3 new roles
Jun 2
2 new roles
9
1 new role
16
2 new roles
23
3 new roles
30
4 new roles
Jul 7
1 new role
14
2 new roles
28
4 new roles
Aug 11
6 new roles
18
2 new roles
25
3 new roles
Sep 1
8 new roles
15
3 new roles
22
6 new roles
29
2 new roles
Oct 6
2 new roles
13
3 new roles
20
6 new roles
27
9 new roles
Nov 3
8 new roles
10
8 new roles
17
4 new roles
24
11 new roles
Dec 1
9 new roles
8
14 new roles
15
10 new roles
22
8 new roles
29
107 new roles
Jan 5
22 new roles
12
45 new roles
19
32 new roles
26
59 new roles
Feb 2
64 new roles
9
63 new roles
16
83 new roles
23
83 new roles
Mar 2
88 new roles
9
97 new roles
16
72 new roles
23
215 new roles
30
158 new roles
Apr 6
250 new roles
13
199 new roles
20
332 new roles
27
304 new roles
May 4
189 new roles
11
131 new roles
18
102 new roles
25
129 new roles
Jun 1
122 new roles
8
49 new roles
15
40 new roles
22
NVIDIA

NVIDIA

Semiconductors

HQ
California, US
Founded
1993
Size
35,000+
Website
nvidia.com

NVIDIA currently has 496 active AI-related job listings. The majority of these roles, 52%, are focused on serving infrastructure, with agents representing another significant segment at 23%. Engineering is the dominant function, with 441 positions. The United States leads hiring geographies with 287 roles, followed by China with 64. Frequent tech tags include model_serving, inference_infra, and agent_orchestration, suggesting a focus on deployment and management of AI models. Over the last 30 days, NVIDIA posted 214 new AI roles, a 27% decrease compared to the previous 30-day period.

Auto-generated from active job postings · last refreshed 2026-05-24

Frequently asked questions

  • What AI roles is NVIDIA hiring for?

    NVIDIA currently has 487 active AI-related roles in our index. The most common open titles are: Deep Learning Performance Architect (4), Senior Deep Learning Performance Architect (4), AI Research Scientist (3), Developer Technology Engineer - AI (3), Manager, Deep Learning Algorithms (3). Most positions are in Engineering and Research.

  • What stage of AI development does NVIDIA focus on?

    NVIDIA's active AI hiring is concentrated in: serving infrastructure (54%), agents (21%), application (8%). These categories follow a seven-stage AI lifecycle: data, pre-training, post-training, serving infrastructure, agents, evaluation, and application.

  • Where is NVIDIA hiring AI talent?

    NVIDIA is hiring AI talent in: United States (286 roles), China (59 roles), Israel (50 roles), Germany (21 roles).

  • What technologies does NVIDIA's AI team work with?

    Job postings at NVIDIA most frequently reference: model serving, inference infra, agent orchestration, llm observability, multimodal.

  • How many AI roles has NVIDIA posted recently?

    In the past 30 days, NVIDIA has posted 110 new AI-related roles. That is a -50% change versus the prior 30 days (218 → 110).

Jobs (723)

434 AI · 1824 total active
Show
Active onlyAI only (≥ 7)
Stage
AllData · 28Pretrain · 30Post-train · 51Serve · 356Agent · 192Eval Gate · 11Ship · 55
Function
AllEngineering · 627Research · 82Product · 14
Country
AllUnited States · 439China · 93Israel · 54Germany · 36Switzerland · 31India · 26United Kingdom · 24Poland · 17Vietnam · 13Canada · 12Singapore · 11France · 10Netherlands · 9Italy · 8Taiwan · 6Hong Kong · 4Japan · 4Spain · 3Australia · 2Czech Republic · 2Finland · 2Hungary · 2South Korea · 2Armenia · 1Brazil · 1Mexico · 1Romania · 1Saudi Arabia · 1Sweden · 1United Arab Emirates · 1
Sort
AI scoreRecentTitle
TitleStageFunctionLocationFirst seenAI score
Research Scientist, Deep Learning and Computer Vision - New College Graduate
Research Scientist role at NVIDIA focusing on deep learning and computer vision, with an emphasis on novel methods, generative and multimodal AI, and explainable AI. The role involves research, design, implementation, publication, and technology transfer to product groups. Requires a Ph.D. and a strong publication record in top conferences.
PretrainResearchTaipei, Taiwan2w ago10
Senior Scientist, Synthetic Data Generation
Senior Scientist focused on synthetic data generation for training frontier LLMs, contributing to open-source libraries and advancing multimodal data generation.
DataPost-train
1–50 of 723← Prev12…15Next →
Research
Santa Clara, CA +5 · Remote
2w ago
10
AI Research Scientist
AI Research Scientist at NVIDIA focusing on GPU-accelerated generative AI models for language, vision, and robotics, with a strong emphasis on publishing novel research and efficient model design.
PretrainResearchSingapore, Singapore · Remote4w ago10
AI Research Scientist
AI Research Scientist at NVIDIA focusing on developing and optimizing GPU-accelerated generative AI models for language, vision, and robotics, with a strong emphasis on publishing research and transferring technology to products.
PretrainResearchSingapore, Singapore · Remote7w ago10
Research Scientist
NVIDIA Research Singapore is seeking AI researchers to develop and optimize GPU-accelerated efficient AI computing for language models, visual generation, and robotics, focusing on pushing generative AI boundaries and publishing research.
PretrainResearchSingapore, Singapore · Remote7w ago10
Applied Deep Learning PhD Research Intern, Reinforcement Learning for LLMs - Fall 2026
PhD research intern focused on advancing LLMs using reinforcement learning. Will design, implement, and evaluate new RL-based methods for improving LLM behavior, reasoning, alignment, and task performance. Involves hands-on experimentation and large-scale GPU cluster work.
Post-trainAgentResearchSanta Clara, CA7w ago10
Senior Research Scientist, Multimodal Foundation Models and Robotics
Research Scientist role focused on building multimodal foundation models and systems for humanoid robots and embodied agents, involving algorithm design, large-scale training/inference, and deployment on physical hardware and simulations.
Post-trainAgentResearchSanta Clara, CA8w ago10
Senior Deep Learning Researcher, Diffusion
Senior Deep Learning Researcher at NVIDIA focusing on diffusion-based technologies and multi-modal learning. The role involves inventing and building new techniques, combining diffusion models with LLMs, and publishing research findings. It requires a PhD, research experience, and a strong publication record in leading AI conferences and journals.
PretrainPost-trainResearchTel Aviv, Israel8w ago10
Senior Research Scientist, Fundamental Generative AI
Senior Research Scientist focused on fundamental generative AI research, particularly for biomolecular design and scientific applications. The role involves designing and implementing novel, large-scale generative models, publishing research, and transferring technology to product groups.
PretrainResearchSanta Clara, CAApr 410
AI Research Scientist
AI Research Scientist at NVIDIA focusing on developing and optimizing GPU-accelerated generative AI models for language, vision, and robotics, with a strong emphasis on publishing research and transferring technology to products.
PretrainResearchSingapore, Singapore · RemoteApr 410
Research Scientist, Fundamental Generative AI - New College Grad 2026
Research Scientist role focused on fundamental generative AI research, pushing boundaries in image, video, 3D, or scientific applications. Requires a strong mathematical foundation, novel model development, and publication in top ML venues. The role involves designing and implementing large-scale generative AI methods, publishing original research, and collaborating with research and product teams.
PretrainResearchSanta Clara, CAApr 410
Research Scientist, Fundamental Generative AI - New College Grad 2026
Research Scientist role focused on fundamental generative AI research, particularly for scientific applications like biomolecular design. The role involves developing novel, large-scale generative models, publishing research, and potentially transferring technology to product groups. Requires a strong theoretical and practical understanding of generative AI and deep learning.
PretrainResearchSanta Clara, CAApr 410
Research Scientist, Generalist Embodied Agent Research - PhD New College Grad 2026
Research Scientist role focused on building humanoid robot foundation models and general-purpose embodied agents. Involves designing and implementing novel AI algorithms, developing large-scale training and inference methods, and deploying models in simulation and on hardware. Requires strong experience in LLMs, multimodal foundation models, reinforcement learning, agent learning, and applied robotics, with a PhD and publication record.
AgentDataResearchSanta Clara, CA +1 · RemoteJan 910
Senior Research Scientist, Multimodal Foundation Models and Robotics
Research Scientist role focused on developing multimodal foundation models and systems for general-purpose humanoid robots and embodied agents, involving algorithm design, large-scale training/inference, and deployment on physical hardware and simulations.
Post-trainAgentResearchSanta Clara, CAJan 910
Deep Learning Senior Engineer, End-To-End Autonomous Driving
NVIDIA is seeking a Deep Learning Senior Engineer to design, implement, and deploy end-to-end autonomous driving systems. The role focuses on AI 2.0, leveraging LLMs, VLMs, and VLAs for reasoning and planning in autonomous vehicles and robotics. Responsibilities include training large-scale models, building and fine-tuning LLM/VLM/VLA systems, exploring data generation strategies, and deploying models in production environments, integrating them with vehicle firmware.
Post-trainAgentEngineeringShanghai, China +11w ago9
Senior Research Scientist, Nemotron Post-training
Research Scientist/Engineer at NVIDIA focused on building Nemotron models, specifically working on post-training pipelines, synthetic data, agentic RL, data/training infrastructure, and large-scale model post-training. The role involves advancing open-source foundation models, developing training data, benchmarks, LLMs, and software, and solving end-to-end foundation model post-training challenges. Requires a Master's/PhD and 5+ years of experience in model post-training, RL, and agentic systems, with experience in data curation, model training, and inference/deployment environments.
Post-trainAgentResearchSanta Clara, CA +1 · Remote1w ago9
High-Performance LLM Training Engineer - New College Grad 2026
NVIDIA is seeking an experienced engineer to optimize LLM training workloads on high-performance computing systems, focusing on software stack optimization for thousands of GPUs and influencing future hardware roadmaps. The role involves performance analysis, profiling, and implementation across the deep learning platform, from drivers to frameworks, and contributing to MLPerf benchmarks.
DataEngineeringSanta Clara, CA1w ago9
Research Scientist, Efficient Deep Learning - New College Grad 2026
Research Scientist role focused on efficient deep learning methods, including post-training optimization, efficient architecture design, and resource-efficient training/finetuning. Requires a Ph.D. or equivalent research experience, strong Python/PyTorch skills, and experience with large-scale model training and large vision-language models. The role involves research, implementation, publication, and technology transfer.
Post-trainServeResearchSanta Clara, CA2w ago9
Deep Learning Performance Architect
NVIDIA is seeking a Deep Learning Performance Architect to analyze, model, and optimize deep learning system performance, particularly for LLM workloads, on state-of-the-art hardware architectures. This role influences future hardware and software design by collaborating with various internal teams.
ServeEngineeringShanghai, China +12w ago9
Deep Learning Performance Architect
NVIDIA is seeking a Deep Learning Performance Architect to optimize deep learning hardware and software architectures for edge devices, workstations, and data center GPUs. The role involves benchmarking, performance modeling, bottleneck identification, and exploring new hardware/software capabilities, with a focus on LLMs and generative AI. Experience with AI agents for engineering workflows is also mentioned.
ServePost-trainEngineeringShanghai, China2w ago9
Senior Scientist, Synthetic Data and Privacy
Senior Scientist role focused on building LLM-based methods for synthetic data generation and privacy-preserving AI, contributing to open-source libraries within the NVIDIA NeMo ecosystem. The role involves applied research, software engineering, and optimizing LLMs for inference, with a strong emphasis on publishing original research.
DataServeResearchSanta Clara, CA +5 · Remote2w ago9
Senior Quantum Applied Research Scientist, Calibration and Decoding
Research Scientist at NVIDIA focusing on developing AI models for quantum system calibration and decoding. This role involves building physics-informed synthetic data generation pipelines, developing surrogate models of quantum hardware, and architecting real-time AI systems. The work also includes applying reinforcement learning and online learning methods for optimization, with a strong emphasis on GPU acceleration and collaboration across Product, Engineering, and Applied Research teams to advance fault-tolerant quantum computing.
Post-trainDataResearchRedmond, WA +2 · Remote2w ago9
Senior Research Manager, World Model Evaluation
Lead a research team focused on world-model evaluation and benchmarking for NVIDIA's Physical AI portfolio, defining the scientific roadmap for closed-system and open-system evaluations, developing benchmarks for various physical AI capabilities, and driving evaluation-to-model-improvement loops. The role requires publishing high-quality papers and establishing rigorous standards.
Eval GatePost-trainResearchSanta Clara, CA3w ago9
Senior Systems Software Engineer, AI Stack and Performance - DGX Station
Senior Systems Software Engineer focused on optimizing AI stack performance and readiness on NVIDIA's DGX Station, a workstation-class AI computer. The role involves profiling, identifying bottlenecks, and driving optimizations across the full stack from GPU kernels to applications, ensuring AI workloads like LLM inference and agents run efficiently in multi-GPU, multi-user configurations. Collaboration with framework, compiler, and GPU architecture teams is critical.
ServeShipEngineeringSanta Clara, CA +1 · Remote3w ago9
Applied AI Researcher - World Reconstruction and Generation
NVIDIA is seeking an Applied AI Researcher to work on NuRec-related research in world reconstruction and generation, developing and adapting Deep Learning-based methods for tasks like novel view synthesis, generative modeling, and neural rendering. The role involves prototyping with Python/PyTorch, building evaluation and agentic AI-assisted research workflows, and turning research into usable technology. Requires a PhD or MS with significant experience in ML/DL, Computer Graphics, Computer Vision, or 3D reconstruction, with strong Python/PyTorch skills.
Post-trainAgentResearchGermany +3 · Remote3w ago9
Senior Machine Learning Engineer, Perception - Autonomous Driving
NVIDIA is seeking a Senior Machine Learning Engineer for their Autonomous Driving Perception team. The role involves designing and developing end-to-end deep learning solutions for perception modules, focusing on road layout detection, lane structures, and other critical driving components. The engineer will also drive data-driven development, leverage simulation and augmentation, and productize solutions meeting safety and latency requirements. Experience with deep learning frameworks, Python/C++, and perception for autonomous driving or robotics is essential.
ShipDataEngineeringSanta Clara, CA +2 · Remote3w ago9
Senior Software Engineer, DGX Cloud AI Infrastructure
Senior Software Engineer to lead the bring-up, triage, benchmarking, analysis, and optimization of distributed training and inference workloads across NVIDIA GPU platforms at scale. This role involves setting technical direction for communication libraries, model frameworks, and inference/training stacks, leading performance and reliability investigations, defining benchmarking and qualification processes, and building resilience capabilities for large clusters.
ServePost-trainEngineeringSanta Clara, CA +4 · Remote3w ago9
Software Engineer, DGX Cloud AI Infrastructure
Software Engineer role focused on AI infrastructure, specifically distributed training and inference workloads on NVIDIA GPU platforms. Responsibilities include bring-up, triage, benchmarking, analysis, and optimization of these workloads at scale. Requires experience with multi-GPU/multi-node systems, debugging distributed environments, and strong Python/C++ skills.
ServePost-trainEngineeringSanta Clara, CA +4 · Remote3w ago9
Senior Deep Learning Performance Architect
NVIDIA is seeking a Senior Deep Learning Performance Architect to analyze and develop next-generation architectures for AI and HPC applications. The role involves developing innovative architectures, analyzing performance/cost/power trade-offs using models and simulators, understanding hardware/software interplay, and evaluating PPA for architectural decisions. Collaboration with software, product, and research teams is key. Requires MS/PhD, 6+ years experience, strong background in GPU/Deep Learning ASIC architecture for distributed training/inference, performance modeling, and ML/DL fundamentals, particularly transformer architectures. Proficiency in Python, C, C++ is essential.
ServeEngineeringSanta Clara, CA +13w ago9
Research Scientist, Electronic Design Automation - New College Grad 2026
Research Scientist role focused on applying AI/ML techniques, including supervised, unsupervised, reinforcement learning, and agentic AI, to Electronic Design Automation (EDA) and VLSI design. The role involves defining and conducting original research, innovating in EDA software and algorithms, and applying deep learning to improve chip design tools, with a strong emphasis on publication and collaboration.
Post-trainAgentResearchSanta Clara, CA3w ago9
AI Inference Performance Engineer - New College Grad 2026
NVIDIA is seeking an AI Inference Performance Engineer to optimize and benchmark GenAI inference on their accelerators, working with frameworks like TensorRT-LLM, SGLang, and vLLM. The role involves driving industry benchmark results, defining cutting-edge workloads, architecting distributed inference, establishing performance methodology, and influencing the ecosystem through open-source contributions and cross-functional partnerships. Requires strong programming skills, DL framework expertise, and a deep understanding of LLM inference mechanics.
ServeEngineeringSanta Clara, CA3w ago9
Senior Software Engineer, Generative AI Research
NVIDIA is seeking a Senior Software Engineer for Generative AI Research to build and operate scalable infrastructure for training their world foundation model for physical AI, Cosmos. This role involves designing and developing high-throughput systems for data processing, retrieval, and workflow orchestration, improving system reliability and performance, and contributing to long-term infrastructure strategy for training, data management, and large-scale compute efficiency. The role requires a strong engineering background in distributed systems, ML infrastructure, or large-scale compute/data platforms, proficiency in Python and C++/Go/Rust, and experience with orchestration systems and data pipelines. Experience with large-scale model training infrastructure, distributed compute, synthetic data, or multimodal datasets is a plus.
DataPretrainEngineeringSanta Clara, CA3w ago9
Senior Software Manager, Agentic AI
Senior Software Manager to lead a team building agentic AI solutions for chip design workflows, involving coding agents, custom skills, and integration with enterprise systems. The role requires technical leadership in designing, developing, and deploying AI applications using LLMs and agentic systems, including model customization (fine-tuning, RL, instruction tuning) and overseeing retrieval/generation algorithms for enterprise data. Collaboration with cross-functional teams and ensuring high technical standards for evaluation, guardrails, and monitoring are key.
AgentPost-trainEngineeringSanta Clara, CA3w ago9
Deep Learning Performance Software Engineer
Develops GPU-accelerated deep learning software, including compilers, DSLs, and optimized kernels, for current and next-generation chips, focusing on performance analysis of AI workloads and integration with AI frameworks.
ServeEngineeringShanghai, China3w ago9
Senior Research Scientist
NVIDIA is seeking a Senior Research Scientist to join their applied research team focused on building next-generation Conversational AI systems. The role involves developing new Deep Learning models for speech recognition, speech synthesis, neural machine translation, and natural language processing, designing large-scale training algorithms, and open-sourcing models via the NeMo framework. The position requires a PhD, at least 5 years of research experience in speech recognition or NLP, a strong understanding of Deep Learning in these areas, proficiency in Python and PyTorch, and a strong publication record. Collaboration with academic and product teams, as well as mentoring interns, are also key aspects of the role.
Post-trainPretrainResearchGermany · Remote3w ago9
Senior Data Scientist - Security and Networking Research
Senior Data Scientist role focused on AI cybersecurity, developing agentic AI systems, optimizing models, and leveraging data pipelines for NVIDIA's networking and data center security products.
AgentPost-trainResearchTel Aviv, Israel +34w ago9
AI Computing Architect
NVIDIA is seeking an AI Computing Architect to develop innovative architectures for deep learning performance and efficiency, analyze trade-offs using models and simulators, and prototype algorithms. The role requires strong programming skills, computer architecture background, and a foundation in machine learning.
ServePost-trainEngineeringShanghai, China +14w ago9
Senior Software Engineer - Agentic AI
Senior Software Engineer role focused on leading Agentic AI solutions, including sophisticated AI agents and fine-tuning, integrating them with enterprise production systems. The role involves designing, developing, and deploying AI applications using LLMs, Agentic frameworks, and optimizing retrieval/generation algorithms for enterprise data (text, code, images) to build advanced AI applications for engineering assistants and multi-turn, multi-modal dialogue systems, ultimately solving complex problems in chip design.
AgentPost-trainEngineeringSanta Clara, CA4w ago9
AI Workload and Networking Research Architect
Research Architect role focused on optimizing AI workloads and networking infrastructure for NVIDIA's AI computing platforms, involving modeling, analysis, and influencing future product roadmaps.
ServePost-trainResearchYokneam, Israel +14w ago9
Senior Systems Software Engineer, Machine Learning
Senior Systems Software Engineer focused on Machine Learning, specifically generative AI, LLMs/VLMs, computer vision, and agentic systems. The role involves converting research into production products, building and shipping ML workflows/pipelines, and leveraging AI in data generation. Key responsibilities include defining evaluation criteria and running offline evals. Experience with multi-agent pipelines, VLMs in production, and shipping AI-powered features to users is highly valued.
AgentDataEngineeringSanta Clara, CA +14w ago9
Senior AI Safety Red Teamer
NVIDIA is seeking a Senior AI Safety Red Teamer to improve the safety and security posture of their AI models, systems, and infrastructure. The role involves hands-on safety and security research, developing tools to expose weaknesses, defining safety standards, and partnering with cross-functional teams. Requires 5+ years of experience in AI safety/security and offensive cybersecurity, with knowledge of AI vulnerabilities, LLMs, MLLMs, Generative AI, Agents, and RAG workflows, and strong Python programming skills.
Eval GateAgentEngineeringTel Aviv, Israel +34w ago9
Senior High Performance AI Engineer
Senior High Performance AI Engineer to build multi-agent systems for the CUDA ecosystem, focusing on agentic runtimes, compiler-integrated orchestration, and GPU acceleration for agent workloads like planning, tool-use, and code generation. Collaborates across the AI stack from hardware to model/agent teams.
AgentServeEngineeringSanta Clara, CA +5 · Remote5w ago9
Senior Quantum AI Research Scientist, Applied Research
NVIDIA is seeking a Senior Quantum AI Research Scientist to architect and build AI solutions for fault-tolerant quantum computing, focusing on quantum error correction, decoding, calibration, and beyond. The role involves researching and developing open AI models, datasets, and benchmarks, fine-tuning models for specific quantum systems, and collaborating with cross-functional teams to integrate AI into quantum supercomputers.
Post-trainDataResearchRedmond, WA +15w ago9
Senior Applied Deep Learning Scientist - Large Vision Language Models
NVIDIA is seeking a Senior Applied Deep Learning Scientist to work on multimodal language models, specifically the Nemotron Omni family. The role involves pushing the boundaries of these models for downstream applications, preparing large-scale multimodal datasets, and collaborating globally to turn research into impactful products. The position spans the full pipeline from pre-training to post-training, with a focus on open models, weights, and data for real-world applications.
Post-trainDataResearchZurich, Switzerland +3 · Remote5w ago9
Senior LLM Agents Architect
Senior LLM Agents Architect role focused on designing and building agentic AI systems to optimize GPU compute kernels, analyze architectural simulations, and drive improvements in hardware design and developer efficiency. The role involves hands-on CUDA programming, collaboration with hardware architects, and building automated agentic workflows for performance forensics and architectural studies.
AgentEngineeringYokneam, Israel +15w ago9
Senior Research Scientist
Research Scientist role focused on core AI research in machine learning methods for efficient, reliable, and adaptive AI systems. The role involves developing new approaches for efficient model training, reasoning models, retrieval-based AI systems, and agentic learning, with a strong emphasis on experimental validation and publication, and potential translation into practical NVIDIA AI systems.
PretrainAgentResearchYokneam, Israel5w ago9
Senior Performance Architect, Nemotron
NVIDIA is seeking a Senior Performance Architect for Nemotron to focus on deep model-system-hardware co-design. The role involves developing high-fidelity performance models to evaluate architectural choices, predict deployment efficiency, and ensure Pareto-optimal trade-offs for future Nemotron models. This position will guide future software and hardware roadmaps by modeling end-to-end performance impact of GenAI workflows and collaborating with research, framework, compiler, and hardware teams.
ServeEngineeringSanta Clara, CA +25w ago9
Software Engineer, AI and DL Kernel Libraries - New College Grad 2026
Software Engineer role focused on developing AI systems software for efficient inference, including libraries, code generators, and GPU kernels for NVIDIA's hardware. The role involves designing abstractions, optimizing kernels, building LLM serving runtimes, and contributing to open-source projects like FlashInfer and vLLM.
ServeEngineeringSanta Clara, CA +1 · Remote5w ago9
Senior Research Scientist, Post-Training LLM and DLM
Senior Research Scientist focused on post-training algorithms for LLMs and DLMs, system optimization for training and serving, and developing evaluation frameworks. The role involves translating research ideas into production-ready implementations and contributing to open-source communities.
Post-trainServeResearchSanta Clara, CA +1 · Remote6w ago9
Senior Machine Learning Engineer - Physical AI and Synthetic Data Generation
NVIDIA is seeking a Senior Machine Learning Engineer to join their Physical AI team. The role focuses on architecting and developing generative pipelines for high-fidelity synthetic data using multimodal and diffusion models. Responsibilities include building and fine-tuning large-scale models, applying user controls for data synthesis, establishing quality assurance pipelines, and leading the generation of massive training datasets. The role requires deep technical knowledge in image/video synthesis, strong programming skills, and experience in assessing synthetic data impact on model performance.
DataPost-trainEngineeringSanta Clara, CA6w ago9