Research Labs
All Research Labs
Spatial Intelligence
Applied Research
Autonomous Vehicles
Deep Imagination
Publications
AI Playground
New and Featured
AI Art Gallery
NGC Demos
Research Areas
AI & Machine Learning
3D Deep Learning
Computer Vision
Robotics
All Areas
Careers
Academic Collaborations
Government Collaborations
Graduate Fellowship
Internships
Research Openings
Research Scientists
Meet the Team
Licensing
Skip to main content
Artificial Intelligence Computing Leadership from NVIDIA
Login
Research Labs
All Research Labs
Spatial Intelligence
Applied Research
Autonomous Vehicles
Deep Imagination
Publications
AI Playground
New and Featured
AI Art Gallery
NGC Demos
Research Areas
AI & Machine Learning
3D Deep Learning
Computer Vision
Robotics
All Areas
Careers
Academic Collaborations
Government Collaborations
Graduate Fellowship
Internships
Research Openings
Research Scientists
Meet the Team
Licensing
Search
Search
Enter the terms you wish to search for.
Publications
Our publications provide insight into some of our leading-edge research.
Filters
Search
Apply
Filters
Filters
Publication Year
2026
(1)
2025
(10)
2024
(1)
2023
(5)
2022
(5)
2021
(9)
2020
(6)
2019
(15)
2018
(14)
2017
(15)
2016
(9)
2015
(4)
2014
(1)
2012
(4)
2011
(4)
2010
(3)
2009
(2)
2008
(1)
2005
(1)
Facet Publication Year
Research Areas
High Performance Computing
(30)
Algorithms and Numerical Methods
(10)
Artificial Intelligence and Machine Learning
(8)
Computer Architecture
(8)
Computer Graphics
(7)
Programming Languages, Systems and Tools
(5)
Real-Time Rendering
(5)
Resilience and Safety
(3)
VR, AR and Display Technology
(2)
Human Computer Interaction
(1)
Networking
(1)
Events
No Results Available
30 results found
High Performance Computing
Clear all
2019
2017
High Performance Computing
2019
Near-Memory Data Transformation for Efficient Sparse Matrix Multi-Vector Multiplication
Daichi Fujiki,
Niladrish Chatterjee
,
Donghyuk Lee
,
Mike O'Connor
Highly-scalable, Physics-informed GANs for Learning Solutions of Stochastic PDEs
Liu Yang, Sean Treichler, Thorsten Kurth, Keno Fischer, David Barajas-Solano, Josh Romero, Valentin Churavy, Alexandre Tartakovsky, Michael Houston, Prabhat, George Karniadakis
Exascale Deep Learning for Scientific Inverse Problems
Nouamane Laanait, Joshua Romero, Junqi Yin, M. Todd Young, Sean Treichler, Vitalii Starchenko, Albina Borisevich, Alex Sergeev, Michael Matheson
Task Bench: A Parameterized Benchmark for Evaluating Parallel Runtime Performance
Elliott Slaughter, Wei Wu, Yuankun Fu, Legend Brandenburg, Nicolai Garcia, Wilhem Kautz, Emily Marx, Kaleb S. Morris, Qinglei Cao, George Bosilca, Seema Mirchandaney, Wonchan Lee, Sean Treichler, Patrick McCormick, Alex Aiken
Optimizing Multi-GPU Parallelization Strategies for Deep Learning Training
Saptadeep Pal, Eiman Ebrahimi, Arslan Zulfiqar,
Yaosheng Fu
, Victor Zhang, Szymon Migacz,
David Nellans
, Puneet Gupta
GPU-Accelerated Atari Emulation for Reinforcement Learning
Steven Dalton
,
Iuri Frosio
,
Michael Garland
GPU Snapshot: Checkpoint Offloading for GPU-Dense Systems
Kyushick Lee,
Michael B. Sullivan
,
Siva Hari
, Timothy Tsai,
Steve Keckler
, Mattan Erez
On the Trend of Resilience for GPU-Dense Systems
Kyushick Lee,
Michael B. Sullivan
,
Siva Hari
, Timothy Tsai,
Steve Keckler
, Mattan Erez
Best of SELSE (Workshop on Silicon Errors in Logic - System Effects)
NVGaze: An Anatomically-Informed Dataset for Low-Latency, Near-Eye Gaze Estimation
Joohwan Kim
,
Michael Stengel
, Alexander Majercik,
Shalini De Mello
, David Dunn,
Samuli Laine
, Morgan McGuire,
David Luebke
Buddy Compression: Enabling Larger Memory for Deep Learning and HPC Workloads on GPUs
Esha Choukse,
Michael B. Sullivan
,
Mike O'Connor
, Mattan Erez, Jeff Pool,
David Nellans
, Stephen W. Keckler
DeLTA: GPU Performance Model for Deep Learning Applications with In-depth Memory System Traffic Analysis
Sangkug Lym,
Donghyuk Lee
,
Mike O'Connor
,
Niladrish Chatterjee
, Mattan Erez
A Fast and Robust Method for Avoiding Self-Intersection
Carsten Wächter,
Nikolaus Binder
Massively Parallel Path Space Filtering
Nikolaus Binder
, Sascha Fricke,
Alex Keller
Metaoptimization on a Distributed System for Deep Reinforcement Learning
Greg Heinrich,
Iuri Frosio
Massively Parallel Construction of Radix Tree Forests for the Efficient Sampling of Discrete Probability Distributions
Nikolaus Binder
,
Alex Keller
2017
Integrating External Resources with a Task-Based Programming Model
Zhihao Jia, Sean Treichler, Galen Shipman,
Michael Bauer
, Noah Watkins, Carlos Maltzahn, Patrick McCormick, Alex Aiken
AdaBatch: Adaptive Batch Sizes for Training Deep Neural Networks
Aditya Devarakonda, Maxim Naumov,
Michael Garland
Near-eye Light Field Holographic Rendering with Spherical Waves for Wide Field of View Interactive 3D Computer Graphics
Liang Shi, Fu-Chung Huang,
Ward Lopes
, Wojciech Matusik,
David Luebke
A Novel Shard-Based Approach for Asynchronous Many-Task Models for In Situ Analysis
Philippe P. Pébaÿ, Giulio Borghesi, Hemanth Kolla, Janine C. Bennett, Sean Treichler
Control Replication: Compiling Implicit Parallelism to Efficient SPMD with Logical Regions
Elliott Slaughter, Wonchan Lee, Sean Treichler, Wen Zhang,
Michael Bauer
, Galen Shipman, Patrick McCormick, Alex Aiken
Low Communication FMM-Accelerated FFT on GPUs
Cris Cecka
Parallel Jaccard and Related Graph Clustering Techniques
Alexandre Fender, Nahid Emad, Serge Petiton, Joe Eaton, Maxim Naumov
Fine-Grained DRAM: Energy-Efficient DRAM for Extreme Bandwidth Systems
Mike O'Connor
,
Niladrish Chatterjee
,
Donghyuk Lee
,
John Wilson
, Aditya Agrawal,
Steve Keckler
,
William Dally
Feedforward and Recurrent Neural Networks Backward Propagation and Hessian in Matrix Form
Maxim Naumov
Exploiting Budan-Fourier and Vincent’s Theorems for Ray Tracing 3D Bézier Curves
Alexander Reshetov
Parallel Modularity Clustering
Alexandre Fender, Nahid Emad, Serge Petiton, Maxim Naumov
Relaxations for High-Performance Message Passing on Massively Parallel SIMT Processors
Benjamin Klenk, Holger Fröning,
Hans Eberle
,
Larry Dennison
Best Paper Award
The Iray Light Transport Simulation and Rendering System
Alex Keller
, Carsten Wächter, Matthias Raab, Daniel Seibert, Dietger van Antwerpen, Johann Korndörfer, Lutz Kettner
SASSIFI: An Architecture-level Fault Injection Tool for GPU Application Resilience Evaluation
Siva Hari
, Timothy Tsai,
Mark Stephenson
,
Steve Keckler
,
Joel Emer
Parallel Depth-First Search for Directed Acyclic Graphs
Maxim Naumov, Alysson Vrielink,
Michael Garland