Research Labs
All Research Labs
3D Deep Learning
Applied Research
Autonomous Vehicles
Deep Imagination
Publications
AI Playground
New and Featured
AI Art Gallery
NGC Demos
Research Areas
AI & Machine Learning
3D Deep Learning
Computer Vision
Robotics
All Areas
Careers
Academic Collaborations
Government Collaborations
Graduate Fellowship
Internships
Research Openings
Research Scientists
Meet the Team
Licensing
Skip to main content
Artificial Intelligence Computing Leadership from NVIDIA
Login
Research Labs
All Research Labs
3D Deep Learning
Applied Research
Autonomous Vehicles
Deep Imagination
Publications
AI Playground
New and Featured
AI Art Gallery
NGC Demos
Research Areas
AI & Machine Learning
3D Deep Learning
Computer Vision
Robotics
All Areas
Careers
Academic Collaborations
Government Collaborations
Graduate Fellowship
Internships
Research Openings
Research Scientists
Meet the Team
Licensing
Search
Search
Enter the terms you wish to search for.
Publications
Our publications provide insight into some of our leading-edge research.
Filters
Search
Apply
Filters
Filters
Publication Year
2025
(7)
2024
(1)
2023
(5)
2022
(5)
2021
(9)
2020
(6)
2019
(15)
2018
(14)
2017
(15)
2016
(9)
2015
(4)
2014
(1)
2012
(4)
2011
(4)
2010
(3)
2009
(2)
2008
(1)
2005
(1)
Facet Publication Year
Research Areas
High Performance Computing
(20)
Artificial Intelligence and Machine Learning
(5)
Computer Architecture
(5)
Programming Languages, Systems and Tools
(5)
Algorithms and Numerical Methods
(3)
Computer Graphics
(3)
Real-Time Rendering
(3)
Resilience and Safety
(3)
Networking
(2)
Climate Simulation
(1)
Events
No Results Available
20 results found
High Performance Computing
Clear all
2020
2018
High Performance Computing
2020
Accelerating Reinforcement Learning through GPU Atari Emulation
Iuri Frosio
,
Steven Dalton
Locality-Centric Data and Threadblock Management for Massive GPUs
Mahmoud Khairy, Vadim Nikiforov,
David Nellans
, Timothy G. Rogers
EMOGI: Efficient Memory-access for Out-of-memory Graph-traversal In GPUs
Seung Won Min, Vikram Sharma Mailthody, Zaid Qureshi, Jinjun Xiong, Eiman Ebrahimi, Wen-mei Hwu
Buddy Compression: Enabling Larger Memory for Deep Learning and HPC Workloads on GPUs
Esha Chouske,
Michael B. Sullivan
,
Mike O'Connor
, Mattan Erez, Jeff Pool,
David Nellans
,
Steve Keckler
An In-Network Architecture for Accelerating Shared-Memory Multiprocessor Collectives
Benjamin Klenk
,
Ted Jiang
, Greg Thorson,
Larry Dennison
NWChem: Past, Present, and Future
Edoardo Aprà, Many others,
Oreste Villa
, Many others
2018
Dynamic Tracing: Memoization of Task Graphs for Dynamic Task-based Runtimes
Wonchan Lee, Elliott Slaughter,
Michael Bauer
, Sean Treichler, Todd Warszawski,
Michael Garland
, Alex Aiken
Exascale Deep Learning for Climate Analytics
Thorsten Kurth, Sean Treichler, Joshua Romero, Mayur Mudigonda, Nathan Luehr, Everett Phillips, Ankur Mahesh, Michael Matheson, Jack Deslippe, Massimiliano Fatica, Prabhat, Michael Houston
Evaluating and Accelerating High-Fidelity Error Injection for HPC
Chun-Kai Chang, Sangkug Lym, Nicholas Kelly,
Michael B. Sullivan
, Mattan Erez
Exploiting Idle Resources in a High-Radix Switch for Supplemental Storage
Matthias Blumrich
,
Ted Jiang
,
Larry Dennison
Fast, High Precision Ray/Fiber Intersection using Tight, Disjoint Bounding Volumes
Nikolaus Binder
,
Alex Keller
Massively Parallel Stackless Ray Tracing of Catmull-Clark Subdivision Surfaces
Nikolaus Binder
,
Alex Keller
Exascale Deep Learning for Climate Analytics
Thorsten Kurth, Sean Treichler, Joshua Romero, Mayur Mudigonda, Nathan Luehr, Everett Phillips, Ankur Mahesh, Michael Matheson, Jack Deslippe, Massimiliano Fatica, Prabhat, Michael Houston
CRUM: Checkpoint-Restart Support for CUDA's Unified Memory
Rohan Garg, Apoorve Mohan,
Michael B. Sullivan
, Gene Cooperman
Phantom Ray-Hair Intersector
Alexander Reshetov
,
David Luebke
Hamartia: A Fast and Accurate Error Injection Framework
Chun-Kai Chang, Sangkug Lym, Nicholas Kelly,
Michael B. Sullivan
, Mattan Erez
Isometry: A Path-Based Distributed Data Transfer System
Zhihao Jia, Sean Treichler, Galen Shipman, Patrick McCormick, Alex Aiken
Structurally Sparsified Backward Propagation for Faster Long Short-Term Memory Training
Maohua Zhu,
Jason Clemons
, Jeff Pool, Minsoo Rhu,
Steve Keckler
, Yuan Xie
Scalable Collectives for Distributed Asynchronous Many-Task Runtimes
Matthew Whitlock, Hemanth Kolla, Sean Treichler, Philippe Pebay, Janine C. Bennett
BabelFlow: An Embedded Domain Specific Language for Parallel Analysis and Visualization
Steve Petruzza, Sean Treichler, Valerio Pascucci, Peer-Timo Bremer