Throughput-oriented GPU memory allocation
Published • Feb 5, 2019
NobleIDNI9P35W08R45S24
Authors:,
Isaac Gelado
Michael Garland
Abstract
Throughput-oriented architectures, such as GPUs, can sustain three orders of magnitude more concurrent threads than multicore architectures. This level of concurrency pushes typical synchronization primitives (e.g., mutexes) over their scalability limits, creating significant performance bottlenecks...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!