NobleBlocks
Public

Throughput-oriented GPU memory allocation

Published • Feb 5, 2019
NobleIDNI9P35W08R45S24
Authors:
Isaac Gelado
,
Michael Garland

Abstract

Throughput-oriented architectures, such as GPUs, can sustain three orders of magnitude more concurrent threads than multicore architectures. This level of concurrency pushes typical synchronization primitives (e.g., mutexes) over their scalability limits, creating significant performance bottlenecks...

Finding related papers...

Discussions

(0)

No comments yet

Be the first to share your thoughts!