NobleBlocks
Public

Code generation and runtime techniques for enabling data-efficient deep learning training on GPUs

Published in Zenodo (CERN European Organization for Nuclear Research) • Dec 5, 2024
Authors:
Wu, Kun

Abstract

As deep learning models scale, their training cost has surged significantly. Due to both hardware advancements and limitations in current software stacks, the need for data efficiency has risen. Data efficiency refers to the effective hiding of data access latency and the avoidance of unnecessary da...

Finding related papers...

Discussions

(0)

No comments yet

Be the first to share your thoughts!