Code generation and runtime techniques for enabling data-efficient deep learning training on GPUs
Published in Zenodo (CERN European Organization for Nuclear Research) • Dec 5, 2024
Authors:
Wu, Kun
Abstract
As deep learning models scale, their training cost has surged significantly. Due to both hardware advancements and limitations in current software stacks, the need for data efficiency has risen. Data efficiency refers to the effective hiding of data access latency and the avoidance of unnecessary da...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!