Code generation and runtime techniques for enabling data-efficient deep learning training on GPUs
Published in arXiv (Cornell University) • Dec 6, 2024
Authors:
Kun Wu
Abstract
As deep learning models scale, their training cost has surged significantly. Due to both hardware advancements and limitations in current software stacks, the need for data efficiency has risen. Data efficiency refers to the effective hiding of data access latency and the avoidance of unnecessary da...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!