Indi
std::bodun::blog bodunhu.com

This paper aims to reduce GPU memory usage during DNN training. Capuchin achieves this goal though swapping and recomputation , using tensor as unit of operation. The major question is how to balance …

讨论

还没有评论,来说第一句吧。

std::bodun::blog 的其他文章