A Mathematical Theory of Reusable Neural Bases for Network Compression
As large AI models become increasingly prevalent across a wide range of applications, memory cost has become a critical bottleneck in both training and inference. To mitigate this issue, we introduce the Linear Reusable Neural Bases Architecture (LRNBA), a novel framework aimed at improving parameter efficiency and reducing memory cost. Inspired by recurrent neural network (RNN) designs, the core…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.