Memory-augmented decoder leads when neural data are scarce
A preprint reports higher VN-SST scores than a Transformer on three neural benchmarks, with the clearest gap under limited training data.
Developing Light ยท https://developinglight.com/editorial/developing-light