Memoria: Hebbian Memory Architecture for Human-Like Sequential Processing
Sangjun Park, JinYeong Bak
OpenReview ground truth
TL;DR — Memoria applying Hebbian theory to enhance long-term memory capabilities.
Abstract
Transformers have demonstrated their success in various domains and tasks. However, Transformers struggle with long input sequences due to their limited capacity. While one solution is to increase input length, endlessly stretching the length is unrealistic. Furthermore, humans selectively remember and use only relevant information from inputs, unlike Transformers which process all raw data from start to end. We introduce Memoria, a general memory network that applies Hebbian theory which is a major theory explaining human memory formulation to enhance long-term dependencies in neural networks. Memoria stores and retrieves information called engram at multiple memory levels of working memory, short-term memory, and long-term memory, using connection weights that change according to Hebb's rule. Through experiments with popular Transformer-based models like BERT and GPT, we present that Memoria significantly improves the ability to consider long-term dependencies in various tasks. Results show that Memoria outperformed existing methodologies in sorting and language modeling, and long text classification.
Author context
Most prolific author: 2 submissions (credibility 1.00).
No mass-submission penalty for this paper (authors within normal submission volume).
Aggregate statistics only — no individual author rankings.
Ranking trajectory
Percentile by tournament round — convergence indicates rating stability.
Battle history — 36 comparisons
Ranked above opponent in 41% of matchups.
- ▲ beat Deep Learning-based Discrimination of Paus… ×8
- ▼ lost to MT-Ranker: Reference-free machine translat… ×6
- ▲ beat Enhancing Fine-Tuning Performance of Large… ×6
- ▼ lost to Efficient ConvBN Blocks for Transfer Learn… ×4
- ▼ lost to BiDST: Dynamic Sparse Training is a Bi-Lev… ×4
Judge assessments
Mean overall score 0.0 ± 0.0 (n = 36)