Bridging the gap between offline and online continual learning
Yaqian Zhang, Eibe Frank, Bernhard Pfahringer, Albert Bifet
OpenReview ground truth
TL;DR — This work provides a theoretical framework to unify online and offline continual learning showing online CL leads to tighter generalization bound.
Abstract
Instead of training deep neural networks offline with a large static dataset, continual learning (CL) considers a new learning paradigm, which continually trains the deep networks from a non-stationary data stream on the fly. Despite the recent progress, continual learning remains an open challenge. Many CL techniques still require offline training of large batches of data chunks (i.e., tasks) over multiple epochs. Conventional wisdom holds that online continual learning, which assumes single-pass data, is strictly harder than offline continual learning, due to the combined challenges of catastrophic forgetting and underfitting within a single training epoch. Here, we challenge this assumption by empirically demonstrating that online CL can match or exceed the performance of its offline counterpart given equivalent memory and computational resources. This finding is further verified across different CL approaches and benchmarks. To better understand these counterintuitive experimental findings, we design a framework to unify and interpolate between online and offline CL and provide a theoretical analysis showing that online CL can yield a tighter generalization bound than offline CL.
Author context
Most prolific author: 2 submissions (credibility 1.00).
No mass-submission penalty for this paper (authors within normal submission volume).
Aggregate statistics only — no individual author rankings.
Ranking trajectory
Percentile by tournament round — convergence indicates rating stability.
Battle history — 44 comparisons
Ranked above opponent in 61% of matchups.
- ▲ beat Navigating the Design Space of Equivariant… ×6
- ▲ beat Simple-TTS: End-to-End Text-to-Speech Synt… ×6
- ▲ beat Continual Memory Neurons ×4
- ▲ beat BOWLL: A DECEPTIVELY SIMPLE OPEN WORLD LIF… ×4
- ▲ beat Federated Causal Discovery from Heterogene… ×4
Judge assessments
Mean overall score 0.0 ± 0.0 (n = 44)