This section reviews literature related to Instance-Incremental Learning (IIL), contrasting it with the more explored Class-Incremental LearningThis section reviews literature related to Instance-Incremental Learning (IIL), contrasting it with the more explored Class-Incremental Learning

Incremental Learning: Comparing Methods for Catastrophic Forgetting and Model Promotion

2025/11/05 02:00

Abstract and 1 Introduction

  1. Related works

  2. Problem setting

  3. Methodology

    4.1. Decision boundary-aware distillation

    4.2. Knowledge consolidation

  4. Experimental results and 5.1. Experiment Setup

    5.2. Comparison with SOTA methods

    5.3. Ablation study

  5. Conclusion and future work and References

    \

Supplementary Material

  1. Details of the theoretical analysis on KCEMA mechanism in IIL
  2. Algorithm overview
  3. Dataset details
  4. Implementation details
  5. Visualization of dusted input images
  6. More experimental results

2. Related works

This paper devotes to the instance-incremental learning which is an associated topic to the CIL but seldom investigated. In the following, related topics on class-incremental learning, continual domain adaptation, and methods based on knowledge distillation (KD) are introduced.

\ Class-incremental learning. CIL is proposed to learn new classes without suffering from the notorious catastrophic forgetting problem and is the main topic that most of works focused on in this area. Methods of CIL can be categorized into three types: 1) important weights regularization [1, 10, 19, 32], which constrains the important weights for old tasks and free those unimportant weights for new task. Freezing the weights limits the ability to learn from new data and always lead to a inferior performance on new classes. 2) Rehearsal or pseudo rehearsal method, which stores a small-size of typical exemplars [2, 4, 9, 22] or relies on a generation network to produce old data [23] for old knowledge retaining. Usually, these methods utilize knowledge distillation and perform over the weight regularization method. Although the prototypes of old classes are efficacy in preserving knowledge, they are unable to promote the model’s performance on hard samples, which is always a problem in real deployment. 3) Dynamic network architecture based method [8, 15, 30, 31], which adaptively expenses the network each time for new knowledge learning. However, deploying a changing neural model in real scenarios is troublesome, especially when it goes too big. Although most CIL methods have strong ability in learning new classes, few of them can be directly utilized in the new IIL setting in our test. The reason is that performance promotion on old classes is less emphasized in CIL.

\ Knowledge distillation-based incremental learning. Most of existing incremental learning works utilize knowledge distillation (KD) to mitigate catastrophic forgetting. LwF [12] is one of the earliest approaches that constrains the prediction of new data through KD. iCarl [22] and many other methods distill knowledge on preserved exemplars to free the learning capability on new data. Zhai et al. [33] and Zhang et al. [34] exploit distillation with augmented data and unlabeled auxiliary data at negligible cost. Different from above distillation at label level, Kang et al. [9] and Douillard [4] proposed to distill knowledge at feature level for CIL. Compared to the aforementioned researches, the proposed decision boundary-aware distillation requires no access to old exemplars and is simple but effective in learning new as well as retaining the old knowledge.

\ Comparison with the CDA and ISL. Rencently, some work of continual domain adptation (CDA) [7, 21, 27] and incremental subpopulation learning (ISL) [13] is proposed and has high similarity with the IIL setting. All of the three settings have a fixed label space. The CDA focus on solving the visual domain variations such as illumination and background. ISL is a specific case of CDA and pays more attention to the subcategories within a class, such as Poodles and Terriers. Compared to them, IIL is a more general setting where the concept drift is not limited to the domain shift in CDA or subpopulation shifting problem in ISL. More importantly, the new IIL not only aims to retain the performance but also has to promote the generalization with several new observations in the whole data space.

\

:::info Authors:

(1) Qiang Nie, Hong Kong University of Science and Technology (Guangzhou);

(2) Weifu Fu, Tencent Youtu Lab;

(3) Yuhuan Lin, Tencent Youtu Lab;

(4) Jialin Li, Tencent Youtu Lab;

(5) Yifeng Zhou, Tencent Youtu Lab;

(6) Yong Liu, Tencent Youtu Lab;

(7) Qiang Nie, Hong Kong University of Science and Technology (Guangzhou);

(8) Chengjie Wang, Tencent Youtu Lab.

:::


:::info This paper is available on arxiv under CC BY-NC-ND 4.0 Deed (Attribution-Noncommercial-Noderivs 4.0 International) license.

:::

\

Disclaimer: The articles reposted on this site are sourced from public platforms and are provided for informational purposes only. They do not necessarily reflect the views of MEXC. All rights remain with the original authors. If you believe any content infringes on third-party rights, please contact [email protected] for removal. MEXC makes no guarantees regarding the accuracy, completeness, or timeliness of the content and is not responsible for any actions taken based on the information provided. The content does not constitute financial, legal, or other professional advice, nor should it be considered a recommendation or endorsement by MEXC.
Share Insights

You May Also Like

Crypto Market Cap Edges Up 2% as Bitcoin Approaches $118K After Fed Rate Trim

Crypto Market Cap Edges Up 2% as Bitcoin Approaches $118K After Fed Rate Trim

The global crypto market cap rose 2% to $4.2 trillion on Thursday, lifted by Bitcoin’s steady climb toward $118,000 after the Fed delivered its first interest rate cut of the year. Gains were measured, however, as investors weighed the central bank’s cautious tone on future policy moves. Bitcoin last traded 1% higher at $117,426. Ether rose 2.8% to $4,609. XRP also gained, rising 2.9% to $3.10. Fed Chair Jerome Powell described Wednesday’s quarter-point reduction as a risk-management step, stressing that policymakers were in no hurry to speed up the easing cycle. His comments dampened expectations of more aggressive cuts, limiting enthusiasm across risk assets. Traders Anticipated Fed Rate Trim, Leaving Little Room for Surprise Rally The Federal Open Market Committee voted 11-to-1 to lower the benchmark lending rate to a range of 4.00% to 4.25%. The sole dissent came from newly appointed governor Stephen Miran, who pushed for a half-point cut. Traders were largely prepared for the move. Futures markets tracked by the CME FedWatch tool had assigned a 96% probability to a 25 basis point cut, making the decision widely anticipated. That advance positioning meant much of the potential boost was already priced in, creating what analysts described as a “buy the rumour, sell the news” environment. Fed Rate Decision Creates Conditions for Crypto, But Traders Still Hold Back Andrew Forson, president of DeFi Technologies, said lower borrowing costs would eventually steer more money toward digital assets. “A lower cost of capital indicates more capital flows into the digital assets space because the risk hurdle rate for money is lower,” he noted. He added that staking products and blockchain projects could become attractive alternatives to traditional bonds, offering both yield and appreciation. Despite the cut, crypto markets remained calm. Open interest in Bitcoin futures held steady and no major liquidation cascades followed the Fed’s decision. Analysts pointed to Powell’s language and upcoming economic data as the key factors for traders before building larger positions. Powell’s Caution Tempers Immediate Impact of Fed Rate Move on Crypto Markets History also suggests crypto rallies after rate cuts often take time. When the Fed eased in Dec. 2024, Bitcoin briefly surged 5% cent before consolidating, with sustained gains arriving only weeks later. This time, market watchers are bracing for a similar pattern. Powell’s insistence on caution, combined with uncertainty around inflation and growth, has kept short-term volatility muted even as sentiment for risk assets improves. BitMine’s Tom Lee this week predicted that Bitcoin and Ether could deliver “monster gains” in the next three months if the Fed continues on an easing path. His view echoes broader expectations that liquidity-sensitive assets will outperform once the cycle gathers pace. For now, the crypto sector has digested the Fed’s move with restraint. Traders remain focused on signals from the central bank’s October meeting to determine whether Wednesday’s step marks the beginning of a broader policy shift or just a one-off adjustment
Share
CryptoNews2025/09/18 13:14