• Home
  • Search
  • Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental Learning
  • Cite Icon2
  • https://doi.org/10.1609/aaai.v39i16.33842Copy DOI Icon

Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental Learning

  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

Logit-based knowledge distillation (KD) is commonly used to mitigate catastrophic forgetting in class-incremental learning (CIL) caused by data distribution shifts. However, the strict match of logit values between student and teacher models conflicts with the cross-entropy (CE) loss objective of learning new classes, leading to significant recency bias (i.e. unfairness). To address this issue, we rethink the overlooked limitations of KD-based methods through empirical analysis. Inspired by our findings, we introduce a plug-and-play pre-process method that normalizes the logits of both the student and teacher across all classes, rather than just the old classes, before distillation. This approach allows the student to focus on both old and new classes, capturing intrinsic inter-class relations from the teacher. By doing so, our method avoids the inherent conflict between KD and CE, maintaining fairness between old and new classes. Additionally, recognizing that overconfident teacher predictions can hinder the transfer of inter-class relations (i.e., dark knowledge), we extend our method to capture intra-class relations among different instances, ensuring fairness within old classes. Our method integrates seamlessly with existing logit-based KD approaches, consistently enhancing their performance across multiple CIL benchmarks without incurring additional training costs.

Similar Papers
  • Conference Article
  • Citations2

Improving Feature Generalizability with Multitask Learning in Class Incremental Learning

  • May 23, 2022
  • Dong Ma +2
  • Research Article
  • Citations36

Curiosity-Driven Class-Incremental Learning via Adaptive Sample Selection

  • Dec 01, 2022
  • IEEE Transactions on Circuits and Systems for Video Technology
  • Qinghua Hu +2
  • Conference Article
  • Citations2

KABI: Class-Incremental Learning via knowledge Amalgamation and Batch Identification

  • Mar 05, 2021
  • Caixia Li +3
  • Research Article
  • Citations34

A Class-Incremental Approach With Self-Training and Prototype Augmentation for Specific Emitter Identification

  • Jan 01, 2024
  • IEEE Transactions on Information Forensics and Security
  • Dingzhao Li +4
  • Book Chapter
  • Citations79

Few-Shot Class-Incremental Learning via Entropy-Regularized Data-Free Replay

  • Jan 01, 2022
  • Huan Liu +6
  • Research Article

FeTT: Class-Incremental Learning with Feature Transformation Tuning

  • Mar 27, 2025
  • Mathematics
  • Sunyuan Qiang +1
  • Research Article
  • Citations2

Class incremental learning via dynamic regeneration with task-adaptive distillation

  • Dec 28, 2023
  • Computer Communications
  • Hao Yang +4
  • Research Article

Parameter-Free Logit Distillation via Sorting Mechanism

  • Jan 01, 2025
  • IEEE Signal Processing Letters
  • Stephen Ekaputra Limantoro
  • Research Article
  • Citations1

Class Incremental Learning With Large Domain Shift

  • Jan 01, 2024
  • IEEE Access
  • Kamin Lee +4
  • Research Article
  • Citations1

Class Incremental Learning Method Based on Dynamic Structure Extension and Feature enhancement

  • Jul 29, 2024
  • Journal of Computing and Electronic Information Management
  • Zhenghu Li +1
  • Research Article

Dilated Balanced cross entropy loss for medical image segmentation.

  • Feb 25, 2026
  • BMC medical imaging
  • Seyed Mohsen Hosseini +1
  • Conference Article
  • Citations7

Comparative Analysis of Catastrophic Forgetting in Metric Learning

  • Nov 14, 2020
  • Jiahao Huo +1
  • Conference Article

Natural Statistics Of Network Activations And Implications For Knowledge Distillation

  • Sep 19, 2021
  • Michael Rotman +1
  • Research Article
  • Citations2

Enhancing marine oil spill detection through dynamic adaptive knowledge distillation with spectral mask superpixel.

  • Oct 01, 2025
  • Marine pollution bulletin
  • Shuang Dong +4
  • Research Article
  • Citations4

HyperKD: Lifelong Hyperspectral Image Classification With Cross-Spectral–Spatial Knowledge Distillation

  • Jan 01, 2025
  • IEEE Transactions on Geoscience and Remote Sensing
  • Zhenlin Li +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.