• Home
  • Search
  • Parameter-Free Logit Distillation via Sorting Mechanism
  • https://doi.org/10.1109/lsp.2025.3602654Copy DOI Icon

Parameter-Free Logit Distillation via Sorting Mechanism

Show More
  • Abstract
  • Literature Map
  • References
  • Similar Papers
Abstract

Knowledge distillation (KD) aims to distill the knowledge from the teacher (larger) to the student (smaller) model via soft-label for the efficient neural network. In general, the performance of a model is determined by accuracy, which is measured with labels. However, existing KD approaches usually use the teacher with its original distribution, neglecting the potential of incorrect prediction. This may contradict the motivation of hard-label learning through cross-entropy loss, which may lead to sub-optimal knowledge distillation on certain samples. To address this issue, we propose a novel logit processing scheme via a sorting mechanism. Specifically, our method has a two-fold goal: (1) fixing the incorrect prediction of the teacher based on the labels and (2) reordering the distribution in a natural way according to priority rank at once. As an easy-to-use, plug-andplay pre-processing, our sort method can be effectively applied to existing logit-based KD methods. Extensive experiments on the CIFAR-100 and ImageNet datasets demonstrate the effectiveness of our method.

Similar Papers
  • Conference Article
  • Citations7

Distilling deep neural networks with reinforcement learning

  • Aug 01, 2018
  • You Huang +1
  • Conference Article
  • Citations5

Extracurricular Learning: Knowledge Transfer Beyond Empirical Distribution

  • Jun 01, 2021
  • Hadi Pouransari +3
  • Research Article
  • Citations46

A Novel Approach to Classifying Breast Cancer Histopathology Biopsy Images Using Bilateral Knowledge Distillation and Label Smoothing Regularization

  • Oct 20, 2021
  • Computational and Mathematical Methods in Medicine
  • Sushovan Chaudhury +4
  • Conference Article

Triplet Knowledge Distillation Networks for Model Compression

  • Jul 18, 2021
  • Jialiang Tang +3
  • Research Article
  • Citations19

Explainability-based knowledge distillation

  • Oct 30, 2024
  • Pattern Recognition
  • Tianli Sun +3
  • PDF
  • Research Article
  • Citations25

Semantic-aware knowledge distillation with parameter-free feature uniformization

  • May 08, 2023
  • Visual Intelligence
  • Guangyu Guo +4
  • Research Article
  • Citations11

Improving adversarial robustness using knowledge distillation guided by attention information bottleneck

  • Mar 05, 2024
  • Information Sciences
  • Yuxin Gong +4
  • Research Article
  • Citations2

Maintaining Fairness in Logit-based Knowledge Distillation for Class-Incremental Learning

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Zijian Gao +7
  • Conference Article
  • Citations29

Lipschitz Continuity Guided Knowledge Distillation

  • Oct 01, 2021
  • Yuzhang Shang +4
  • Conference Article

Feature Point Detection Algorithm Based on Sequential Neural Networks and Knowledge Distillation

  • Sep 13, 2024
  • Mu He
  • Book Chapter
  • Citations44

Differentiable Feature Aggregation Search for Knowledge Distillation

  • Jan 01, 2020
  • Yushuo Guan +6
  • Research Article
  • Citations2

Enhancing marine oil spill detection through dynamic adaptive knowledge distillation with spectral mask superpixel.

  • Oct 01, 2025
  • Marine pollution bulletin
  • Shuang Dong +4
  • Conference Article
  • Citations13

Rectifying the Data Bias in Knowledge Distillation

  • Oct 01, 2021
  • Boxiao Liu +4
  • Research Article
  • Citations4

Robust knowledge distillation based on feature variance against backdoored teacher model

  • Jun 24, 2024
  • Applied Soft Computing
  • Jinyin Chen +5
  • Conference Article
  • Citations1

Lightweight Network for Vietnamese Landmark Recognition based on Knowledge Distillation

  • Nov 26, 2020
  • Viet T Tran +3
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.