样本-特征协同的长尾识别算法
作者:
作者单位:

作者简介:

通讯作者:

中图分类号:

基金项目:

国家自然科学基金(62376126); 航空发动机及燃气轮机重大专项基础研究项目(J2019-IV-0018-0086)


Sample-feature Collaborative Long-tailed Recognition Algorithm
Author:
Affiliation:

Fund Project:

  • 摘要
  • |
  • 图/表
  • |
  • 访问统计
  • |
  • 参考文献
  • |
  • 相似文献
  • |
  • 引证文献
  • |
  • 资源附件
  • |
  • 文章评论
    摘要:

    现实场景中的数据分布普遍呈现长尾模式, 导致所训练的深度模型常遭遇头部偏置困境, 即模型偏向于头部类而在尾部类上表现欠佳. 一种简单且有效的策略是通过增广尾部类样本来平衡数据分布. 尽管多数基于该策略的方法能在样本数量上实现重平衡, 但生成的样本仍存在语义漂移和多样性不足的风险, 导致特征分布松散和决策边界有偏, 限制了模型性能. 为此, 提出一种样本-特征协同的长尾识别算法, 旨在构建对数据分布不敏感的分类模型. 具体而言, 在样本层面, 基于傅里叶变换提出“幅值增广”策略, 借助幅值迁移来转换增广样本的风格, 能够在丰富样本多样性的同时保留原有语义信息; 在特征层面, 依据神经坍塌理论提出“特征坍塌”损失, 将类原型对齐至具有最大可分性的等角紧框架, 并促使特征收敛至对应的类原型, 实现最大的类间间隔. 该方法从样本和特征两个层面缓解头部偏置问题, 进而增强类内紧凑性并校准决策边界. 多个基准数据集上的实验结果表明该方法可显著提高长尾识别性能.

    Abstract:

    Data distributions in real-world scenarios commonly exhibit long-tail patterns, causing deep models to suffer from head bias, where performance is biased toward head classes while tail classes are poorly recognized. A simple and effective strategy to alleviate this problem is to balance data distributions by augmenting tail-class samples. Although most methods following this strategy achieve a quantitative rebalancing, the generated samples often suffer from semantic shift and insufficient diversity, which leads to dispersed feature distributions and biased decision boundaries, thereby limiting overall model performance. To address these issues, this study proposes a sample-feature collaborative learning framework for long-tailed recognition, aiming to construct classification models that are insensitive to data distribution imbalance. At the sample level, a “magnitude augmentation” strategy based on the Fourier transform is introduced, in which amplitude shifting is employed to modify the style of augmented samples, while preserving their original semantic information. At the feature level, a “feature collapse” loss inspired by neural collapse theory is proposed to align class prototypes into an equiangular tight frame with maximum separability and to encourage features to converge toward their corresponding class prototypes, achieving maximum inter-class separation. By jointly addressing head bias from both the sample and feature perspectives, the proposed framework enhances intra-class compactness and calibrates decision boundaries. Experimental results across multiple benchmark datasets demonstrate that the proposed method significantly improves long-tailed recognition performance.

    参考文献
    相似文献
    引证文献
引用本文

张恩豪,李超华,王志华,陈松灿.样本-特征协同的长尾识别算法.软件学报,,():1-14

复制
相关视频

分享
文章指标
  • 点击次数:
  • 下载次数:
  • HTML阅读次数:
  • 引用次数:
历史
  • 收稿日期:2025-09-30
  • 最后修改日期:2025-12-01
  • 录用日期:
  • 在线发布日期: 2026-07-08
  • 出版日期:
文章二维码
您是第位访问者
版权所有:中国科学院软件研究所 京ICP备05046678号-3
地址:北京市海淀区中关村南四街4号,邮政编码:100190
电话:010-62562563 传真:010-62562533 Email:jos@iscas.ac.cn
技术支持:北京勤云科技发展有限公司

京公网安备 11040202500063号