跳到主要导航 跳到搜索 跳到主要内容

Unified Attention Distillation with Decoupled Progressive Scheduling for Object Detection

  • Center for Interdisciplinary Research and Innovation
  • Westlake University

科研成果: 期刊稿件文章同行评审

摘要

Knowledge distillation (KD) has become a key technique for compressing object detection models while maintaining accuracy, enabling deployment in latency-sensitive scenarios such as autonomous driving and drone surveillance. However, existing detection-oriented KD approaches face two major limitations: first, attention modules transplanted from classification tasks fail to capture fine-grained localization cues and perform poorly under occlusion; second, the naive combination of feature imitation and response mimicking often leads to performance degradation and slower training. To address these issues, we propose a Unified Attention Distillation (UAD) framework that integrates masked local attention to enhance textures and edges with a Row-Column Masked Global Attention (RCMGA) module for long-range dependency modeling and occlusion-aware suppression. Furthermore, we introduce a Decoupled Progressive Schedule (DPS) that progressively transitions from unified to global attention in terms of spatial granularity, and from feature imitation to response mimicking in terms of knowledge hierarchy, thereby harmonizing the dynamics of knowledge transfer. Extensive experiments on COCO and PASCAL VOC demonstrate that UAD combined with DPS achieves state-of-the-art detection performance while reducing training time by 11.9% compared to existing distillation methods.

源语言英语
期刊IEEE Transactions on Multimedia
DOI
出版状态已接受/待刊 - 2026

学术指纹

探究 'Unified Attention Distillation with Decoupled Progressive Scheduling for Object Detection' 的科研主题。它们共同构成独一无二的学术指纹。

引用此