跳到主要导航 跳到搜索 跳到主要内容

Towards Effective Foundation Model Adaptation for Extreme Cross-Domain Few-Shot Learning

  • Northwestern Polytechnical University Xian
  • University of Electronic Science and Technology of China
  • Xi'an Institute of Posts and Telecommunications
  • Nanyang Technological University

科研成果: 书/报告/会议事项章节会议稿件同行评审

2 引用 (Scopus)

摘要

Large-scale pre-trained foundation models have demonstrated remarkable generalization capabilities across diverse computer vision tasks through fine-tuning. However, existing fine-tuning approaches often encounter challenges in extreme cross-domain few-shot learning scenarios, primarily due to the significant domain shift between pre-training data and target tasks, as well as the scarcity of annotated target samples. To mitigate this issue, we propose a novel absorption adaptation learning framework which meticulously regularizes the fine-tuning procedure of foundation model using an expert model with the same architecture but trained from scratch on the targeted data in two aspects. On one hand, we first design a masked cross-model unidirectional reconstruction scheme, which forces the foundation model to recover the intermediate feature of the expert model in a randomly masked manner. On the other hand, a decision graph association loss is developed to encourage the consistency of token similarity matrix between these two models. By doing these, the task-relevant semantic knowledge in the expert model from both intermediate feature and the final decision levels are appropriately extracted and absorbed by the foundation model during its fine-tuning, thus mitigating the performance drop caused by domain gap and limited annotation. Sufficient experiments with further observations and analyses underpin our observation and argument. The code is available at https://github.com/NWPUZhoufei/FMA.

源语言英语
主期刊名Proceedings - 2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025
出版商Institute of Electrical and Electronics Engineers Inc.
4582-4593
页数12
ISBN(电子版)9798331587758
DOI
出版状态已出版 - 2025
活动2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025 - Honolulu, 美国
期限: 19 10月 202523 10月 2025

丛书

姓名Proceedings of the IEEE International Conference on Computer Vision
ISSN(印刷版)1550-5499
ISSN(电子版)2380-7504

会议

会议2025 IEEE/CVF International Conference on Computer Vision, ICCV 2025
国家/地区美国
Honolulu
时期19/10/2523/10/25

学术指纹

探究 'Towards Effective Foundation Model Adaptation for Extreme Cross-Domain Few-Shot Learning' 的科研主题。它们共同构成独一无二的学术指纹。

引用此