跳到主要导航 跳到搜索 跳到主要内容

GeoMoE: Geometry-Driven prompts with adaptive Mixture-of-Experts for remote sensing image semantic segmentation

  • Zhen Wang
  • , Ke Wang
  • , Jiayuan Li
  • , Xiao Sun
  • , Nan Xu
  • , Zhuhong You
  • Xijing University
  • Northwestern Polytechnical University Xian
  • Xi'an University of Technology
  • Shenzhen University

科研成果: 期刊稿件文章同行评审

摘要

Recent progress in prompt-driven vision transformers has markedly enhanced remote sensing image (RSI) semantic segmentation. However, existing methods often fail to incorporate fine-grained geometric structure and to effectively address the spatial heterogeneity present in complex aerial scenes. To overcome these limitations, we propose GeoMoE, a geometry-driven adaptive mixture-of-experts framework tailored for precise and efficient semantic segmentation. Building upon a frozen vision transformer backbone, GeoMoE introduces three novel components, 1) a token-field hybrid adapter (TFH-Adapter) that enables geometry-aware feature adaptation without modifying backbone parameters; 2) a geometry-contextual prompt generator (Geo-Prompt) that integrates multi-scale shape prototypes and contextual cues into expressive prompt embeddings; and 3) a geometry mixture-of-complexity-experts (Geo-MoCE) decoder which dynamically routes spatial regions to specialized experts based on local geometric complexity. This unified architecture allows explicit modeling of geometric information and flexible allocation of decoding capacity, resulting in more accurate segmentation of structurally complex and heterogeneous regions. Extensive experiments on several benchmark remote sensing datasets demonstrate that GeoMoE achieves state-of-the-art performance in segmentation accuracy, model efficiency, and boundary delineation.

源语言英语
文章编号133600
期刊Expert Systems with Applications
332
DOI
出版状态已出版 - 1 1月 2027

学术指纹

探究 'GeoMoE: Geometry-Driven prompts with adaptive Mixture-of-Experts for remote sensing image semantic segmentation' 的科研主题。它们共同构成独一无二的学术指纹。

引用此