Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

May 02, 2025

Boyuan Meng, Xiaohan Zhang, Peilin Li, Zhe Wu, Yiming Li, Wenkai Zhao, Beinan Yu, Hui-Liang Shen

Figure 1 for CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

Figure 2 for CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

Figure 3 for CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

Figure 4 for CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

Share this with someone who'll enjoy it:

Abstract:Cross-domain few-shot object detection (CD-FSOD) aims to detect novel objects across different domains with limited class instances. Feature confusion, including object-background confusion and object-object confusion, presents significant challenges in both cross-domain and few-shot settings. In this work, we introduce CDFormer, a cross-domain few-shot object detection transformer against feature confusion, to address these challenges. The method specifically tackles feature confusion through two key modules: object-background distinguishing (OBD) and object-object distinguishing (OOD). The OBD module leverages a learnable background token to differentiate between objects and background, while the OOD module enhances the distinction between objects of different classes. Experimental results demonstrate that CDFormer outperforms previous state-of-the-art approaches, achieving 12.9% mAP, 11.0% mAP, and 10.4% mAP improvements under the 1/5/10 shot settings, respectively, when fine-tuned.

View paper on

Share this with someone who'll enjoy it:

Title:CDFormer: Cross-Domain Few-Shot Object Detection Transformer Against Feature Confusion

Paper and Code