Similar Items: Multi‐Grained Vision–Language Alignment for Domain Generalised Person Re‐Identification
- RainReID: Person Re‐Identification in Rainy Weather and a Large‐Scale Dataset
- OSS–CAEA: Bridging Vision and Language for Open‐Vocabulary Semantic Segmentation via Collaborative Attention and Embedding Alignment
- AT‐ViT: Area‐Targeted Multi‐View Vision Transformer With Cross‐Attention and Multi‐Scale Patching for Plant Trait Recognition in Herbarium Images
- ST‐LoRA: SVD‐Guided Sparse Low‐Rank Adaptation With Trainable Masks for Large Language and Vision Models
- Towards Robust Multimodal Detection via Progressive Cross‐Domain Feature Fusion
- ESFFA: Early‐Stage Feature Frequency Attack in Cross‐Domain Few‐Shot Learning