Skip to content

Similar papers

Open access Aug 2026

Semantic Topological Multi-Scale Part Network for Fine-Grained Visual Classification

A Prior-Guided Part Aggregator (PGA) is designed, which leverages the foreground prior provided by foundation models to guide discriminative part discovery, enhancing target region responses while suppressing background interference, and a Topology-Informed Semantic Graph Convolutional Network (TIS-GCN) is designed to...

Xue-Rong Liu, Min Zhi, Yan-Jun Yin et al. · 0 citations
Sep 2026

Multimodal-guided self-distillation for unified person search.

Person search is challenging due to limitations in identity representation. Existing methods rely on one-hot encoding, ignoring semantic relationships among pedestrians. This leads to a fragmented feature space and reduces generalization ability, especially in large-scale scenarios with a significant proportion of unla...

Xi Yang, He-Xun Zhou, Hai-Yang Zhu et al. · 0 citations
Aug 2026

CLIP-SGI: A Semantic-Guided and Instance-Consistent Framework for Generalizable Person Re-Identification

This work proposes CLIP-SGI, a semantic-guided and instance-consistent framework for generalizable person ReID that combines semantic guidance, domain-aware representation learning, and instance consistency to improve robustness under domain shifts.

Dai-Xin Liu, Yu Yang, Linlin Tang et al. · 0 citations
Conference Aug 2026

Segmentation-Guided Scene Perturbation for Passage-Based Top-View Person Re-Identification

Passage-based top-view person re-identification aims to match individuals across short overhead walking videos. This privacy-preserving setting is challenging because overhead cameras suppress facial and body-part cues, compress pedestrian appearance, and make models vulnerable to scene shortcuts from floors, ramps, an...

Hien Pham Duy, Bao Tran, Tien Do et al. · 0 citations
Preprint Sep 2026

FineHOI: Part-Aware Dense Representations for Zero-Shot Human-Object Interaction Detection

Human-Object Interaction (HOI) detection aims to localize humans and objects in images and classify their interactions. Zero-shot HOI focuses on recognizing interactions that are not observed during training, requiring models to generalize beyond seen verb-object compositions. Recent approaches leverage Vision-Language...

Francesco Tonini, Lorenzo Vaquero, Mohammad Mahdi Derakhshani et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.