Aug 2026
Fadet: a fusion-aware 3D detection network with cascaded feature enhancement for small object detection in autonomous driving
Changhong Yu, Shaoshi Luo, Wenli Shen
· Multimedia Systems · 0 citations
2 papers indexed here
We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.
Not the right person? Other researchers publish under this name.
This work proposes a novel framework that employs the pre-trained vision-language model BLIP (Bootstrapping Language-Image Pre-training) to generate descriptive image captions and introduces an aspect-guided soft prompt mechanism that enables dynamic interaction between aspect terms and multimodal features, thereby mitigating the effects of structural irregularities.