Skip to content

Author

Dingwen Zhang

We have 3 of 44 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

P2Fusion: Prompt-based Progressive Infrared-Visible Image Fusion via Dual-Prior Distillation

Infrared-visible image fusion (IVIF) is pivotal for multimodal perception, yet reconciling the inherent information disparity between thermal and textural features remains a fundamental challenge. Existing prior-guided methods often rely on static constraints that induce optimization conflicts or utilize extrinsic sema...

Yi Shi, Huichao Xie, Yuqing Wang et al. · 0 citations
Jul 2026

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery

This work introduces Universal Referring, a generalized UAV referring task that jointly expands the query modality and the output cardinality, and presents UAV-URNet, a detection-style baseline that maps heterogeneous queries into a shared query space and predicts variable-size target sets through set prediction.

Haibin Tian, Huichao Xie, Xue-Lin Qian et al. · 0 citations
Preprint Aug 2026

Annotations as Rollouts: Efficient and Scalable Reinforcement Learning for Video MLLMs

The sample efficiency and scalability of RL post-training for video MLLMs and introduces OraRL, a decoupled advantage estimator that scales with model size and data, surpassing its backbone from 0.8B to 9B and GRPO up to 100k prompts.

Yunheng Li, Guo-Hong Mu, Hao Li et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.