Skip to content

Author

Tianrui Zhu

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning

Variance Recovery Policy Optimization (VRPO) is introduced, which retains and progressively expands groups to recover informative signals from prompts that are difficult yet solvable and retains and progressively expands these groups to recover informative signals from prompts that are difficult yet solvable.

Jingqi Tian, Haoji Zhang, Lin Chen et al. · 0 citations