Skip to content

Author

Weihua Luo

We have 6 of 78 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Test-Time Scaling for Video Diffusion Models via Diagnosis-Guided Candidate Recycling

Recent video diffusion models have achieved remarkable generation quality, but high-fidelity results still largely depend on closed-source systems or costly large-scale infrastructure. Test-time scaling (TTS) offers a training-free way to improve lightweight generators by spending additional inference compute, yet exis...

Hangzhou He, Lun-Hao Duan, Shan-Shan Zhao et al. · 0 citations
Preprint Aug 2026

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-evolution in expert domains. Existing self-evolving skill methods construct skills internally from the model's parametric knowledge or trajectories, and are therefore bo...

Muyang Ye, Tian Lan, Feihu Jiang et al. · 1 citation
#artificial intelligence Preprint Sep 2026

CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning

Credit-Aware Hierarchical Memory Evolution (CHIME), a self-evolving memory framework that maintains a separate planning bank and execution bank and follows an attribute-before-memorize principle, which shows that CHIME consistently outperforms state-of-the-art training-based and self-evolving memory baselines.

Yongshi Ye, Tian Lan, Feihu Jiang et al. · 0 citations
#natural language process... Preprint Aug 2026

WnW: Waxing-and-Waning KV Cache for Long-Form Speech LLMs

WnW (Waxing-and-Waning KV cache), which classifies KV-heads into anchor, tidal, and fixed roles via offline calibration, and preserves near-Full-Cache accuracy while keeping only 20% of audio tokens on GPU, where prefill-only baselines fail to terminate.

Yi-Ming Yao, Chenyang Lyu, Xuan-Fan Ni et al. · 0 citations
Preprint Jul 2026

OvisOCR2 Technical Report

This work introduces OvisOCR2, a 0.8B document parsing model designed as an end-to-end parser that combines filtered real-document annotations with synthetic pages whose rendered images and Markdown targets are derived from the same HTML source.

Shiyin Lu, Yinglun Li, Yu Xia et al. · 1 citation
2025

Alleviating Hallucinations in Large Language Models through Multi-Model Contrastive Decoding and Dynamic Hallucination Detection

This work proposes M ulti-Model C ontrastive D ecoding (MCD), which integrates a pretrained language model with an evil model and a truthful model for contrastive decoding and effectively reduces hallucinations in LLMs and outperforms state-of-the-art methods across various benchmarks.

Chenyu Zhu, Yefeng Liu, Hao Zhang et al. · 7 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.