Skip to content

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

VicEdit: Learning to Edit Videos from Visual In-Context Examples

This work proposes Visual In-context Editing, a new paradigm elevating video editing from textual instructions to multi-modal visual guidance encompassing single image, image pair, and video pair, and curates VicEdit-400K, the first large-scale dataset for visual in-context video editing.

Yuji Wang, Teng Hu, Yuheng Chen et al. · 0 citations
Preprint Aug 2026

Efficient Audio-Visual Generation via Synchrony-Aware Cross-Modal Sparse Attention

This work presents a synchronization-aware acceleration framework for efficient audio-visual generation by explicitly accounting for cross-modal dependence during acceleration, and improves inference efficiency while keeping video quality, audio quality, and audio-video synchronization.

Sheng-Chuan Gao, Teng Hu, Bohao Feng et al. · 0 citations
Preprint Jul 2026

Cycle-World: Mitigating Error Accumulation in Long-term Video World Models via Reverse-Prediction Cycle Consistency

This work proposes Cycle-World, a novel framework designed for stable and temporally consistent long-video generation that tackles error drift by enforcing strict temporal reversibility across both the training and inference phases, and demonstrates that forward generative drift can be strictly bottlenecked by a cycle-consistency objective.

Zihan Su, Teng Hu, Jiangning Zhang et al. · 1 citation