Preprint
Aug 2026
Persistent Object Narratives for Token-Efficient Video Language Models
Experimental results establish persistent object narratives as a compact, structured, and temporally organized visual interface for Video-LLMs as a favorable trade-off between accuracy and visual-token count compared with prior compact Video-LLM interfaces.
Junzhe Chen, Siyuan Meng, Xiaojie Guo
· 0 citations