In this work, we propose an enhanced framework named Unified Sequential Search and Recommendation (UnifiedSSR+) for joint modeling of user behavior across both search and recommendation scenarios. Specifically, we consider user-interacted products in the recommendation scenario, user-interacted products and user-issued...
Qian-Qian Zhang, Jiayi Xie, Yi-Kang Zhou et al.· ACM Transactions on Intellig...· 0 citations
XL-DocBench is introduced, a fully human-verified benchmark for extra-long document understanding, with 1,519 retained questions from six professional domains and contexts up to 2,303 pages, and fills a gap left by prior single-page, short multi-page, or text-only long-context benchmarks.
Hongchen Wei, Yuanzhe Wang, Bei Liu et al.· 0 citations
DocAtlas is presented, a system that treats long-document understanding as a mutable-state information-seeking process and instantiate DocAtlas as a mutable document harness: an external environment that determines what document information is searched, read, stored, reviewed, and shown to the model at each step.
Hongchen Wei, Yuanzhe Wang, Bei Liu et al.· 0 citations
DiffVC-ONE, a diffusion-based generative video compression framework built on a one-step Video Diffusion Transformer, is proposed and a Unified Unidirectional Latent Compressor that uses a shared model to efficiently and uniformly compress compact latent slices is introduced.
Coding Prior-enhanced Dual Conditioning branches are designed to jointly model compressed video and coding prior conditions, where coding priors including residuals and motion vectors provide complementary structural and motion guidance during the diffusion denoising process.
Wenqiang Xiao, Wenzhuo Ma, Junxi Zhang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.