Skip to content

Author

Xin-Ting Hu

We have 3 of 13 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

VLM4Cluster: Benchmarking Deep Clustering In the Era of Vision-Language Pre-training

Vision-language pre-training has reshaped image clustering, giving rise to language-assisted image clustering (LaIC), which leverages textual semantics to complement visual representations. Despite the rapid proliferation of LaIC methods, it remains unclear how much LaIC has actually advanced image clustering, as exist...

Yuan Hu, Bo Peng, Yu-Heng Jia et al. · 0 citations
Preprint Sep 2026

SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models

We introduce SolarWM, a fully open foundation for building interactive video world models from data preparation through long-horizon inference. Training across heterogeneous data sources and video backbones is challenging: datasets differ in temporal scale, camera geometry, visual quality, motion, and captioning styles...

Jun-Chao Huang, Gui-An Fang, Sheng-Ju Qian et al. · 1 citation
Jul 2026

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders

VideoRAE is introduced, a representation autoencoder that converts features from a frozen video foundation model into compact, reconstruction-capable latents for video generation, establishing frozen video foundation representations as compact, versatile, and generation-friendly video latents.

Zhihao Xie, Junfeng Wu, Xinting Hu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.