Skip to content

Author

Hong-Yi Fu

We have 5 of 5 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Guide, Then Let Go: Gap-Adaptive Teacher Scheduling for Sparse-Reward Agentic RL

Reinforcement learning for long-horizon agents typically relies on sparse outcome-based rewards. This leads to a severe cold-start problem, as early-stage policies often fail to solve sampled tasks, leaving little useful reward signal for learning. To mitigate this problem, we use on-policy distillation (OPD) to provid...

You-Ling Huang, Tian-Kuo Xu, Jia-Ji Liu et al. · 0 citations
#natural language process... Preprint Sep 2026

GSM: Efficient Language Modeling with Shared Global State

The Global State Model (GSM), a causal encoder--decoder architecture that concentrates the selection and aggregation of long-range information in the encoding stage, is introduced, offering a shared-state architecture for efficient language modeling.

Yun Zheng, Bin Wen, Xiao-Jie Wang et al. · 0 citations
Preprint Aug 2026

Thinking Beyond Videos: Unifying Video Reasoning and Deep Research for Open-World Video Agents

Open-world video understanding often requires a model to locate sparse visual evidence and acquire external knowledge that is absent from the video and its parametric memory. While Thinking-with-Videos enables active temporal perception and Deep Research supports multi-step information seeking, the two capabilities are...

Wenqi Liu, Shijie Ma, Yun-Xiao Wang et al. · 0 citations
#natural language process... Preprint Sep 2026

Lngram v2: Latent N-Gram Memory with Interpretable Discrete Representations

Transformers lack a native lookup mechanism, requiring repeated dense computation to recognize and reuse local static patterns. Lngram v1 introduces tokenizer-independent conditional memory through discrete latent n-gram addressing, but its memory capacity is coupled with the backbone width, limiting scalability due to...

Yun Zheng, Bin Wen, Xiao-Jie Wang et al. · 0 citations
Preprint Jul 2026

KAT-Coder-V2.5 Technical Report

We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-turn code generator. Its capability is bottlenecked less by model scale than by the scarcity of reproducible environments, verifiable rewards, and high-value trajectories, w...

Bofeng Huang, Fengxiang Li, Hao Xu et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.