Skip to content

Author

Guanhua Chen

7 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

JustMem: Just-Enough Memory Access for Long-Term Conversations

Efficient long-term conversational memory requires retrieving sufficient evidence without indiscriminately expanding the context presented to the language model. This is challenging because relevant evidence may be distributed across multiple sessions, while compression may discard details needed for answering. Differe...

Guanhua Chen, Yan-Ting Wang, Wen-Jing Zhi et al. · 0 citations
Book Open access Aug 2026

SAFT: Safety-Preserving Adaptation via Fine-Tuning Transfer for Large Language Models

Adapting instruction-tuned large language models (LLMs) to downstream domains is increasingly common, yet fine-tuning on imperfect data can erode the safety alignment learned during post-training. Existing safety-preserving fine-tuning methods typically optimize the aligned instruction model directly, which can destabi...

Zhiwen Ruan, Yan Yang, Zhuocheng Liang et al. · 0 citations
#natural language process... Preprint Sep 2026

ROAM: Robust Organization of Atomic Memories for Agents through Semantic Relations

Long-term language-model agents rely on external memory across interactions. Atomic memories are particularly useful: their fine-grained semantic boundaries enable precise retrieval and direct comparison between observations. Yet accumulating atoms inevitably become redundant, overlapping, or conflicting. Existing meth...

Jianjie Zheng, Peng Lai, Sijie Cheng et al. · 0 citations
#artificial intelligence Preprint Sep 2026

UniRRM: Unified Reasoning Reward Models Across Languages and Evaluation Paradigms

Reinforcement learning (RL) excels on tasks with verifiable rewards, but in open-ended tasks, the reliability of reward models remains a key challenge. Existing solutions either depend on costly proprietary LLM-as-a-Judge systems or opaque scalar reward models that lack interpretability. Recent works on generative rewa...

Peng Lai, Yi-Chao Du, Junchao Wu et al. · 1 citation
#artificial intelligence Preprint Sep 2026

AlignDiff: Exploiting Model-Intrinsic Information for Better Preference Data Selection

AlignDiff, a preference data filtering framework driven by intrinsic model signals, first identifies samples with clear preferences using both positive and inverse signals, then prioritizes the more challenging samples based on the average negative log-likelihood gap, encouraging the model to learn richer information f...

Peng Lai, He Zhu, Zhiwen Ruan et al. · 1 citation
Preprint Aug 2026

Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction

CAST (Concept-guided Artifact Suppression Tuning), an SAE-based framework for auditable clinical text classification, improves over its corresponding fine-tuned encoder baselines and remains competitive with strong LLM baselines, while producing a feature-level audit trail of the clinical concepts that support each pre...

Jin Mu, Guan-Hua Chen · 0 citations
Preprint Aug 2026

Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing

P-Bench is built, a benchmark comprising 425 open-ended, realistic hypothesis-testing tasks spanning economics, biology, and medicine and introduces Fisher-R1, an open-weight LLM agent trained for rigorous hypothesis testing using synthetic tasks and reinforcement learning.

Jia-Cheng Miao, Jin Mu, Guanhua Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.