Skip to content

Author

Juanzi Li

We have 8 of 50 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

SAEVerbalizer: Generating Explanations for Sparse Autoencoder Features via Representation Verbalization

Sparse autoencoders (SAEs) are proposed to extract numerous features from large language model (LLM) representations, yet explaining these features still relies primarily on external observation. This reliance leads to superficial explanations inferred from observed model behavior and computational inefficiency from co...

Weihang Meng, Hongzhu Guo, Yi Jing et al. · 0 citations
Preprint Aug 2026

SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs

Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) exhibit fundamentally different behaviors in enhancing multi-task reasoning for large language models (LLMs). Our preliminary experiments revealed a phenomenon: SFT suffers from severe task conflicts under multi-stage training, whereas RL enables stable coexi...

Kejian Zhu, Zhuo-Ran Jin, Shangqing Tu et al. · 0 citations
Book Open access Aug 2026

SurveyReview: A Reviewer-Aligned Benchmark for Survey Evaluators

This work proposes SurveyReview, a reviewer-aligned, multi-dimensional benchmark and dataset for survey evaluation, and develops a strong baseline evaluator that substantially improves alignment with human reviewers, providing a competitive reference for future research.

Yuheng Zhang, Yuanchun Wang, Fanjin Zhang et al. · 0 citations

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios?

RuVerBench is introduced, the first benchmark for assessing LaaJ reliability in rubric verification for agentic scenarios, and the impact of key LaaJ strategies, including prompt design, batching, and majority voting, on rubric verification is analyzed.

Yangda Peng, Yunjia Qi, Hao Peng et al. · 2 citations
Book Open access Aug 2026

SurveyReview: A Reviewer-Aligned Benchmark for Survey Evaluators

The rapid advancement of large language models has transformed survey writing from a months-long manual effort into an automated process. As generation scales, reliable evaluation becomes the bottleneck, and LLMs are increasingly used as survey evaluators. However, existing approaches largely rely on off-the-shelf LLM-...

Yuheng Zhang, Yuanchun Wang, Fanjin Zhang et al. · 0 citations

Where Steering Signals Come From: Activation Source Selection in Activation Steering

Tail subtraction is introduced, which removes shared prompt and continuation semantics from boundary states and yields cleaner, more stable steering signals, and suggests that steering depends on representations of what the model is about to do, not merely on what has already appeared.

Jiaran Ye, Lingxu Ran, Zijun Yao et al. · 2 citations
Preprint Aug 2026

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

TrajDebug is proposed, an error-lifecycle tracing framework that addresses long-trajectory error discovery with multi-granularity history compression and evidence-based error identification, and supports critical attribution by tracing each error's resolution status and terminal impact.

Yunjia Qi, Zehua Yin, Xin Shi et al. · 4 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.