Skip to content

Author

Hai-Bo Shi

We have 5 of 17 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

AdaTutoRank: Learning to Rerank Document Sets via Adaptive Tutoring Optimization for RAG and Deep Research

AdaTutoRank is proposed, a setwise reranker trained with Adaptive Tutoring Optimization under a three-level hierarchy of nine rubric dimensions, which supplies silver labels for the cold start, rewards for reinforcement learning, and hints for distillation.

Kai-Lin Jiang, Lei Liu, Jian-Fei Xi et al. · 0 citations
Jul 2026

Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking

Rubric4Setwise is proposed, a training-free method that converts rubric-based evaluation criteria into document set selection signals, achieving the best downstream generation performance with fewer documents and search rounds, validating the effectiveness of closing the loop from evaluation to optimization.

Kai-Lin Jiang, Lei Liu, Jian-Fei Xi et al. · 3 citations
#natural language process... Preprint Jul 2026

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing external rewards provide either sparse outcome supervision or richer feedback from process annotations and LLM judges. Outcome rewards scale readily but cannot distinguish g...

Ruoxi Cheng, Haoxuan Ma, Hongyi Zhang et al. · 0 citations
Jul 2026

OPOD: On-Policy Omni Distillation

On-Policy Omni Distillation (OPOD), which consolidates text, image, and audio teachers into one omni model, and surpasses the base model and pooled RL training on all twelve benchmarks, and ranks first or second on eleven even when the teachers are included.

Tong Zhao, Yuyang Hu, Reed Li et al. · 0 citations
Preprint Aug 2026

ForestBench: A Unified Graph Framework for Evaluating Multi-Agent Collaboration

This work introduces a generalizable evaluation framework that maps native MAS traces into a shared space of unified collaboration graphs, enabling different methods to be evaluated under the same representation, reference set, and metric panel.

Guo Chen, Ziwen Li, Reed Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.