Skip to content

Author

Jian-Hao Yan

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

DeepInstructor: An Agentic AI Instructor for Experience-Driven Idea Evaluation

As automated scientific discovery advances, Large Language Models (LLMs) can now generate research ideas at an unprecedented scale, shifting the bottleneck from idea generation to idea evaluation. Existing evaluators mainly rely on parametric LLM knowledge or unstructured retrieval, producing judgments that lack the ex...

Rong-Can Pei, Fang Guo, Qing-Lin Qi et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening

Value Flattening is identified as an important yet overlooked failure mode of critic learning in standard PPO and a simple sparse supervision strategy can mitigate it; SParse Proximal Policy Optimization is introduced, which applies the value loss to only a few well-separated states in each response to mitigate both ef...

Yi-Zhuo Li, Jian-Hao Yan, Yun Luo et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.