Skip to content

Author

Jiaxin Ding

We have 2 of 47 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

Beyond the Mean: Multi-Moment Policy Optimization for LLM Reasoning

A moment-based perspective on policy optimization for LLM reasoning is introduced by treating the failure probability of a randomly sampled problem as a random variable and characterizing optimization objectives through its moments, leading to MMPO, a novel policy optimization framework that jointly minimizes multiple...

Yijun Zhang, Yule Xie, Jia-Xin Ding et al. · 0 citations
Jul 2026

Information Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM Agents

Information Gain-based Rollout Policy Optimization (IGRPO) is proposed, a policy optimization framework that treats intermediate-state informativeness as the organizing principle of rollout collection and performs budget-aware tree-structured rollouts by allocating expansion budget according to node-level informativene...

Yijun Zhang, Fan Xu, Jiaxin Ding et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.