Skip to content

Author

Qi-Yong Zhong

We have 3 of 11 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

MAS-OPD: On-Policy Distillation for Multi-agent Systems

Multi-agent systems (MAS) split a task across specialized roles and are promising on complex tasks, yet a prevailing approach relies on inference-time orchestration alone. General-purpose APIs are costly and hard to customize, while small models with role prompts rarely develop stable role competence or reliable collab...

Qi-Yong Zhong, Mao Zheng, Ming-Yang Song et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Data-free On-policy Distillation

On-policy distillation (OPD) has become a standard component of frontier post-training pipelines, yet how much its training data actually contributes has gone largely unexamined. On the two teacher--student pairings most common in practice, we find OPD almost indifferent to its data: eight prompts already match a 17k-p...

Gengsheng Li, Mao Zheng, Ming-Yang Song et al. · 0 citations
Jul 2026

EasyOPD: An Easy-to-use On-Policy Distillation Framework for Large Language Models

Experiments on reasoning, code-generation, scientific-knowledge, scientific-knowledge, and tool-use benchmarks show that these implementations can be executed through the same verl-based backend while retaining their method-specific objectives and task-dependent performance profiles.

Jie Sun, Mao Zheng, Mingyang Song et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.