The rapid integration of Generative Artificial Intelligence (GenAI) and Large Language Models (LLMs) into educational environments has fundamentally transformed instructional design building upon earlier AI research in educational psychology. However, there is growing concern that computational innovation is outstrippi...
Shu-Jin Zhong, J. Simons· Frontiers in Psychology· 0 citations
Reinforcement learning enables large language model (LLM) agents to learn multi-step behaviors through interaction with their environments. However, rewards in many interactive tasks reflect only the final outcome, providing limited guidance on which intermediate decisions advance the task. Successful training trajecto...
Bo-Wen Zhang, Jun-Wei He, Mao-Qi Liu et al.· 0 citations
Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit criteria. Yet in GRPO-style pipelines, these structured judgments are reduced to a scalar response-level reward and converted into a response-level advantage, which is broadcast uniformly to all generated t...
Bo-wen Zhang, Junwei He, Wen Wang et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.