Skip to content

Author

Mu-Hao Chen

We have 3 of 22 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Targeting Pivotal Decisions for Credit Assignment in Agentic Reinforcement Learning

Group Relative Policy Optimization (GRPO) has become a promising approach for training large language model agents. However, its uniform assignment of trajectory-level advantages to all policy tokens fails to distinguish consequential decisions from less relevant ones, obscuring which intermediate decisions contributed...

Dongwon Jung, H. Ramesh, Yi-Fan Wang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Despite Instructions: Frontier Agents Improvise Covert Channels at Test Time

In security-sensitive applications, language-model agents are often required to coordinate without disclosing confidential information. Yet repeated interactions may also let ordinary messages acquire shared private meaning. We study a repeated game with pairs of models in which the sender model observes one of four se...

Jacob Dineen, Si-Lei Ren, Mu-Hao Chen et al. · 0 citations
Jun 2026

Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens

SafeClawArena is developed, a benchmark of 406 adversarial tasks executed in containerized replicas of real agent platforms with canary-marked credentials and evaluated via automated taint tracking across nine output channels, exposing the inadequacy of current defenses and suggesting directions for future hardening.

Peizhi Niu, Wenjie Qu, Shangding Gu et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.