Skip to content

Author

Chun-Yi Zhou

We have 6 of 40 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

2026

Action-Level Backdoor Attacks Against Deep Reinforcement Learning Systems via Adaptive Reward Exploration

Deep Reinforcement Learning (DRL) has demonstrated remarkable capabilities in domains such as robotics, finance, and autonomous systems. With the increasing cost of training, DRL models are increasingly shared and reused via model marketplaces, cloud platforms, and open-source repositories. This trend exposes DRL syste...

Ou-Bo Ma, L. Du, Yang Dai et al. · 0 citations
Review Open access Aug 2026

A Comparative Survey of Security Risks in AI Systems: From LLMs to AI Agents and Embodied Agents

Rapid AI development across industries raises pressing security and privacy risks. This work presents a unified comparison of large language models, AI agents, and embodied agents, introducing a taxonomy of risks spanning data, models, systems, content, and applications, alongside a catalog of 24 specific threats. We c...

Baiqi Wu, Qing-Ming Li, Chun-Yi Zhou et al. · 0 citations
2026

PREFed: An Effective and Stealthy Static-Anchor Backdoor Attack via Trigger Pre-Optimization in Federated Learning

Existing Federated Learning (FL) backdoor attacks commonly employ round-wise proximity strategies, dynamically adapting malicious updates to mimic benign ones in order to evade detection. However, such adaptive mechanisms often introduce instability, increase computational overhead, and create temporal patterns that ma...

Xi Chen, Rui Zeng, Chun-Yi Zhou et al. · 0 citations
Preprint Sep 2026

A Finger on the Scale: Covert Policy Steering through Agentic Skills

Reusable agent skills extend large language model (LLM) agents with task procedures, tool-use guidance, and output constraints. Yet these skills also act as externalized behavioral policies, which create a supply-chain risk: a third-party skill may preserve the declared task and valid output interface while covertly re...

Jia-Rui Li, Jia-Hao Chen, Chun-Yi Zhou et al. · 0 citations
Jul 2026

Lilith: Backdoor Generalization under Training-Inference Trigger Shift

This work forms this problem as backdoor generalization under training--inference trigger shift and introduces Lilith, a black-box anchor-to-family framework that achieves high family-wise attack success with limited utility degradation and a small trigger generalization gap.

Zhou Feng, Jia-Hao Chen, Chun-Yi Zhou et al. · 0 citations
Book Open access Aug 2026

The Boy Who Cried Wolf: Adversarial Misclassification of Safe Inputs as Unsafe in Multimodal Guardrails

Unsafe Semantic Distillation is proposed, which aligns adversarial perturbations with distributional representations of unsafe content rather than prompt-specific instances, and achieves 84% attack success rates, outperforming existing methods and exposing fundamental vulnerabilities in current multimodal safety archit...

Shuo Shi, Ruiping Yin, Na-En Xu et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.