Skip to content

4 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

2026

Action-Level Backdoor Attacks Against Deep Reinforcement Learning Systems via Adaptive Reward Exploration

Deep Reinforcement Learning (DRL) has demonstrated remarkable capabilities in domains such as robotics, finance, and autonomous systems. With the increasing cost of training, DRL models are increasingly shared and reused via model marketplaces, cloud platforms, and open-source repositories. This trend exposes DRL syste...

Ou-Bo Ma, L. Du, Yang Dai et al. · 0 citations
Preprint Sep 2026

The Shape of Ownership: Verifying LLM Provenance through Semantic Structures

As large language models (LLMs) are increasingly redistributed, adapted, and served behind opaque APIs, model ownership can no longer be established reliably by inspecting model internals or deployment records. This creates a need for behavioral signatures that remain observable through black-box interaction. Yet most...

Zhong-Rui Sun, Jia-Hao Chen, Ou-Bo Ma et al. · 0 citations
Preprint Sep 2026

A Finger on the Scale: Covert Policy Steering through Agentic Skills

Reusable agent skills extend large language model (LLM) agents with task procedures, tool-use guidance, and output constraints. Yet these skills also act as externalized behavioral policies, which create a supply-chain risk: a third-party skill may preserve the declared task and valid output interface while covertly re...

Jia-Rui Li, Jia-Hao Chen, Chun-Yi Zhou et al. · 0 citations
#artificial intelligence Book Feb 2024

SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learning Systems

This study unveils the capability of attackers to generate adversarial policies even when restricted to partial observations of the victims in multi-agent competitive environments, and proposes a novel black-box attack (SUB-PLAY) that incorporates the concept of constructing multiple subgames to mitigate the impact of...

Oubo Ma, Yuwen Pu, L. Du et al. · 16 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.