Skip to content

Author

Tian-Wei Zhang

We have 7 of 369 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

RAPID: A Real-Time Defense Against Unauthorized Model Distillation for Text-to-Image Services

Diffusion-based text-to-image (T2I) models are increasingly used for visual content creation, making their generation capability a valuable intellectual property asset. However, this capability is vulnerable to black-box output-based distillation, where an adversary queries the service, collects prompt-image pairs, and...

Zi-Han Wang, Bo-Heng Li, Rui Zhang et al. · 0 citations
Preprint Aug 2026

ECLIPSE: Self-Evolving Stealthy Prompt Injection Attack against Long-Horizon Agentic Systems

Recently, large language model (LLM) agents, such as Codex, Claude Code, and OpenClaw, have become capable of planning and executing long-horizon tasks through repeated tool calls. This capability also creates new opportunities for prompt injection. Existing attacks either place the malicious objective in one explicit...

Shiqian Zhao, Yang-Fan Zhou, Xin-Feng Li et al. · 0 citations
Preprint Aug 2026

ASCon: A Direction-Aware Reciprocal Agent--Step Contextualization Model for Failure Attribution in Multi-Agent Systems

ASCon is proposed, a direction-aware reciprocal \textbf{A}gent--\textbf{S}tep \textbf{Con}textualization model for multiple failure attribution targets that introduces direction-aware graph attention to model execution context, masked step-to-agent attention to construct behavior-aware agent representations, and agent-...

Shuyu Jiang, Yue Ran, Kaiyu Xu et al. · 0 citations
Preprint Aug 2026

The Model's Tell: Measuring Context-Leakage Attack Signals with Behavior Gauges

This paper proposes LeakGauge, which probes this response by appending a suffix that gauges leakage behavior and mapping its prefill token probabilities to an attack-risk score, and shows that the risk score is sensitive to an internal leakage-related direction.

Maosen Zhang, Jianshuo Dong, Bo-Han Lu et al. · 1 citation
#artificial intelligence Review Aug 2026

Auditing Chinese Web-scale Corpora via Sampled BPE Token Statistics

This work proposes Sampled-BPE, a lightweight token-level auditing pipeline that sample a small subset and train BPE tokenizer to surface polluted tokens, and releases a hierarchical Chinese web token dataset with 660k+ token records, organized as trees to support review and tracing of pollution.

Qingjie Zhang, Ziqi Tang, Jie Zhang et al. · 0 citations
Preprint Aug 2026

Your Agentic LLMs Secretly Encode Indirect Prompt-Injection Exposure in Hidden States

A probe-gated reasoning-based defense is introduced to bridge a knowledge-action gap in agentic LLMs when they are exposed to IPI attacks, and an analysis framework is introduced that identifies natural-language explanations strongly correlated with probe-captured signals.

Jianshuo Dong, Yiming Liu, Maosen Zhang et al. · 1 citation · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.