Skip to content

Author

Tian-Wei Zhang

We have 8 of 369 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Sep 2026

Practical Secrets Extraction against Black-box LLMs

Large language models (LLMs) increasingly power autonomous coding agents such as Codex and Claude Code, yet their training corpora may contain confidential credentials exposed in public repositories or collected from private development artifacts, creating risks of memorization and subsequent leakage. Existing extracti...

Shi-Qian Zhao, Si-Wei Jiang, Xin-Feng Li et al. · 0 citations
Preprint Sep 2026

RAPID: A Real-Time Defense Against Unauthorized Model Distillation for Text-to-Image Services

Diffusion-based text-to-image (T2I) models are increasingly used for visual content creation, making their generation capability a valuable intellectual property asset. However, this capability is vulnerable to black-box output-based distillation, where an adversary queries the service, collects prompt-image pairs, and...

Zi-Han Wang, Bo-Heng Li, Rui Zhang et al. · 0 citations
Preprint Aug 2026

ECLIPSE: Self-Evolving Stealthy Prompt Injection Attack against Long-Horizon Agentic Systems

Recently, large language model (LLM) agents, such as Codex, Claude Code, and OpenClaw, have become capable of planning and executing long-horizon tasks through repeated tool calls. This capability also creates new opportunities for prompt injection. Existing attacks either place the malicious objective in one explicit...

Shiqian Zhao, Yang-Fan Zhou, Xin-Feng Li et al. · 0 citations
Preprint Aug 2026

ASCon: A Direction-Aware Reciprocal Agent--Step Contextualization Model for Failure Attribution in Multi-Agent Systems

ASCon is proposed, a direction-aware reciprocal \textbf{A}gent--\textbf{S}tep \textbf{Con}textualization model for multiple failure attribution targets that introduces direction-aware graph attention to model execution context, masked step-to-agent attention to construct behavior-aware agent representations, and agent-...

Shuyu Jiang, Yue Ran, Kaiyu Xu et al. · 0 citations
Preprint Aug 2026

The Model's Tell: Measuring Context-Leakage Attack Signals with Behavior Gauges

This paper proposes LeakGauge, which probes this response by appending a suffix that gauges leakage behavior and mapping its prefill token probabilities to an attack-risk score, and shows that the risk score is sensitive to an internal leakage-related direction.

Maosen Zhang, Jianshuo Dong, Bo-Han Lu et al. · 1 citation
#artificial intelligence Review Aug 2026

Auditing Chinese Web-scale Corpora via Sampled BPE Token Statistics

This work proposes Sampled-BPE, a lightweight token-level auditing pipeline that sample a small subset and train BPE tokenizer to surface polluted tokens, and releases a hierarchical Chinese web token dataset with 660k+ token records, organized as trees to support review and tracing of pollution.

Qingjie Zhang, Ziqi Tang, Jie Zhang et al. · 0 citations
Preprint Aug 2026

Your Agentic LLMs Secretly Encode Indirect Prompt-Injection Exposure in Hidden States

A probe-gated reasoning-based defense is introduced to bridge a knowledge-action gap in agentic LLMs when they are exposed to IPI attacks, and an analysis framework is introduced that identifies natural-language explanations strongly correlated with probe-captured signals.

Jianshuo Dong, Yiming Liu, Maosen Zhang et al. · 1 citation · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.