Skip to content

Author

Zijian Wang

We have 3 of 53 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

E2E-SWE: Benchmarking LLMs on Building Working Codebases from Scratch

Coding agents powered by large language models (LLMs) are evolving from making localized code changes to developing complete software repositories. However, evaluating repository-scale generation remains challenging: tasks must demand system-level reasoning while ensuring that all evaluated behaviors are precisely spec...

Hantian Ding, Chloe Bi, Jia-Cheng Zhu et al. · 0 citations

Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking Attacks

JAWS-BENCH(Jailbreaks Across WorkSpaces), a benchmark spanning three escalating workspace regimes mirroring attacker capability, is presented, indicating that JAWS-BENCH can be reused across multiple agent frameworks.

Shoumik Saha, Jifan Chen, Sam Mayers et al. · 8 citations · ⚡2

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.