With the rapid evolution of Large Language Models (LLMs), automated software testing is witnessing a paradigm shift. While proprietary models like GPT-4o demonstrate impressive capabilities, their high deployment costs and data privacy concerns make open-source LLMs the practical imperative for many academic and indust...
Guo-Qing Wang, Cheng-Ran Yang, Xiao-Xuan Zhou et al.· Proceedings of the ACM on so...· 0 citations
A typed domain-specific language that captures recurring tensor structures, such as repeated regions and floating-point fields, through a set of reversible operators, is designed, which formulates lossless tensor compression as program synthesis.
Jie-Ke Shi, Jun-Da He, Wenjia Jiang et al.· 0 citations
Accurate vulnerability severity assessment is essential for prioritizing remediation, yet manually assessing Common Vulnerability Scoring System (CVSS) base metrics remains labor-intensive. Existing automated approaches often fail to capture the repository-level evidence required for assessing many CVSS base metrics. S...
Jinfeng Jiang, Yikun Li, Chengran Yang et al.· 0 citations
CyberFactory is introduced, a unified open-source framework that connects data construction, trajectory synthesis, and model training across proof-of-concept (PoC) generation, vulnerability patching, and cybersecurity question answering (CyberQA).
Jian Yang, Haau-Sing Li, Shawn Guo et al.· 0 citations
This paper proposes AgentExecutor, a novel multi-agent framework for partial code execution that is Supported by the power of LLM agents who can think, act, and get feedback iteratively, and is able to autonomously explore a richer action space, enabling diverse operations such as creating resource files and resolving...
Junkai Chen, Cheng-Ran Yang, Xing Hu et al.· 0 citations
The first systematic study of model editing as a model-level hardening mechanism for secure code generation is conducted, evaluating 3 state-of-the-art editing methods across diverse LLM families and comparing them with CoSec, a representative inference-time approach, focusing on security, robustness, generalization, a...
Wei-Feng Sun, Quan-Jun Zhang, Yuchen Chen et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.