Evaluating how LLM agents recover from mid-task failures is central to deploying reliable agentic systems. Existing checkpoint-based benchmarks measure recovery by comparing which action is selected as best across independent runs, a quantity known as set agreement. However, set agreement is a purely ordinal measure th...
Dong Xu, Zhang-Fan Yang, Jian-Tao Wu et al.· 0 citations
DegradeQuery, a context-aware prediction framework that converts label-missing records into a pretraining signal, is introduced and demonstrates that incompletely labeled PROTAC databases contain useful relational supervision and provide a practical route for learning context-aware degradation predictors from scarce ex...
Dong Xu, Zhang-Fan Yang, Jian-Tao Wu et al.· 0 citations
Structure-based drug design (SBDD) models are central to modern pharmaceutical research, enabling the rational exploration of protein-ligand interactions at atomic resolution. However, most existing approaches frame molecular generation as an isolated optimization or a one-to-one matching task, overlooking the shared b...
Dong Xu, Zhang-Fan Yang, Junchuang Cai et al.· IEEE transactions on computa...· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.