Skip to content

Author

Xipeng Qiu

We have 7 of 12 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#computer vision Preprint Sep 2026

VehicleArena: A Realistic Urban Environment for Multi-Agent Driving

Real-world embodied agents often pursue independent objectives within a shared physical environment, where their actions can alter the conditions faced by others. Existing benchmarks, however, typically assume shared goals or explicitly prescribed interaction protocols, leaving such emergent physical coupling underexpl...

Jie Yang, Jia-Jun Chen, Jia-Zheng Zhou et al. · 0 citations
#machine learning Preprint Sep 2026

MARCO: Multi-Round Agentic Reinforcement for Conditional Molecular Optimization

Molecular optimization is inherently iterative: a candidate is proposed, evaluated against several objectives, and revised while preserving a relationship to the source molecule. Most instruction-following models instead emit one edited molecule, forcing validity, property improvement, and similarity control into a sin...

Shi-Cheng Fang, Yu-Xin Wang, Zhuo Yang et al. · 0 citations
#machine learning Preprint Sep 2026

ORPG: Reconciling Multiple Reward Objectives through Objective-wise Policy Gradients

Objective-wise Reconciled Policy Gradient (ORPG), which constructs a separate clipped policy objective for each reward and reconciles the resulting gradients into one policy update, achieves the highest average full-budget accuracy and three-budget hypervolume among the compared methods.

Shi-Cheng Fang, Yi-Wen Zhao, Wen-Bo Tian et al. · 0 citations
Preprint Aug 2026

SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science?

A paired ablation that removes explicit scientific guidance while preserving the repository and executable engineering context shows that scientific knowledge is not uniformly beneficial: well-grounded information can constrain repair and improve average performance and token efficiency, whereas poorly aligned guidance...

Zhi-Peng Xu, Jia-Hao Lu, Yi-Ning Zheng et al. · 4 citations
Preprint Aug 2026

ETA: A New Agentic Paradigm for Embodied Tasks

The Embodied Task Agent is introduced, a new paradigm for extending digital agents into the physical world, and OpenETA is released as its open-source implementation, which provides replaceable Planners, composable Tools and Skills, auditable memory, replayable trajectories, and common interfaces for simulation and rea...

Yi-Tong Chen, Zezheng Huai, Si-Xian Li et al. · 5 citations · ⚡1
#machine learning Preprint Jul 2026

EvoCUA-1.5: Online Reinforcement Learning for Multi-turn Computer-Use Agents

EvoCUA-1.5 extends self-evolving computer-use agents from offline experience learning to online reinforcement learning, where policies interact with executable sandbox environments and improve from verifiable task outcomes and provides a practical framework for scaling online RL in multi-turn computer-use agents.

Mianqiu Huang, Taofeng Xue, Chong Peng et al. · 1 citation
Preprint Jul 2026

Rethinking Scientific Discovery in the Agentic Era

Applications in materials analysis, molecule design, and protein or antibody screening, together with experiments on scientific reading, idea generation, molecule generation, and antibody screening, show that SCION outperforms existing autonomous research-agent baselines, especially in decomposition, verification, refi...

Y. Zheng, Yuxin Wang, Jiahao Lu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.