Skip to content

Author

Ze-Run Ma

We have 4 of 18 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Review Sep 2026

Atria Dawn: The Dawn of Agentic Superintelligence

As AI agents become participants in the development of their successors, they reshape both the production of intelligence and the role of human researchers. We introduce Atria Dawn Preview, a foundation agentic language model designed for scientific research and engineering workflows, with the goal of expanding the fro...

Honglin Guo, Tao Gui, Kun Cai et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents

SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering agents on challenging repository-level tasks. However, our analysis work show that its evaluation is undermined by two sources of unreliability: reward hacking, enabled by leakage of gold solutions or hidden evaluation information, an...

Pujun Zheng, Zi-Xin Shang, Shufan Jiang et al. · 1 citation
Jul 2026

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

AgentCompass is introduced, an open-source, lightweight, and extensible infrastructure for evaluating LLM-based agents that organizes the evaluation process around three independent components, thereby enabling flexible configurations without requiring the reimplementation of complex execution logic.

Zichen Ding, Jiaye Ge, Shufan Jiang et al. · 2 citations
Review Aug 2026

Intern-S2-Preview: Scientific Agentic Foundation Model

Evaluations across scientific, multimodal, agentic, and general-purpose benchmarks show that Intern-S2-Preview-397B achieves competitive or leading results in multiple settings.

Lei Bai, Jiaqi Cao, Chiyu Chen et al. · 3 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.