Skip to content

Author

Dongsheng Zhu

We have 3 of 9 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Review Sep 2026

Atria Dawn: The Dawn of Agentic Superintelligence

Atria Dawn Preview is introduced, a foundation agentic language model designed for scientific research and engineering workflows, with the goal of expanding the frontier of agent productivity in the real world.

Honglin Guo, Tao Gui, Kun Cai et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents

SWE-Bench Pro Verified offers a more trustworthy benchmark for assessing software engineering agents, which combines anti-hacking safeguards that eliminate major leakage channels without disrupting normal agent functionality, with task refinement that minimally corrects inconsistencies within flawed instances.

Pujun Zheng, Zi-Xin Shang, Shufan Jiang et al. · 1 citation
Jul 2026

AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities

AgentCompass is introduced, an open-source, lightweight, and extensible infrastructure for evaluating LLM-based agents that organizes the evaluation process around three independent components, thereby enabling flexible configurations without requiring the reimplementation of complex execution logic.

Zichen Ding, Jiaye Ge, Shufan Jiang et al. · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.