A sound judgment applies the law to established facts and weighs the circumstances in which they arose. However, existing methods swing between rigid statute matching and ungrounded discretion, benchmarks score a label or a rubric, and the experience that would supply the balance stays unverified. We formalize legal ju...
Zheng-Kai Tu, Ming-Da Zhang, Zi-Jia Wang et al.· 0 citations
Recursive self-improvement (RSI) lets a system improve from its own outcomes; in LLM-based multi-agent systems, Agents refine one another within a task, and outcomes improve how they collaborate across tasks. However, existing multi-agent collaboration leaves this loop open: collaboration is pre-defined at the operator...
Xiao Huang, Ming-Da Zhang, Jun-Ming Zhang et al.· 0 citations
R$^2$ Flow is introduced, a recursive self-improvement framework that alternates policy learning, independent verification, and versioned skill-library updates on a shared-state orchestration graph that improves task accuracy and library-edit precision over heuristic orchestration, reinforcement learning, and skill-evo...
Ming-Da Zhang, Qian-Shuo Huang, Yan-Jin Li et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.