To study this phenomenon, SocioHack is introduced, a sandbox of 72 societal environments, and it is found that within these environments, reward hacking naturally emerges and leads to regulatory loophole discovery.
Wei Liu, Xinyi Mou, Hanqi Yan et al.· arXiv.org· 3 citations
This work extends SocioVerse 1.0 into a human-AI co-evolutionary paradigm built from two loops and one infrastructure, and validates SocioVerse2 across three case families and seven case studies, from reproducing canonical agent-based models to modeling policy processes on real records and nowcasting macro-economic ind...
Xin-Nong Zhang, Jiayu Lin, Jia Wang et al.· 0 citations
Scientific papers may relate by problem, method, result, or contribution, but document-level retrievers collapse these into a single similarity score without saying why they are related. Citation- and similarity-based retrieval alone also confines search to the neighbourhood of what is already known, whereas generative...
Italo Luis da Silva, Hanqi Yan, Yujing Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.