Scientific datasets, such as materials and molecular datasets, are often large, complex, and open-ended, posing a core challenge for data efficiency and model training. While data attribution (DA) offers a principled way to score and select samples for efficient learning, we identify a fundamental misalignment between...
Jianpeng Chen, Wangzhi Zhan, Haohui Wang et al.· Proceedings of the 32nd ACM...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.