Goal-conditioned reinforcement learning (GCRL) relies heavily on how target goals are represented to the policy. While recent methods encode goals via temporal distance, occupancy, or controllability, it remains unclear how much downstream performance actually depends on representation quality. We study this in offline...
Syed Nazmus Sakib, Abdul Monaf Chowdhury, Nafiul Haque et al.· 0 citations
Reinforcement learning with verifiable rewards (RLVR) has become an important approach for improving reasoning during post-training. Recent work suggests that some difficult prompts remain resistant to learning even when they occasionally produce correct solutions. We revisit this unlearnability phenomenon and find tha...
The role of artificial intelligence (AI) in project evaluation, capital allocation, and sustainability performance has gained attention, yet the firm-level channel through green investment remains underexplored. This study examines whether AI-driven decision capability increases green investment intensity and whether s...
M. Qamruzzaman, A. Almulhim, Syed Nazmus Sakib et al.· Journal of Intelligent Decis...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.