SPACE induces two-level programmatic skills from successful trajectories, where subskill boundaries serve as direct chunk-boundary supervision from trajectory-induced programmatic skills, and distilled into a primitive-chunk policy via hybrid on-/off-policy optimization.
Yan-Ting Yang, Can Jin, Jin-Man Zhao et al.· 1 citation
A cross-fidelity knowledge distillation and adaptive fusion network (CFKD-AFN), which leverages abundant but low-fidelity simulation data to enhance the prediction on scarce but high-fidelity trial data, and is extended to an interpretable variant for exploratory analysis of feature-attribution patterns associated with...
Wen-Jing Chen, Lian-Sheng Zhuang, Zi-Ying Luo et al.· arXiv.org· 0 citations
A novel FL framework is presented, FedPhoenix, that stochastically re-sets partial parameters in each round to destroy some features of the global model, guiding FL training to learn multiple generalized features for inference rather than specific overfitting features.
Jia-Hao Wu, Ming Hu, Yanxin Yang et al.· Neural Information Processin...· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.