北京大学校徽SEKE研究组
登录

← 返回论文列表

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

Huanyu Liu, Jia Li, Chang Yu, Taozhi Chen, Yihong Dong, Lecheng Wang, Yongding Tao, Xiaolong Hu, Ge Li

[ACL 2026] Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (ACL 2026), San Diego, California, USA.

PDF

组内作者