ملف الباحث
Lan, Tian
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Multi-agent Deep Covering Skill Discovery
2022 · arXiv (Cornell University)
The use of skills (a.k.a., options) can greatly accelerate exploration in reinforcement learning, especially when only sparse reward signals are available. While option discovery methods have been proposed for individual agents, in multi-agent reinforcement learning …
-
Momentum Decoding: Open-ended Text Generation As Graph Exploration
2022 · arXiv (Cornell University)
Open-ended text generation with autoregressive language models (LMs) is one of the core tasks in natural language processing. However, maximization-based decoding methods (e.g., greedy/beam search) often lead to the degeneration problem, i.e., the generated text …