ملف الباحث
Yunzhe Tao
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Stochastic Training of Residual Networks: a Differential Equation\n Viewpoint
2018 · arXiv (Cornell University)
During the last few years, significant attention has been paid to the\nstochastic training of artificial neural networks, which is known as an\neffective regularization approach that helps improve the generalization\ncapability of trained models. In this work, …
-
$\mathbf{(N,K)}$-Puzzle: A Cost-Efficient Testbed for Benchmarking Reinforcement Learning Algorithms in Generative Language Model
2024 · arXiv (Cornell University)
Recent advances in reinforcement learning (RL) algorithms aim to enhance the performance of language models at scale. Yet, there is a noticeable absence of a cost-effective and standardized testbed tailored to evaluating and comparing these …