Researcher profile

Yunzhe Tao

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Stochastic Training of Residual Networks: a Differential Equation\n Viewpoint

    2018 · arXiv (Cornell University)

    During the last few years, significant attention has been paid to the\nstochastic training of artificial neural networks, which is known as an\neffective regularization approach that helps improve the generalization\ncapability of trained models. In this work, …

  2. $\mathbf{(N,K)}$-Puzzle: A Cost-Efficient Testbed for Benchmarking Reinforcement Learning Algorithms in Generative Language Model

    2024 · arXiv (Cornell University)

    Recent advances in reinforcement learning (RL) algorithms aim to enhance the performance of language models at scale. Yet, there is a noticeable absence of a cost-effective and standardized testbed tailored to evaluating and comparing these …