Researcher profile

Guanzhi Wang

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search

    2024 · arXiv (Cornell University)

    Traditional reinforcement learning and planning typically requires vast amounts of data and training to develop effective policies. In contrast, large language models (LLMs) exhibit strong generalization and zero-shot capabilities, but struggle with tasks that require …