ملف الباحث

Yang Ouyang

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Min-K%++: Improved Baseline for Detecting Pre-Training Data from Large Language Models

    2024 · arXiv (Cornell University)

    The problem of pre-training data detection for large language models (LLMs) has received growing attention due to its implications in critical issues like copyright violation and test data contamination. Despite improved performance, existing methods (including …

  2. Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning

    2025 · arXiv (Cornell University)

    Large language models (LLMs) demonstrate strong reasoning abilities via Chain-of-Thought (CoT), but their token-level generation encourages local decisions and lacks global planning, often leading to redundant or inaccurate reasoning. Existing methods, such as tree-based search …