Researcher profile

Yuchen Zhang

6 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Macro Grammars and Holistic Triggering for Efficient Semantic Parsing

    2017 · arXiv (Cornell University)

    To learn a semantic parser from denotations, a learning algorithm must search over a combinatorially large space of logical forms for ones consistent with the annotated denotations. We propose a new online learning algorithm that …

  2. $\ell_1$-regularized Neural Networks are Improperly Learnable in Polynomial Time

    2015 · arXiv (Cornell University)

    We study the improper learning of multi-layer neural networks. Suppose that the neural network to be learned has $k$ hidden layers and that the $\ell_1$-norm of the incoming weights of any neuron is bounded by …

  3. The Influence of Data Pre-processing and Post-processing on Long Document Summarization

    2021 · arXiv (Cornell University)

    Long document summarization is an important and hard task in the field of natural language processing. A good performance of the long document summarization reveals the model has a decent understanding of the human language. …

  4. DDIN: Deep Disentangled Interest Network for Click-Through Rate Prediction

    2023

    Click-Through Rate(CTR) prediction aims to predict the possibility of users clicking on products, which has become the core task of advertising recommendation systems. Due to the richness of user historical behavior, a key to making …

  5. HiPhO: How Far Are (M)LLMs from Humans in the Latest High School Physics Olympiad Benchmark?

    2025 · arXiv (Cornell University)

    Recently, the physical capabilities of (M)LLMs have garnered increasing attention. However, existing benchmarks for physics suffer from two major gaps: they neither provide systematic and up-to-date coverage of real-world physics competitions such as physics Olympiads, …

  6. Llama 2: Open Foundation and Fine-Tuned Chat Models

    2023 · arXiv (Cornell University)

    In this work, we develop and release Llama 2, a collection of pretrained and fine-tuned large language models (LLMs) ranging in scale from 7 billion to 70 billion parameters. Our fine-tuned LLMs, called Llama 2-Chat, …