ملف الباحث

K. Wang

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Learn to Walk with Continuous-action for Knowledge-enhanced Recommendation System

    2024

    Knowledge graphs are more widely utilized to enhance recommendability and explainability. Reinforcement learning agents built to wander around the knowledge graph have been successfully applied in recommendation systems in a form of multi-hop relation reasoning. …

  2. LoRC: Low-Rank Compression for LLMs KV Cache with a Progressive Compression Strategy

    2024 · arXiv (Cornell University)

    The Key-Value (KV) cache is a crucial component in serving transformer-based autoregressive large language models (LLMs), enabling faster inference by storing previously computed KV vectors. However, its memory consumption scales linearly with sequence length and …