Researcher profile

Mark Klibanov

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference

    2024 · arXiv (Cornell University)

    Large language models (LLMs) have triggered a new stream of research focusing on compressing the context length to reduce the computational cost while ensuring the retention of helpful information for LLMs to answer the given …