Researcher profile

Anirudh Atmakuru

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Does quantization affect models' performance on long-context tasks?

    2025 · arXiv (Cornell University)

    Large language models (LLMs) now support context windows exceeding 128K tokens, but this comes with significant memory requirements and high inference latency. Quantization can mitigate these costs, but may degrade performance. In this work, we …