Researcher profile
Tzu-Heng Huang
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
2025 · arXiv (Cornell University)
Large language models (LLMs) are widely used to evaluate the quality of LLM generations and responses, but this leads to significant challenges: high API costs, uncertain reliability, inflexible pipelines, and inherent biases. To address these, …