ملف الباحث

Rafet Sifa

3 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Is Reinforcement Learning (Not) for Natural Language Processing: Benchmarks, Baselines, and Building Blocks for Natural Language Policy Optimization

    2022 · arXiv (Cornell University)

    We tackle the problem of aligning pre-trained large language models (LMs) with human preferences. If we view text generation as a sequential decision-making problem, reinforcement learning (RL) appears to be a natural conceptual framework. However, …

  2. KPI-EDGAR: A Novel Dataset and Accompanying Metric for Relation Extraction from Financial Documents

    2022

    We introduce KPI-EDGAR, a novel dataset for Joint Named Entity Recognition and Relation Extraction building on financial reports uploaded to the Electronic Data Gathering, Analysis, and Retrieval (EDGAR) system, where the main objective is to …

  3. Judging Quality Across Languages: A Multilingual Approach to Pretraining Data Filtering with Language Models

    2025 · arXiv (Cornell University)

    High-quality multilingual training data is essential for effectively pretraining large language models (LLMs). Yet, the availability of suitable open-source multilingual datasets remains limited. Existing state-of-the-art datasets mostly rely on heuristic filtering methods, restricting both their …