Xiaocheng Feng
8 papers in the PaperMetrix corpus
Papers by this author
-
Topic-to-Essay Generation with Neural Networks
2018
We focus on essay generation, which is a challenging task that generates a paragraph-level text with multiple topics.Progress towards understanding different topics and expressing diversity in this task requires more powerful generators and richer training …
-
MSAMSum: Towards Benchmarking Multi-lingual Dialogue Summarization
2022
Dialogue summarization helps users capture salient information from various types of dialogues has received much attention recently. However, current works mainly focus on English dialogue summarization, leaving other languages less well explored. Therefore, we present …
-
Trends in Integration of Knowledge and Large Language Models: A Survey and Taxonomy of Methods, Benchmarks, and Applications
2023 · arXiv (Cornell University)
Large language models (LLMs) exhibit superior performance on various natural language tasks, but they are susceptible to issues stemming from outdated data and domain-specific limitations. In order to address these challenges, researchers have pursued two …
-
Aligning Translation-Specific Understanding to General Understanding in Large Language Models
2024 · arXiv (Cornell University)
Large Language models (LLMs) have exhibited remarkable abilities in understanding complex texts, offering a promising path towards human-like translation performance. However, this study reveals the misalignment between the translation-specific understanding and the general understanding inside …
-
FroM: Frobenius Norm-Based Data-Free Adaptive Model Merging
2025 · arXiv (Cornell University)
With the development of large language models, fine-tuning has emerged as an effective method to enhance performance in specific scenarios by injecting domain-specific knowledge. In this context, model merging techniques provide a solution for fusing …
-
Improving Low Resource Named Entity Recognition using Cross-lingual Knowledge Transfer
2018
Neural networks have been widely used for high resource language (e.g. English) named entity recognition (NER) and have shown state-of-the-art results.However, for low resource languages, such as Dutch, Spanish, due to the limitation of resources …
-
CodeBERT: A Pre-Trained Model for Programming and Natural Languages
2020
Zhangyin Feng, Daya Guo, Duyu Tang, Nan Duan, Xiaocheng Feng, Ming Gong, Linjun Shou, Bing Qin, Ting Liu, Daxin Jiang, Ming Zhou. Findings of the Association for Computational Linguistics: EMNLP 2020. 2020.
-
Language Model as an Annotator: Exploring DialoGPT for Dialogue Summarization
2021
Xiachong Feng, Xiaocheng Feng, Libo Qin, Bing Qin, Ting Liu. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long …