Recommending Missed Citations Identified by Reviewers: A New Task, Dataset and Baselines

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Long, Kehan, Li, Shasha, Wang, Pancheng, Bao, Chenlong, Tang, Jintao, Wang, Ting
Format: Preprint
Publié: 2024
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866909127053148160
author Long, Kehan
Li, Shasha
Wang, Pancheng
Bao, Chenlong
Tang, Jintao
Wang, Ting
author_facet Long, Kehan
Li, Shasha
Wang, Pancheng
Bao, Chenlong
Tang, Jintao
Wang, Ting
contents Citing comprehensively and appropriately has become a challenging task with the explosive growth of scientific publications. Current citation recommendation systems aim to recommend a list of scientific papers for a given text context or a draft paper. However, none of the existing work focuses on already included citations of full papers, which are imperfect and still have much room for improvement. In the scenario of peer reviewing, it is a common phenomenon that submissions are identified as missing vital citations by reviewers. This may lead to a negative impact on the credibility and validity of the research presented. To help improve citations of full papers, we first define a novel task of Recommending Missed Citations Identified by Reviewers (RMC) and construct a corresponding expert-labeled dataset called CitationR. We conduct an extensive evaluation of several state-of-the-art methods on CitationR. Furthermore, we propose a new framework RMCNet with an Attentive Reference Encoder module mining the relevance between papers, already-made citations, and missed citations. Empirical results prove that RMC is challenging, with the proposed architecture outperforming previous methods in all metrics. We release our dataset and benchmark models to motivate future research on this challenging new task.
format Preprint
id arxiv_https___arxiv_org_abs_2403_01873
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Recommending Missed Citations Identified by Reviewers: A New Task, Dataset and Baselines
Long, Kehan
Li, Shasha
Wang, Pancheng
Bao, Chenlong
Tang, Jintao
Wang, Ting
Information Retrieval
Citing comprehensively and appropriately has become a challenging task with the explosive growth of scientific publications. Current citation recommendation systems aim to recommend a list of scientific papers for a given text context or a draft paper. However, none of the existing work focuses on already included citations of full papers, which are imperfect and still have much room for improvement. In the scenario of peer reviewing, it is a common phenomenon that submissions are identified as missing vital citations by reviewers. This may lead to a negative impact on the credibility and validity of the research presented. To help improve citations of full papers, we first define a novel task of Recommending Missed Citations Identified by Reviewers (RMC) and construct a corresponding expert-labeled dataset called CitationR. We conduct an extensive evaluation of several state-of-the-art methods on CitationR. Furthermore, we propose a new framework RMCNet with an Attentive Reference Encoder module mining the relevance between papers, already-made citations, and missed citations. Empirical results prove that RMC is challenging, with the proposed architecture outperforming previous methods in all metrics. We release our dataset and benchmark models to motivate future research on this challenging new task.
title Recommending Missed Citations Identified by Reviewers: A New Task, Dataset and Baselines
topic Information Retrieval
url https://arxiv.org/abs/2403.01873