Pasar al contenido principal

A Multi-level Annotated Corpus of Scientific Papers for Scientific Document Summarization and Cross-document Relation Discovery

Tipo
Paper de conferencia
Año
2020
Lugar publicado
Marseille, France
Publisher
European Language Resources Association
Páginas
6672
Abstract

Related work sections or literature reviews are an essential part of every scientific article being crucial for paper reviewing and assessment. The automatic generation of related work sections can be considered an instance of the multi-document summarization problem. In order to allow the study of this specific problem, we have developed a manually annotated, machine readable data-set of related work sections, cited papers (e.g. references) and sentences, together with an additional layer of papers citing the references. We additionally present experiments on the identification of cited sentences, using as input citation contexts. The corpus alongside the gold standard are made available for use by the scientific community.

Autores

Luis Chiruzzo
Ahmed AbuRaéd
Horacio Saggion
Citekey
aburaed-saggion-chiruzzo:2020:LREC
Keywords