Identifying Informational Sources in News Articles

Alexander Spangher, Nanyun Peng, Emilio Ferrara, Jonathan May

Abstract

News articles are driven by the informational sources journalists use in reporting. Modeling when, how and why sources get used together in stories can help us better understand the information we consume and even help journalists with the task of producing it. In this work, we take steps toward this goal by constructing the largest and widest-ranging annotated dataset, to date, of informational sources used in news writing. We first show that our dataset can be used to train high-performing models for information detection and source attribution. Then, we introduce a novel task, source prediction, to study the compositionality of sources in news articles – i.e. how they are chosen to complement each other. We show good modeling performance on this task, indicating that there is a pattern to the way different sources are used together in news storytelling. This insight opens the door for a focus on sources in narrative science (i.e. planning-based language generation) and computational journalism (i.e. a source-recommendation system to aid journalists writing stories). All data and model code can be found at https://github.com/alex2awesome/source-exploration.

Anthology ID:: 2023.emnlp-main.221
Volume:: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing
Month:: December
Year:: 2023
Address:: Singapore
Editors:: Houda Bouamor, Juan Pino, Kalika Bali
Venue:: EMNLP
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 3626–3639
Language:
URL:: https://aclanthology.org/2023.emnlp-main.221
DOI:: 10.18653/v1/2023.emnlp-main.221
Bibkey:
Cite (ACL):: Alexander Spangher, Nanyun Peng, Emilio Ferrara, and Jonathan May. 2023. Identifying Informational Sources in News Articles. In Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 3626–3639, Singapore. Association for Computational Linguistics.
Cite (Informal):: Identifying Informational Sources in News Articles (Spangher et al., EMNLP 2023)
Copy Citation:
PDF:: https://aclanthology.org/2023.emnlp-main.221.pdf
Video:: https://aclanthology.org/2023.emnlp-main.221.mp4

PDF Cite Search Video