Building the Great Recession News Corpus (GRNC): A contemporary diachronic corpus of economy news in English

Autor: Antonio Moreno-Ortiz, Javier Fernández-Cruz
Rok vydání: 2020
Předmět:
Zdroj: Research in Corpus Linguistics. 8:28-45
ISSN: 2243-4712
DOI: 10.32714/ricl.08.02.02
Popis: The paper describes the process involved in developing the Great Recession News Corpus (GRNC); a specialized web corpus, which contains a wide range of written texts obtained from the Business section of The Guardian and The New York Times between 2007 and 2015. The corpus was compiled as the main resource in a sentiment analysis project on the economic/financial domain. In this paper we describe its design, compilation criteria and methodological approach, as well as the description of the overall creation process. Although the corpus can be used for a variety of purposes, we include a sentiment analysis study on the evolution of the sentiment conveyed by the word credit during the years of the Great Recession which we think provides validation of the corpus.
Databáze: OpenAIRE