Building the Great Recession News Corpus (GRNC): A contemporary diachronic corpus of economy news in English

The paper describes the process involved in developing the Great Recession News Corpus (GRNC); a specialized web corpus, which contains a wide range of written texts obtained from the Business section of The Guardian and The New York Times between 2007 and 2015. The corpus was compiled as the main r...

Full description

Saved in:
Bibliographic Details
Published inResearch in corpus linguistics Vol. 8; no. 2; pp. 28 - 45
Main Authors Fernández-Cruz, Javier, Moreno-Ortiz, Antonio
Format Journal Article
LanguageEnglish
Published 2020
Online AccessGet full text
ISSN2243-4712
2243-4712
DOI10.32714/ricl.08.02.02

Cover

More Information
Summary:The paper describes the process involved in developing the Great Recession News Corpus (GRNC); a specialized web corpus, which contains a wide range of written texts obtained from the Business section of The Guardian and The New York Times between 2007 and 2015. The corpus was compiled as the main resource in a sentiment analysis project on the economic/financial domain. In this paper we describe its design, compilation criteria and methodological approach, as well as the description of the overall creation process. Although the corpus can be used for a variety of purposes, we include a sentiment analysis study on the evolution of the sentiment conveyed by the word credit during the years of the Great Recession which we think provides validation of the corpus.
ISSN:2243-4712
2243-4712
DOI:10.32714/ricl.08.02.02