Publications

Publication details [#59387]

Davies, Mark. 2014. Making Google Books n-grams useful for a wide range of research on language change. International Journal of Corpus Linguistics 19 (3) : 401–416.
Publication type
Article in journal
Publication language
English
Language as a subject
Place, Publisher
John Benjamins
Journal DOI
10.1075/ijcl

Annotation

The “standard” Google Books n-grams were released by Google in 2010, and they include more than 155 billion words of data for the American English data alone. Unfortunately, the standard interface is far too simplistic to allow many types of useful research on this massive dataset. This paper discusses an alternative “advanced” architecture and interface for these datasets, which is freely available at googlebooks.byu.edu. This resource allows for a wide range of research on lexical, phraseological, syntactic, and semantic changes in English, in ways that would not be possible with the standard interface. With this new resource, researchers now have access to hundreds of billions of words of data, and can map out changes in English in ways that were not previously possible.