Language varies not only between countries, but also along regional and sociodemographic lines. This variation is one of the driving factors behind language change. However, investigating language variation is a complex undertaking: the more factors we want to consider, the more data we need. Traditional qualitative methods are not well-suited to do this, an thus restricted to isolated factors. This reduction limits the potential insights, and risks attributing undue importance to easily observed factors. While there is a large interest in linguistics to improve upon such studies, it requires training in both variational linguistics and computational methods, a combination that is still not common. We take a first step here to alleviate the problem by providing an interface to explore large-scale language variation along several socio-demographic factors without programming knowledge. It makes use of large amounts of data and provides statistical analyses, maps, and interactive features that enable scholars to explore language variation in a data-driven way.
No takes yet. Share an insight, caveat, or question.
Hovy et al. (2016) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: