PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 24, 2017Social Science Computer Review45 citations

Nonprobability Sampling and Twitter

View Full Paper
PRPatrick Rafail

Key Points

Key points are not available for this paper at this time.

Abstract

Twitter data are widely used in the social sciences. The Twitter Application Programming Interface (API) allows researchers to build large databases of user activity efficiently. Despite the potential of Twitter as a data source, less attention has been paid to issues of sampling, and in particular, the implications of different sampling strategies on overall data quality. This research proposes a set of conceptual distinctions between four types of populations that emerge when analyzing Twitter data and suggests sampling strategies that facilitate more comprehensive data collection from the Twitter API. Using three applications drawn from large databases of Twitter activity, this research also compares the results from the proposed sampling strategies, which provide defensible representations of the population of activity, to those collected with more frequently used hashtag samples. The results suggest that hashtag samples misrepresent important aspects of Twitter activity and may lead researchers to erroneous conclusions.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Patrick Rafail (2017) studied this question.

synapsesocial.com/papers/6a0f7a54fa36b6e053fcb607https://doi.org/10.1177/0894439317709431
Ask AI
Helpful
Bookmark
Share
View Full Paper