An iterative topic model filtering framework for short and noisy user-generated data: analyzing conspiracy theories

Gillian Kant1, Levin Wiebelt1, Christoph Weisser2

  • 1University of Göttingen, Göttingen, Germany.

International Journal of Data Science and Analytics
|May 11, 2022
PubMed
Summary

This study introduces Iterative Filtering and Hashtag Pooling to analyze Twitter data on conspiracy theories. During late 2020, "Election Fraud" and "Covid-19-hoax" theories were prominent, linked to US election and pandemic discussions.

Related Concept Videos

Group Polarization01:01

Group Polarization

Group polarization is the strengthening of an original group attitude following the discussion of views within a group (Teger & Pruitt, 1967). That is, if a group initially favors a viewpoint, after discussion the group consensus is likely a stronger endorsement of the viewpoint. Conversely, if the group was initially opposed to a viewpoint, group discussion would likely lead to stronger opposition.
35.8K
Filtration00:53

Filtration

Filtration is a physical separation process that involves passing a suspension through a porous medium to separate solids from fluids. During filtration, solids collect on the porous medium while liquids, also collectively known as the filtrate, pass through. The filtration medium is selected based on the filtration purpose, quantity, and nature of the precipitate. The general criteria for a suitable filtering medium are that it is inert, mechanically strong, nonabsorbent toward dissolved...
1.0K
Sampling Theorem01:15

Sampling Theorem

In signal processing, the analysis of continuous-time signals, denoted as x(t), often involves sampling techniques to convert these signals into discrete-time signals. This process is essential for digital representation and manipulation. A critical component in sampling is the train of impulses, characterized by the sampling interval and the sampling frequency. The relationship between these parameters and the original signal's properties dictates the success of the sampling process.
813
Outliers and Influential Points01:08

Outliers and Influential Points

An outlier is an observation of data that does not fit the rest of the data. It is sometimes called an extreme value. When you graph an outlier, it will appear not to fit the pattern of the graph. Some outliers are due to mistakes (for example, writing down 50 instead of 500), while others may indicate that something unusual is happening. Outliers are present far from the least squares line in the vertical direction. They have large "errors," where the "error" or residual is the...
4.3K
Null and Alternative Hypotheses01:16

Null and Alternative Hypotheses

The actual hypothesis testing begins by considering two hypotheses. They are termed  the null hypothesis and the alternative hypothesis. These hypotheses contain opposing viewpoints.
The null hypothesis, denoted by H0 is a statement of no difference between the variables—they are not related. This can often be considered the status quo. As  a result if you cannot accept the null, it requires some action.
The alternative hypothesis, denoted by H1 or Ha, is a claim about the...
10.3K
Statistical Hypothesis Testing01:16

Statistical Hypothesis Testing

Hypothesis testing is a critical statistical procedure facilitating informed, evidence-based decisions. It begins with a hypothesis, which is a tentative explanation, or a prediction about a population parameter. This hypothesis can be either a null hypothesis (H0), indicating no effect or difference, or an alternative hypothesis (Ha), suggesting an effect or difference.
Statistical significance measures the probability that an observed result occurred by chance. If this probability, known as...
2.1K