This repository allows you to replicate my cursory analysis of @wikileaks over the 2016 election cycle.
See the medium post: Who Watches Wikileaks?
First, setup the project and environment.
git clone https://github.com/jbn/Wikileaks_Analysis.git
cd Wikileaks_Analysis
conda create --name wikileaks_analysis python=3.5 # Or, virtual_env equiv.
source activate wikileaks_analysis
pip install -r requirements.txtThen you have to edit config/credentials.json.orig, adding your Twitter Developer API credentials. Then, rename that file to config/credentials.json.
Then, run
maketo collect the data.
Notebook to get you started with the data is in notebooks.
The original tweet ids are in the original_tweet_ids.txt for exact replication, if desired.
- Bayesian switch-point analysis over various metrics (esp. over the retweet ratio) for detecting changes in interaction patterns, outside of just electoral leaking.
- Linguistic analysis for detecting a regime change in content.
- The collector collects the 100 most recent retweeters for each tweet, along with their user information. This affords the opportunity to look at the alter interactions of wikileaks, overtime. That's want I really wanted to do, but I ran out of time. The 100 limit is an API constraint. (And, sometimes it's less, for deleted tweets, which offers some signal, too).
- Submit a PR for more ideas.