social-twitter
IA Twitter Stream Grab
A set of Internet Archive bulk dumps of the old Twitter public sample stream, stored as monthly JSON files and covering billions of tweets from roughly 2011 to 2023. You access it by downloading the dataset files, which are listed via the item's metadata API. It is meant for bulk, offline analysis of historical Twitter data rather than looking up an individual tweet; the dataset is frozen, since the source stream no longer exists.
Чем полезен и как работает
The Internet Archive's Twitter Stream Grab is a research dataset, not a search tool. It holds bulk monthly downloads of Twitter's old public sample stream, covering billions of tweets from roughly 2011 through early 2023, after which the source stream was shut down. Researchers reach for it to study language, events, or trends across that era at scale: you download the compressed monthly files and process them offline. Echo links you to the Internet Archive collection page, where you can browse and download individual files.
Что внутри
Over 9.4 billion tweets spanning roughly 2011 to early 2023, stored as compressed monthly JSON archives. The dataset is frozen and will not grow further, since the Twitter public sample stream no longer exists.
Доступ по API
https://archive.org/metadata/twitterstream ; files via https://archive.org/download/ <item>/<file>
Доступ
Programmatic API access (a key may be required, see the API tag).