Archivarix

social-twitter

IA Twitter Stream Grab

A set of Internet Archive bulk dumps of the old Twitter public sample stream, stored as monthly JSON files and covering billions of tweets from roughly 2011 to 2023. You access it by downloading the dataset files, which are listed via the item's metadata API. It is meant for bulk, offline analysis of historical Twitter data rather than looking up an individual tweet; the dataset is frozen, since the source stream no longer exists.

Social API

Por que é útil e como funciona

The Internet Archive's Twitter Stream Grab is a research dataset, not a search tool. It holds bulk monthly downloads of Twitter's old public sample stream, covering billions of tweets from roughly 2011 through early 2023, after which the source stream was shut down. Researchers reach for it to study language, events, or trends across that era at scale: you download the compressed monthly files and process them offline. Echo links you to the Internet Archive collection page, where you can browse and download individual files.

O que há dentro

Over 9.4 billion tweets spanning roughly 2011 to early 2023, stored as compressed monthly JSON archives. The dataset is frozen and will not grow further, since the Twitter public sample stream no longer exists.

Acesso via API

https://archive.org/metadata/twitterstream ; files via https://archive.org/download/ <item>/<file>

Acesso

Programmatic API access (a key may be required, see the API tag).

Página inicial

https://archive.org/details/twitterstream