Archivarix

web-archive

Archive-It (Internet Archive)

The Internet Archive's subscription archiving service, used by more than a thousand libraries, governments, and universities to build curated web collections that together hold tens of billions of documents. Captures are viewed through replay links on wayback.archive-it.org, organized by collection. It pays off when a specific institution's collection covers your subject; coverage is collection-by-collection, so it is not the place to look up arbitrary URLs.

Web pages / sites
Pesquisar neste arquivo

Sem verificação programática, abre a própria busca do arquivo.

Por que é útil e como funciona

Archive-It is where institutions such as universities, government bodies, and libraries build and maintain their own focused web collections, so it excels when your subject falls within a specific organization's collecting area. Over 1,000 partner institutions worldwide have created more than 10,000 public collections, all of which are full-text searchable on archive-it.org. You can browse by organization or topic, then view captures through a Wayback-style replay interface. Because coverage is collection-by-collection, it works best when you have identified a relevant collection rather than when searching for an arbitrary URL.

O que há dentro

Partners have archived tens of billions of documents collectively. The service has been growing since 2006 and the combined holdings span petabyte scale, though a single aggregate figure is not published separately from the broader Internet Archive.

Acesso via API

Per-collection CDX https://wayback.archive-it.org/ <collId>/timemap/cdx?url= (works) ; replay https://wayback.archive-it.org/ <collId>/<ts>/<url> . Aggregate /all/ CDX is blocked.

Acesso

Programmatic API access (a key may be required, see the API tag).

Página inicial

https://archive-it.org/