Archivarix

web-archive

Archive-It (Internet Archive)

The Internet Archive's subscription archiving service, used by more than a thousand libraries, governments, and universities to build curated web collections that together hold tens of billions of documents. Captures are viewed through replay links on wayback.archive-it.org, organized by collection. It pays off when a specific institution's collection covers your subject; coverage is collection-by-collection, so it is not the place to look up arbitrary URLs.

Web pages / sites
Buscar en este archivo

Sin comprobación programática; abre la propia búsqueda del archivo.

Por qué es útil y cómo funciona

Archive-It is where institutions such as universities, government bodies, and libraries build and maintain their own focused web collections, so it excels when your subject falls within a specific organization's collecting area. Over 1,000 partner institutions worldwide have created more than 10,000 public collections, all of which are full-text searchable on archive-it.org. You can browse by organization or topic, then view captures through a Wayback-style replay interface. Because coverage is collection-by-collection, it works best when you have identified a relevant collection rather than when searching for an arbitrary URL.

Qué contiene

Partners have archived tens of billions of documents collectively. The service has been growing since 2006 and the combined holdings span petabyte scale, though a single aggregate figure is not published separately from the broader Internet Archive.

Acceso por API

Per-collection CDX https://wayback.archive-it.org/ <collId>/timemap/cdx?url= (works) ; replay https://wayback.archive-it.org/ <collId>/<ts>/<url> . Aggregate /all/ CDX is blocked.

Acceso

Programmatic API access (a key may be required, see the API tag).

Página principal

https://archive-it.org/