web-archive
Archive-It (Internet Archive)
The Internet Archive's subscription archiving service, used by more than a thousand libraries, governments, and universities to build curated web collections that together hold tens of billions of documents. Captures are viewed through replay links on wayback.archive-it.org, organized by collection. It pays off when a specific institution's collection covers your subject; coverage is collection-by-collection, so it is not the place to look up arbitrary URLs.
Программная проверка недоступна, откроется собственный поиск архива.
Чем полезен и как работает
Archive-It is where institutions such as universities, government bodies, and libraries build and maintain their own focused web collections, so it excels when your subject falls within a specific organization's collecting area. Over 1,000 partner institutions worldwide have created more than 10,000 public collections, all of which are full-text searchable on archive-it.org. You can browse by organization or topic, then view captures through a Wayback-style replay interface. Because coverage is collection-by-collection, it works best when you have identified a relevant collection rather than when searching for an arbitrary URL.
Что внутри
Partners have archived tens of billions of documents collectively. The service has been growing since 2006 and the combined holdings span petabyte scale, though a single aggregate figure is not published separately from the broader Internet Archive.
Доступ по API
Per-collection CDX https://wayback.archive-it.org/ <collId>/timemap/cdx?url= (works) ; replay https://wayback.archive-it.org/ <collId>/<ts>/<url> . Aggregate /all/ CDX is blocked.
Доступ
Programmatic API access (a key may be required, see the API tag).