Archivarix

web-archive

Webrecorder / Browsertrix

The maker of Browsertrix and the successor toolchain to Conifer: self-hostable or hosted crawling software that produces WACZ/WARC web-archive files. There is no central public corpus to search, every deployment holds its own captures. Relevant when you want to create high-fidelity archives of sites yourself rather than search existing ones.

Pesquisar neste arquivo

Sem verificação programática, abre a própria busca do arquivo.

Por que é útil e como funciona

Webrecorder is the open-source project behind Browsertrix and the WACZ file format, which succeeded Conifer as the toolchain for high-fidelity web archiving. It is a platform for creating archives, not a searchable public collection. If your goal is to look up existing archived pages, Webrecorder itself is not the right destination. Echo links you to its site as a resource for those who want to understand the technology or run their own crawls.

O que há dentro

No central public corpus. Every organisation or individual that deploys Browsertrix holds its own private or institutional archive; there is no shared searchable collection.

Acesso via API

Per-deployment REST API (token).

É necessária uma chave de API (geralmente gratuita); veja os endpoints acima para saber onde obtê-la.

Acesso

Catalogue link only, open the archive to search it yourself.

Página inicial

https://webrecorder.net/