web-archive
Webrecorder / Browsertrix
The maker of Browsertrix and the successor toolchain to Conifer: self-hostable or hosted crawling software that produces WACZ/WARC web-archive files. There is no central public corpus to search, every deployment holds its own captures. Relevant when you want to create high-fidelity archives of sites yourself rather than search existing ones.
Sem verificação programática, abre a própria busca do arquivo.
Por que é útil e como funciona
Webrecorder is the open-source project behind Browsertrix and the WACZ file format, which succeeded Conifer as the toolchain for high-fidelity web archiving. It is a platform for creating archives, not a searchable public collection. If your goal is to look up existing archived pages, Webrecorder itself is not the right destination. Echo links you to its site as a resource for those who want to understand the technology or run their own crawls.
O que há dentro
No central public corpus. Every organisation or individual that deploys Browsertrix holds its own private or institutional archive; there is no shared searchable collection.
Acesso via API
Per-deployment REST API (token).
É necessária uma chave de API (geralmente gratuita); veja os endpoints acima para saber onde obtê-la.
Acesso
Catalogue link only, open the archive to search it yourself.