Archivarix

web-archive

Trove / Australian Web Archive (NLA)

The National Library of Australia's web archive. PANDORA, the Australian Government Web Archive, and .au domain crawls combined, holding billions of files and searched through the Trove portal. Captures open as replay links on webarchive.nla.gov.au; the Trove API does not cover archived websites, so searching happens on the site itself. The primary source for the history of Australian websites.

Web pages / sites
Pesquisar neste arquivo

Sem verificação programática, abre a própria busca do arquivo.

Por que é útil e como funciona

The Australian Web Archive is the primary historical record of Australian websites, combining the longstanding PANDORA selective archive, the Australian Government Web Archive, and broad crawls of the .au domain. You search and discover archived Australian content through the Trove portal at trove.nla.gov.au, and the captured pages themselves open through the National Library's replay viewer at webarchive.nla.gov.au. Because the Trove search API does not expose archived website records, Echo links you directly to the replay address for a given URL so you can view it there.

O que há dentro

The combined collection held close to 700 terabytes and over 10 billion files as of 2019, and has continued to grow. The curated PANDORA component alone holds over 1.3 billion files across nearly 80,000 archived titles.

Acesso via API

Trove API v3 (key) does NOT expose archived-website records. Replay https://webarchive.nla.gov.au/awa/ <ts>/<url>.

É necessária uma chave de API (geralmente gratuita); veja os endpoints acima para saber onde obtê-la.

Acesso

Catalogue link only, open the archive to search it yourself.

Página inicial

https://trove.nla.gov.au/