Erratum
WARC revisit metadata records
Originally reported by
.
The revisit records in the Common Crawl WARC
archives in all crawls from CC-MAIN-2018-34
to CC-MAIN-2024-46
(since Aug 2018) lack the metadata record which is attached to all response records. Fixed with CC-MAIN-2024-51
, see commoncrawl/nutch#33. Note: before CC-MAIN-2018-34
, WARC
revisit records were not stored at all.
Affected Crawls
Affected Web Graphs
No items found.