Tech article
Timestamping a Giant Record of the Web
No preview is available. Read the original article for the full story.
Hacker News | Oct 9, 2026 | arthuredelstein
Automated excerpt
It happens to be the number of web page snapshots captured by the Common Crawl project over the past 18 years. Common Crawl's hundreds of billions of snapshots save web pages precisely as they existed over the years, totalling more than 10 petabytes of snapshot data. Fortunately, Common Crawl had already hashed all web page snapshots and collected the hashes in blocks of up to 3000 index entries.
Selected automatically from source text; not independently written or fact-checked. Read the original for full context.