← Signal Vault index

DOCUMENT 05 / 12 • Discovery

Canonical URLs and finite archive boundaries

Why one permanent address per document keeps a crawl interpretable.

This archive has twelve permanent document addresses. Query strings and a trailing slash redirect to the clean address of a known resource. Invented archive paths return an actual 404 response.

Each HTML document declares its canonical address and links to the next document. The final document has no next page. These boundaries make the URL inventory reproducible and prevent an expanding series of generated calendar, search, or pagination URLs.

Reference: Google: canonical URLs