archive.org
Knowledge Base

Frequently Asked Questions

Comprehensive answers to common questions about link rot, web preservation, and how to recover dead links.

Dead / Broken Link & URL Recovery

Restore access to lost web content. Instantly verify if a dead link or broken URL is offline, explore historical snapshots preserved by the Internet Wayback Machine, and perform seamless link recovery and URL recovery.

Recover Link / URL

Understanding Dead & Broken Links

What is a dead or broken link?

A dead link (also called a broken link, dead URL, or rotted link) is a hyperlink on a webpage that leads to a server or page that is permanently or temporarily unreachable. When clicked, it usually results in an HTTP 404 (Not Found), 410 (Gone), 500 (Internal Server Error), 502/503 (Gateway/Service Unavailable), or a DNS failure (NXDOMAIN).

Why do links go dead over time (link rot)?

Link rot occurs naturally as the web evolves. Primary reasons include: website redesigns or migrations where URL paths change without 301 redirects; domain expirations or non-renewals; businesses shutting down; content deletion by authors; and changes to corporate CMS platforms. Studies indicate that after 10 years, over 60% of web citations become dead links.

What is the difference between a soft 404 and a true 404 dead link?

A true 404 occurs when a web server explicitly returns the HTTP 404 status code indicating the resource is missing. A soft 404 occurs when a server returns a 200 OK success status code, but displays an error message or empty template such as "Page Not Found". Broken Link Recovery uses live status probing to help distinguish between responsive servers and dead endpoints.

Recovery Process & Capabilities

How does dead / broken link recovery work?

When you enter a URL into Broken Link Recovery, our system initiates two simultaneous actions: 1) It sends a fast, secure HTTP probe from our edge servers to verify whether the URL is active, redirected, or returning an error code. 2) It queries the Internet Archive Wayback Machine CDX API to fetch an indexed history of all preserved snapshots. You can then view snapshots by date or extract them directly in Reader Mode.

Can I recover a webpage if the entire domain is offline or expired?

Yes. As long as web crawlers like the Internet Archive (or related archiving initiatives) visited and captured that URL while it was online, historical snapshots remain stored in the public archive. You do not need the original host or domain to be active to recover and read the content.

What if no archive snapshot is found for my dead link?

If a URL has never been crawled by the Internet Archive, or if the original site prevented indexing via robots.txt or an authentication wall, snapshots may not exist. In such cases, you can check alternative web archives like archive.today or Google Cache / search engine web caches.

What is Reader Mode and why is it helpful for broken links?

Historical archives often fail to load external styles, images, or JavaScript files because the third-party CDNs hosting them are no longer online. This causes archived pages to display broken layouts or fail to render entirely. Our clean Reader Mode extracts the core article text and headers directly, rendering them cleanly without broken scripts or annoying archive banners.

What are the best alternatives to the Wayback Machine?

When looking for alternatives to the Wayback Machine or a modern wayback machine alternate, Broken Link Recovery offers the fastest solution for dead link recovery and broken URL recovery. Unlike navigating standard waybackmachine calendar interfaces, our platform performs instantaneous live HTTP availability checks, queries the Internet Archive CDX server directly, and formats content into a distraction-free Reader Mode. Other alternatives include Archive.today, Google Cache, and regional digital libraries.

What is the difference between link recovery and URL recovery?

Link recovery and URL recovery describe the process of restoring lost web content when hyperlinks return HTTP 404, 500, or domain expiration errors. While link recovery focuses on reviving citations and inbound references, URL recovery focuses on recovering past versions of a specific broken URL or expired web domain using historical snapshots from the Internet Wayback Machine.

About Broken Link Recovery

A simple utility to check dead links and find archived snapshots from public web archives.