Check before you share
Media Literacy Guide

Tracing Deleted or Edited Pages with the Wayback Machine and archive.today

Step-by-step guide to using the Wayback Machine and archive.today to find deleted or edited web pages, with phone and laptop instructions for verifying what a page originally said.

The web is not a static library. It is a stream, constantly edited, deleted and rewritten. When a page vanishes or changes, the original words often survive in a web archive. Knowing how to pull a deleted page back into view is the single most useful skill for research, fact-checking or simply remembering what a page used to say. Two tools do the heavy lifting: Wayback Machine and archive.today. Neither is magic. Here are the exact steps on a phone and a laptop, what each service can and cannot prove, and how to fit them into the S.U.R.E. framework for source verification.

Open your browser and type web.archive.org. You are now at the Internet Archive, a non-profit library saving the web since 1996. The search bar is the main entrance. Paste the full URL of the deleted or edited page into it. Hit enter. You land on a calendar view showing every snapshot of that page over time. Blue dots mean a copy exists for that day; a solid circle means the capture is a redirect. Look for the earliest date after the page was published and the last date before it disappeared. Click any date to load that snapshot. The URL changes to a long string containing the timestamp in YYYYMMDDHHMMSS format, UTC. That number is your proof of when the page was captured.

Archive tools: Wayback Machine and archive.today for verification
Mojmir Churavy , CC0 via Wikimedia Commons

Navigating The Snapshots

On a laptop, the Wayback Machine shows a toolbar at the top of the archived page with the capture date. On a phone, the toolbar may be hidden; scroll up to see it. If the page was saved, you see the original content, images and links intact. Some interactive elements and JavaScript-rendered content fail in older snapshots. That is a known limit: a snapshot is a static copy of the HTML, not a live simulation. If the page you need is missing, do not give up. Use the site search pattern web.archive.org/web/*/example.com to see every saved page under that domain. This is valuable for tracing a domain's history, watching how a news outlet changed its headline, or verifying whether a claim existed at all.

When The Wayback Machine Has Nothing

For a page deleted recently, or one the Wayback Machine never crawled, you need a different tactic. This is where archive.today comes in. It is a separate service that creates snapshots on demand. Go to archive.today and paste the URL into the form. The service displays one of two things: an existing snapshot, or a prompt asking if you want to save the page now. Click save. The capture takes a few minutes to a few hours depending on server load. The result is a static HTML copy plus a PNG screenshot of the page as it appeared. This screenshot is valuable because it shows the page with its styling and layout, which is often missing from the Wayback Machine's text-based captures.

Pages That Fight Back

archive.today is particularly useful for pages that block the Wayback Machine. Some sites use a robots.txt file to tell crawlers to stay away, and since 2017, the Internet Archive has honoured those requests retroactively. A site owner can scrub past captures. archive.today ignores these requests, capturing content the Wayback Machine cannot. It is also the better choice for a specific social media post or a paywalled article, as it can render the full text where the Wayback Machine only gets a teaser. For geoblocked content, archive.today can capture a version otherwise inaccessible from your region, acting as a location-agnostic browser.

Use Both, Every Time

Both tools serve the same purpose but have different strengths. The Wayback Machine holds the larger database by far, over 916 billion pages, and its API and Memento protocol make it a standard for researchers. But it is not perfect. If a page was never crawled, or was retroactively blocked, the Wayback Machine returns a 404 error. archive.today complements it with a direct save feature and the ability to capture pages that resist other archives. For any critical verification, use both. A page missing from one may be present in the other. That is the practical application of fact-checking methodology: cross-verify with multiple sources, including the archives themselves.

If The Normal Route Is Closed: The One-Click Save

The most common failure case is needing to save the page yourself before it disappears. The Wayback Machine offers a Save Page Now feature at web.archive.org/save. Paste your URL and click save. There is a limit of ten concurrent jobs per user, and a file size limit of 500 MB per capture, which is generous. The capture delay is a few minutes, but can stretch to hours for slow pages. For archive.today, use the bookmarklet: drag the link from the site to your browser bookmarks bar. When you find a page you want to save, click the bookmarklet. It opens a new tab with the archive.today form pre-filled. This is faster and more reliable than copying and pasting URLs, especially on a phone where switching between apps is slow. Do this before you share the link, not after. If you share a URL that is later deleted, you have lost the evidence. If you save it first, you have created a verifiable record.

Browser Extensions and Mobile Workarounds

For laptop users, the official Wayback Machine browser extension for Chrome and Firefox automates much of this. It offers a one-click save function, and detects when you hit a 404 page, offering to show you the archived version instead. It also has a 'first version' and 'recent version' comparison feature, a quick way to see how a page has changed over time. On mobile, the process is less smooth. The browser extension is not available on iOS or Android mobile browsers. Instead, use the share sheet on your phone: navigate to the page, tap share, and choose the bookmarklet (on Android) or the archive.today app (on iOS, not official). This is clunky but workable. The key is to make saving a habit before you rely on finding something that is already gone.

What A Snapshot Actually Proves

Once you have a snapshot, you must know what it can and cannot prove. An archived page proves that a specific URL was reachable and contained certain content at the exact time of the capture. The timestamp is the key. It does not prove that the content was true, only that it was published. A snapshot is not a verification of facts, but a verification of publication. This is a critical distinction for lateral reading: you still need to open new tabs and check the source, the author, and the claims. The archived page is the starting point for your investigation, not the end. It tells you the original context, the original headline, and the original URL, which may have been changed later.

Chain Of Custody For Fact-Checkers

For fact-checkers, the archived page serves as a chain-of-custody document. To verify a claim, you document the original URL, the archive URL, and the capture timestamp. This is the standard for source citation in a fact-check report. If you are verifying a screenshot, the archived page is independent corroboration that the screenshot was not fabricated in an image editor. The screenshot could be fake; the archive is harder to fake. But you must check that the URL in the archive matches the URL in the screenshot. A screenshot of an article with one URL and an archive of a different URL proves nothing.

Archives Have Limits

There is a dark side to the archives: they are not infallible. The Wayback Machine has a retroactive robots.txt policy. A site owner can block past captures by updating their robots.txt file, erasing parts of the digital footprint. archive.today has a different problem: its operator is anonymous, and its servers sit in multiple jurisdictions, so availability can be sporadic. It is occasionally blocked by some internet service providers or suffers slow response times. If you need a page archived for legal reasons, do not rely on a single service. Save it in both. In Singapore, where content may be subject to POFMA correction directions or taken down by platform enforcement, having your own archive is personal record-keeping independent of any platform or government.

What It Costs You

The practical detail matters. The Wayback Machine is free. archive.today is free. There is no paywall, no premium tier. What it costs is time. The Wayback Machine can be slow to load a snapshot, taking 10 to 20 seconds. archive.today can be fast. The failure case arrives when you are on a slow connection or the archive is down. What do you do? You wait, try the other service, or use a different technique: search for a cached version of the page in Google or Bing, though these caches are increasingly discontinued. A better backup is the S.U.R.E. method's principle: use the archive to find the source, then go find the source's original document, even if it is not the exact page you wanted. The URL is a pointer; the archive is a copy of the pointer, and the actual content may live elsewhere.