What the Internet Archive does and why it matters

The Internet Archive is a nonprofit organization that has been photographing the web since 1996. It stores billions of web pages in a searchable library called the Wayback Machine. When you save a page there, you create a permanent record of what that page looked like on a specific date — even if the original site disappears, changes completely, or gets taken down.

This matters because websites vanish. A business closes and its domain expires. A news article gets deleted. A recipe blog moves to a new platform and the old one goes blank. The Wayback Machine lets you retrieve what was there, which is useful for fact-checking, finding old information, or straightforward preserving something you care about.

The Archive also works as a backup for your own bookmarks. If you save a page there and then lose your bookmark list in a device crash or account change, you can still find it through the Wayback Machine's search function. You do not need an account to use it.

Key Takeaways

  • The Wayback Machine at archive.org lets you view snapshots of any website from past dates, going back to 1996 for many sites.
  • You can manually save a page to the Archive by entering its URL on the Wayback Machine homepage, creating a snapshot dated that day.
  • Saved pages show you the exact layout, images, and text from when they were captured, which helps verify what a site said at a particular time.
  • The Archive respects removal requests from site owners, so not every page is available, and some content may be blocked by copyright or legal holds.

How to save a page to the Wayback Machine manually

Go to archive.org and look for the box labeled "Save Page Now" on the homepage. Paste the full URL of the page you want to save — for example, https://www.example.com/article/my-recipe — and click the button. The Archive will visit that page, photograph it, and store it with today's date.

The process takes a few seconds to a few minutes depending on how large the page is and how busy the Archive's servers are. Once it finishes, you will see a confirmation page with a link to your saved snapshot. You can share that link with anyone, and they will see the page exactly as it appeared when you saved it.

Save the confirmation link or bookmark it yourself if you think you might need to find this snapshot again. The Archive keeps the page indefinitely, but having your own bookmark means you do not have to search for it later.

What you can and cannot see in saved pages

Saved pages show text, images, and the page layout from the moment they were captured. However, interactive elements often do not work. If the original page had a search box, a comment form, or a video player, those will appear in the snapshot but will not function — they were live features that cannot be frozen in time.

Some pages do not appear in the Wayback Machine at all. Site owners can request that the Archive remove their pages, and the Archive honors those requests. Pages behind paywalls or login screens are usually not captured. Very new pages may not have been crawled yet. And pages that use heavy JavaScript or dynamic content sometimes do not save well because the Archive captures the page's code, not always what you see on screen.

If you are looking for a specific page and the Wayback Machine shows no snapshots, try searching for a different URL from the same site — sometimes the Archive has captured the homepage or a different article but not the exact page you want.

Finding pages that were saved before you knew about the Archive

The Wayback Machine does not only store pages you personally saved. It also holds snapshots that its automated crawlers captured while visiting websites over the past 25+ years. You can search for any URL and see a calendar of dates when that page was photographed.

Enter a URL into the search box at archive.org and you will see a timeline. Blue dots on the calendar mark dates when snapshots exist. Click a date and you will see that version of the page. This is how you can find old versions of a site that disappeared years ago, or see what a company's homepage looked like in 2010.

The further back you go, the fewer snapshots usually exist — the Archive's crawlers were less frequent in the 1990s and early 2000s. But for most major websites, you can find snapshots from multiple years, sometimes multiple times per year.

Using the Archive as a backup for your own research

If you are writing something that cites a specific web page, saving it to the Wayback Machine creates a dated record of what that page said. This protects you if the page later changes or disappears and someone questions your source. You can point to the snapshot and say, "This is what it said on March 15, 2024."

Journalists, researchers, and fact-checkers use the Wayback Machine this way constantly. It is also useful for tracking how a company's website or messaging has changed over time, or for finding deleted social media posts that were screenshotted and archived.

When you save a page for this reason, note the URL and the date of your snapshot. Store that information alongside your notes or citation. The snapshot itself will remain accessible through archive.org as long as the Archive exists.

What happens when a site owner asks for removal

The Internet Archive respects requests from site owners and copyright holders to remove pages from the Wayback Machine. If you own a website and want your old snapshots deleted, you can submit a removal request through archive.org's exclusion form. The Archive will review it and usually honor it within a few weeks.

This means some pages you are looking for may have been removed at the owner's request, even though they were once in the Archive. If you search for a URL and see a message saying the page is not available, it may have been removed rather than never captured.

The Archive also respects robots.txt files — a standard file that website owners use to tell crawlers which pages they do not want indexed. If a site has always had a robots.txt blocking the Wayback Machine, the Archive will not have captured it.

Combining the Archive with your bookmark system

The Wayback Machine works best alongside your regular bookmarking tools, not instead of them. Use your browser bookmarks or a service like Pocket for pages you visit often or plan to read soon. Use the Wayback Machine when you want a permanent, dated record of a page — especially if you think the page might change or disappear.

If you use a bookmark manager that syncs across devices, you can store the Wayback Machine link instead of the original URL. This gives you two benefits: your bookmark stays in sync across your devices, and you have a snapshot that will not change even if the original site does.

For pages you want to keep long-term, consider doing both: bookmark the original URL in your regular system, and also save a snapshot to the Wayback Machine. That way you have the live page if you want to check for updates, and a frozen copy if the original disappears.

Frequently Asked Questions

Can I remove my own website from the Wayback Machine?

Yes. Go to archive.org/about/exclude.php and fill out the exclusion request form. You will need to verify that you own or represent the website. The Archive usually processes requests within a few weeks and will remove all snapshots of your site.

Does saving a page to the Wayback Machine notify the website owner?

No. The Archive does not send notifications to site owners when pages are saved. Your action is private. However, the Archive's automated crawlers visit websites regularly and capture pages without notifying owners, which is standard practice across the web.

What if the page I want to save is behind a login or paywall?

The Wayback Machine cannot capture pages that require you to log in or pay to view. It will save the login page or paywall message instead. For pages you personally have access to, you can take a screenshot and store it yourself, or use a different archiving tool designed for private content.

How long does the Internet Archive keep pages?

The Archive keeps pages indefinitely unless the site owner requests removal. It is a nonprofit with a mission to preserve digital culture, so pages are not deleted due to age or lack of use. However, the Archive is dependent on funding, so there is always some risk to any single organization's long-term survival.

Can I search the Wayback Machine by date range?

You can see a calendar of available dates and click the one you want, but there is no way to search for "all snapshots between January and March 2020" at once. You have to browse the calendar and click individual dates. For most sites, you can narrow it down by looking at which months have blue dots indicating snapshots.