For the complete documentation index, see llms.txt. This page is also available as Markdown.

Wayback Machine

The Wayback Machine is the Internet Archive's free tool for viewing and saving archived web pages, with over a trillion pages captured, widely used for historical research and digital preservation.

URL

https://web.archive.org/

Description

The Wayback Machine has been archiving the web since 1996 and launched publicly in 2001. The Internet Archive most recent report (October 2025) indicates that it holds more than one trillion archived web pages, totalling over 99 petabytes of data.

For example, here is how amazon.com looked in 1999 versus 2024:

amazon.com in 1999
amazon.com in 1999
amazon.com in 2024
amazon.com in 2024

The Wayback Machine can be accessed not only through the Internet Archive's website, but also through:

Several APIs are also available to access information about Wayback capture data. View the available APIs here.

Learn more about the Wayback Machine.

The Internet Archive

The Internet Archive includes numerous other projects. Some are designed for institutions and organizations with paid tiers, and others are free. View other projects by the Internet Archive.

Cost

Level of difficulty

star

Requirements

The Wayback Machine requires an internet connection to browse archived pages. Separately, the Internet Archive runs an Offline Archive initiative that distriburse its books, audio, video and educational collections to low connectivity regions via local servers and physical media, albeit this doesn't currently extend to offline access of the Wayback Machine's web archieve itself.

Limitations

The Wayback Machine is a powerful tool, but it has some limitations, including:

  • Incomplete Archives: Not all websites or web pages are archived, and some might have gaps in the timeline.

  • Dynamic Content: Interactive elements, dynamic content, and multimedia (such as videos and animations) may not be fully captured or functional.

  • Legal Restrictions: Some websites may block archiving or request the removal of their archived content, limiting access. Learn more about requests to remove content from the Wayback machine here.

  • Loading Issues: Archived pages can load slowly, and some resources (like images or scripts) might be missing.

  • Login Restricted Content: Pages behind paywalls or requiring a username and password generally cannot be crawled, so this content is typically absent from the archive.

  • Service Outages: As a single nonprofit run service, the Wayback Machine can experience extended downtime, including a major cyberattack in Octover 2024, therefore can't be relied on as permanently available.

Ethical Considerations

Using the Wayback Machine may involve several ethical considerations:

  • Accuracy and Context: Archived pages may lack the context in which they were originally presented, potentially leading to misinterpretation or misuse of information.

  • Consent: Some websites may not consent to their content being archived. It's important to consider whether the website owner has explicitly requested exclusion from the archive.

  • Copyright: The Internet Archive's Terms of Use, Privacy Policy, and Copyright Policy states: "You agree to abide by all applicable laws and regulations, including intellectual property laws, in connection with your use of the Archive. In particular, you certify that your use of any part of the Archive's Collections will be limited to noninfringing or fair use under copyright law."

  • Privacy: Archived pages can preserve personal information, for example old posts, photos, or contact information, that an individual no longer wants publically accessible, and this is a seperate concern from site owner consent, since that person affected may not be the one who published the page. The Internet Archive maintains distinct removal channels for provacy and non consensual imagery requests, separate from copyright claims.

  • Potential for Misuse: Because deleted or edited content can remain permanently retrievable through the archive, it can be used to resurface a persons past statements, images, or associations out of context, which raises quesitons of proportionality and fairness independent of whether the original archiving was itself justified.

Guides and articles

Watch an introduction video on How to Use the Wayback Machine.

How to Save Pages With the Wayback Machine:

  1. Go to the Wayback Machine website at https://web.archive.org/

  2. On the homepage of the Wayback Machine, locate the "Save Page Now" section. It's usually found near the top of the page.

Save Page Now section on https://web.archive.org/
Save Page Now section on https://web.archive.org/
  1. Enter the URL of the website you want to archive.

  2. Click the "Save Page" button. This will prompt the Wayback Machine to take a snapshot of the specified webpage.

  3. The Wayback Machine will process the request and capture the webpage. Depending on the page's complexity and size, this might take a few seconds to a minute.

  4. Once the snapshot is complete, the Wayback Machine will provide a link to the archived version of the page. You can click on this link to view the archived page.

  5. Copy the provided link for future reference. This URL is a permanent link to the archived version of the page and can be shared or cited as needed.

Note: This only saves the single page you enter, not the rest of the site. To archive additional pages, repeat the process for each URL.

Learn more about How to save pages with the Wayback Machine.

How to Access Archived Pages Using the Wayback Machine:

  1. Go to the Wayback Machine website at https://web.archive.org/.

  2. Enter the website URL you want to view an archived version of.

  3. After entering the URL, you will be taken to a calendar view. This calendar shows the dates on which the Wayback Machine has website snapshots. Select a year in the timeline above the calendar to narrow down your options.

  4. Once you have selected a year, click on a specific date highlighted on the calendar. These highlighted dates indicate that snapshots of the site are available for that day. They may be different colors, but you will usually want to select the blue dots or links, as they indicate successful responses to the capture. The colors you may see and what they mean include:

    • Blue: The web server returned a successful response (status code 2nn).

    • Green: The web server redirected the request (status code 3nn).

    • Orange: There was a client error (status code 4nn).

    • Red: There was a server error (status code 5nn).

  5. After selecting a date, you can narrow it further by time of day if multiple snapshots are available from the same day. Once you choose, the Wayback Machine will display the archived version of the website as it appeared on that day and time. You can navigate the website as if browsing it on that particular date.

Note: If a linked resoucce on the page wasn't captured at exactly that time, the Wayback Machine will subsititute the closeset available snapshot of it instead, so an archived page can sometimes show a mix of elements from slightly different dates.

Tool provider

The Internet Archive, United States

Advertising Trackers

Page maintainer

Bellingcat Volunteer Team

Last updated

Was this helpful?