methods / archived-copies
Read a blocked site from the archive - no tools at all
Short answer
Read blocked sites using archived copies from the Wayback Machine or archive.today.
Introduction to Archived Copies
When a website is blocked, accessing its content can be challenging. However, archived copies of websites can provide a way to read blocked content without needing any special tools. The Internet Archive's Wayback Machine and archive.today are two services that offer archived copies of websites. These services operate their own servers in their own jurisdictions, which means that if a site is blocked for you, its archived copy is usually still accessible with just one click - your network blocks the target site, not the archive.
The Two Archiving Services
- Wayback Machine (web.archive.org) - This is the big, institutional archive with decades of history. You can reach it by prefixing the URL of the blocked site with
https://web.archive.org/web/https://the-blocked-site.com/, which serves the most recent capture. The Wayback Machine is an initiative of the Internet Archive, a 501(c)(3) non-profit building a digital library of Internet sites and other cultural artifacts in digital form [1]. - archive.today (archive.ph and its mirrors) - This service captures web pages on demand and is better at handling JavaScript-heavy pages. It's particularly useful when the Wayback Machine has no recent snapshot of the page you're trying to access. archive.today provides a short and reliable link to an unalterable record of any web page [2].
What Works Well Through the Archive
- News Articles and Long-form Text: Archived copies are ideal for reading news articles, blogs, and other long-form text content. Since these types of content are mostly static, they are preserved well in archives.
- Documentation, Blogs, Forum Threads: Technical documentation, personal blogs, and forum threads are also well-suited for archiving. These resources often contain valuable information that remains relevant over time.
- Static Reference Pages: Pages that serve as references, such as dictionaries, wikis, and knowledge bases, are another type of content that works well through archives.
What Fails
- Logins and Personalized Pages: The archive shows the public snapshot of a page, not your personalized session. This means you won't be able to access content that requires you to be logged in.
- Live Feeds and Video: Media embeds usually point back to the blocked origin and fail to load in the archived version. This is because the archive can't access the live feed or video content from the original site.
- Commenting, Posting, Anything Interactive: Since archived pages are snapshots of the past, you can't interact with them. This means commenting, posting, or any other interactive feature won't work.
When the Archive is the Wrong Tool
If you need to access the site itself for activities like logins, watching videos, or uploading content, the archive isn't the right tool. In such cases, consider using other methods to unblock the site, such as DoH for name blocks or WARP or Psiphon for connection blocks. The archive is meant as a reading tool, not an access tool, and it's essential to understand its limitations.
A Note on Paywalls
Archived copies of paywalled news pages exist because people snapshot them before or after they become paywalled. Whether reading an archived paywalled article is the right thing to do is a question of the publisher's terms and your own ethics - not something this site takes a position on. The tool itself is neutral and public [1].
Using the Archive Effectively
To get the most out of archived copies, it's crucial to understand how to use them effectively. For instance, if you're trying to access a blocked news site, you can use the Wayback Machine to find the most recent snapshot of the article you're interested in. If the article is not available on the Wayback Machine, you can try searching for it on archive.today.
Searching the Archives
Both the Wayback Machine and archive.today allow you to search for archived copies of websites. On the Wayback Machine, you can enter the URL of the site you're looking for and browse through the available snapshots [1]. On archive.today, you can search for snapshots by entering the URL or a part of the URL, and the service will show you all the available snapshots [2].
Capturing Pages
If you find a page that you want to ensure remains available, you can capture it using archive.today. This service allows you to take a 'snapshot' of a webpage that will always be online, even if the original page disappears [2]. This feature is particularly useful for preserving content that might change soon, such as price lists, job offers, or real estate listings.
Conclusion
Archived copies of websites can be a powerful tool for accessing blocked content. By understanding how to use services like the Wayback Machine and archive.today, you can read blocked news articles, access documentation, and view static reference pages without needing any special tools. Remember, the archive is a reading tool, not an access tool, and it's essential to respect the limitations and terms of use of these services.
Additional Tips
- Always verify the authenticity of the archived content to ensure it hasn't been tampered with.
- Be aware of the potential for archived content to be outdated.
- Consider supporting the Internet Archive and archive.today through donations to help them continue providing these valuable services [1][3].
Archiving Services and Their Roles
The Internet Archive and archive.today play critical roles in preserving the internet's cultural and historical content. By archiving websites, they help ensure that valuable information remains accessible even if the original site is blocked or taken down. These services rely on donations and support from users to continue their work [3].
The Importance of Preservation
Preserving the internet's content is essential for maintaining access to information and knowledge. Archived copies of websites can serve as a window into the past, allowing us to understand how things have changed over time. They can also provide valuable insights into historical events, cultural trends, and social movements.
Future of Archiving
As the internet continues to evolve, the importance of archiving services will only grow. With more content being created and shared online, the need to preserve this content for future generations will become increasingly critical. Services like the Wayback Machine and archive.today will play a vital role in this effort, and it's essential that we support them in their mission to preserve the internet's history.
By understanding the role of archiving services and how to use them effectively, we can ensure that valuable content remains accessible, even in the face of blocking or censorship. Whether you're a researcher, a student, or simply someone interested in preserving the internet's history, archived copies of websites are a powerful tool that can help you achieve your goals.
FAQ
What is the Wayback Machine?
The Wayback Machine is a digital archive of the internet, allowing users to access archived versions of websites.
How do I access archived copies of a website?
You can access archived copies by using the Wayback Machine or archive.today, and entering the URL of the website you want to access.
Can I use archived copies to access paywalled content?
Archived copies of paywalled content may exist, but whether reading them is acceptable depends on the publisher's terms and your ethics.
Why can't I interact with archived pages?
Archived pages are snapshots of the past and do not support interactive features like commenting or logging in.
How can I support archiving services?
You can support services like the Internet Archive and archive.today through donations to help them continue preserving the internet's content.