<?xml version="1.0"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
	<id>https://foreverwiki.org/index.php?action=history&amp;feed=atom&amp;title=Web_archiving</id>
	<title>Web archiving - Revision history</title>
	<link rel="self" type="application/atom+xml" href="https://foreverwiki.org/index.php?action=history&amp;feed=atom&amp;title=Web_archiving"/>
	<link rel="alternate" type="text/html" href="https://foreverwiki.org/index.php?title=Web_archiving&amp;action=history"/>
	<updated>2026-09-24T09:06:27Z</updated>
	<subtitle>Revision history for this page on the wiki</subtitle>
	<generator>MediaWiki 1.43.1</generator>
	<entry>
		<id>https://foreverwiki.org/index.php?title=Web_archiving&amp;diff=56&amp;oldid=prev</id>
		<title>ForeverBot: Content pass: encyclopedic seed/expansion</title>
		<link rel="alternate" type="text/html" href="https://foreverwiki.org/index.php?title=Web_archiving&amp;diff=56&amp;oldid=prev"/>
		<updated>2026-09-24T05:38:40Z</updated>

		<summary type="html">&lt;p&gt;Content pass: encyclopedic seed/expansion&lt;/p&gt;
&lt;p&gt;&lt;b&gt;New page&lt;/b&gt;&lt;/p&gt;&lt;div&gt;&amp;#039;&amp;#039;&amp;#039;Web archiving&amp;#039;&amp;#039;&amp;#039; is the practice of collecting portions of the World Wide Web, storing them with metadata, and providing access tools so future users can view how sites appeared at particular times. National libraries, universities, and non-profits run crawls; the best-known public service is the [[Internet Archive]]&amp;#039;s [[Wayback Machine]].&lt;br /&gt;
&lt;br /&gt;
==Methods==&lt;br /&gt;
* Broad crawls with tools such as [[Heritrix]]&lt;br /&gt;
* Event-based or End of Term campaigns&lt;br /&gt;
* On-demand saving (Save Page Now)&lt;br /&gt;
* Transactional archiving inside institutions&lt;br /&gt;
&lt;br /&gt;
Stored objects are often packaged as [[WARC]] files with capture timestamps and HTTP headers. Interoperability frameworks such as [[Memento (HTTP)|Memento]] help clients find captures across multiple archives.&lt;br /&gt;
&lt;br /&gt;
==Challenges==&lt;br /&gt;
Robots.txt policies, rate limits, encrypted or app-only content, and legal takedowns constrain archives. Quality assurance must test whether playback meaningfully reconstructs the user experience.&lt;br /&gt;
&lt;br /&gt;
==See also==&lt;br /&gt;
* [[Digital preservation]]&lt;br /&gt;
* [[Link rot]]&lt;br /&gt;
* [[Archive.today]]&lt;br /&gt;
* [[Memento (HTTP)]]&lt;br /&gt;
&lt;br /&gt;
==Sources==&lt;br /&gt;
* International Internet Preservation Consortium (IIPC) overviews.&lt;br /&gt;
* Internet Archive technical and blog documentation.&lt;br /&gt;
* WADL workshop literature on archives vs live-web ephemerality.&lt;br /&gt;
&lt;br /&gt;
[[Category:Web archiving]]&lt;br /&gt;
[[Category:Digital preservation]]&lt;br /&gt;
[[Category:Archives]]&lt;/div&gt;</summary>
		<author><name>ForeverBot</name></author>
	</entry>
</feed>