๐ท๏ธ Crawler Tool
Deep-crawl and archive web pages from the early internet era
url
geocities.com
angelfire.com
tripod.com
yahoo.com (1995)
altavista.com
amazon.com (1995)
First Web Page
CERN's original 1990 page
GeoCities Recovery
Mass crawl of GeoCities pages
WebRing Archive
Crawl a WebRing chain
W3C Specifications
HTML 2.0 / 3.2 specs
๐ฏ Era Detection (Optional)
๐
Pre-Web
1989โ1993
๐พ
Web 1.0 Early
1994โ1996
๐
GeoCities Era
1997โ1999
๐ฟ
Dot-Com Boom
2000โ2001
๐ฑ
Post Crash
2002โ2004
๐ฎ
Any Era
No filter
โ Crawl Settings
โพ
Crawl Depth
Max link depth to follow
Max Pages
Cap on total pages to crawl
Politeness Delay
Seconds between requests
User Agent
Browser identity for requests
๐ Content Filters
โพContent Types
HTML
GIF
JPEG
MIDI
AVI
PDF
CSS
JS
Follow Redirects
Respect robots.txt
Archive Subdomains
Detect Era Metadata
โ Crawl in Progress
0
Pages
0 KB
Size
0
Errors
0s
Elapsed
Recent Crawled URLs
[--:--:--]
INFO
[--:--:--]
INFO
[--:--:--]
INFO
[--:--:--]
INFO
๐ Crawl Results 0
| URL | Status | Format | Size | Detected Era | Crawled At | Actions |
|---|
No crawl results yet
Enter a URL above and start a crawl to see archived pages appear here. Each page is preserved with its original format and metadata.