root / archive / internet-history

Internet History Archive

Search, browse, and explore systematically preserved web content from 1990 to 1999. Every URL, format, and metadata record is cryptographically verified and historically contextualized.

4.21M
Pages Indexed
9,847
Unique Domains
1990–1999
Coverage Span
WARC
Storage Format

Curated Collections

View Full Catalog →
🏠

GeoCities & Angelfire

Personal homepages, guestbooks, and early community spaces before social media.

1995–1999182K pages
🏛️

Academic & Government

University departments, research projects, and early .gov/.edu institutional sites.

1991–199864K pages
🛒

Dot-Commerce Era

Early e-commerce catalogs, payment gateways, and pre-Amazon retail experiments.

1996–199941K pages
🔗

WebRings & Directories

Link farms, Yahoo! directories, and community-driven navigation ecosystems.

1994–199929K nodes

Recent Additions

[ Netscape Navigator 3.0 ]
Best viewed 800x600
★ Welcome ★
HTML 3.2142 KB

Midtown/Science Fiction – Personal Archive

http://www.geocities.com/Midtown/SciFi/8842/index.html
Archived: 1997-04-12 Frameset Guestbook CGI
A highly complete preservation of a popular fan site featuring nested frames, MIDI background audio, and an active Perl-based guestbook.
VERIFIED
[ Under Construction ]
Table Layout v2.0
Counter: 004821
HTML 2.086 KB

Digital Dynamics – Early E-Commerce Prototype

http://www.digitaldyn.com/catalog/index.html
Archived: 1998-11-03 CGI Form SSL
Rare capture of a pre-1999 e-commerce catalog featuring server-side Perl forms and early SSL checkout simulation.
VERIFIED
[ University Repository ]
Dept. of Computer Sci
Public Access
HTML 3.054 KB

MIT CSAIL – 1996 Course Materials

http://wwwai.mit.edu/courses/6.034/spring96/
Archived: 1996-09-20 Academic PDF Preview
Institutional archive of open courseware including lecture notes, problem sets, and early digital homework submission portals.
PARTIAL

Preservation Methodology

📡 Crawling & Capture

We use custom HEADLESS-NCSA and early Mosaic-compatible crawlers to render and capture pages exactly as they resolved in their native environments.

  • Respects original robots.txt (historical)
  • Follows 302/301 legacy redirects
  • Captures dynamic CGI states

🗃️ Storage & Format

All artifacts are stored in WARC-0.17 compliant containers with redundant geographic mirroring and checksum verification.

  • WARC + ARC hybrid archives
  • SHA-256 integrity verification
  • Offline tape + encrypted SSD

🔍 Retrieval & Rendering

Our archive reader emulates Netscape 3.x, IE 4.x, and early Opera environments to guarantee accurate visual and behavioral playback.

  • Pixel-accurate font substitution
  • Legacy JavaScript emulation
  • Metadata & provenance tracking
"}{