147
Total Papers
89
Published
23K
Total Citations
12
Active Projects

Publication Types

πŸ“„

Peer-Reviewed Papers

89 publications

πŸ“‹

Whitepapers

32 publications

πŸ“Š

Technical Reports

18 publications

πŸŽ“

Dissertations

8 publications

Archival Science Published

Cryptographic Provenance Chains for Verified Web Archiving: A Proof-of-Existence Framework

Dr. Raj Patel, Sofia Kim, Dr. Alexander Novak

We propose a cryptographic framework that provides tamper-evident proof of existence for archived web content at specific points in time. By leveraging Merkle tree structures and timestamped blockchain anchors, our system enables researchers to verify that a particular webpage configuration existed on a given date. We evaluate the framework across 2.3 million archived pages from the 1990–2005 period, demonstrating sub-second verification times with zero false positives for content manipulation.

πŸ“… January 2024 πŸ“– Proceedings of the Digital Heritage Conference πŸ‘οΈ 3,102 views ⬇️ 892 downloads πŸ“ 28 citations
Preservation Technology Published

Emulating Netscape Navigator 3.0: A Browser-Based Engine for Authentic Rendering of Legacy Web Pages

Dr. Hiroshi Watanabe, Lisa Rodriguez, TomΓ‘Ε‘ NovotnΓ½

This paper introduces WebRender-N3, an open-source browser engine module designed to authentically render web pages as they appeared in Netscape Navigator 3.0 (1996). Our approach reverse-engineers the proprietary layout algorithms, CSS 0.0 rendering quirks, and table-based layout behaviors to produce an emulation layer that achieves 99.2% visual parity with original Netscape output. We benchmark the engine against 50,000 archived pages and discuss the challenges of rendering deprecated HTML elements, JavaScript 1.0 scripts, and plugin-dependent content.

πŸ“… November 2023 πŸ“– ACM Digital Archives, Vol. 8 πŸ‘οΈ 5,440 views ⬇️ 1,870 downloads πŸ“ 35 citations
Social Impact Published

Lost Voices: Documenting the Cultural Impact of Disappeared Early-Web Communities

Dr. Amara Osei, David Park, Dr. Ingrid Larsson

This qualitative study examines the cultural and social significance of decommissioned early-web communities, focusing on the loss of personal narratives, marginalized voices, and grassroots digital cultures that vanished with the closure of GeoCities, Tripod, and Similar forums. Through oral histories of 127 former webmasters and analysis of recovered content, we document patterns of identity construction, community formation, and creative expression that defined the democratized web of the 1990s. Our findings have direct implications for contemporary platform accountability and digital heritage policy.

πŸ“… September 2023 πŸ“– New Media & Society, Vol. 25 πŸ‘οΈ 6,210 views ⬇️ 2,140 downloads πŸ“ 51 citations
Web Preservation Preprint

Scaling Web Archival Infrastructure: Distributed Crawling of 4.2 Million Pages Using a Federated Network of Preservation Nodes

Dr. Chen Wei, Dr. Elena Vasquez, Michael Torres, Priya Sharma

We describe the architecture and operational results of our distributed web archival system, which coordinates 247 preservation nodes across 38 countries to systematically capture and store early web content. Our federated crawling approach achieved 98.3% coverage of identified 1990s-era domains while reducing single-point failure risks by 73%. We present throughput benchmarks, latency measurements, and lessons learned from two years of continuous archival operations across diverse network conditions and geopolitical boundaries.

πŸ“… June 2024 πŸ“– arXiv preprint πŸ‘οΈ 2,890 views ⬇️ 670 downloads πŸ“ Under review
Digital Humanities Published

Mining the Early Web: Computational Analysis of Design Aesthetics, Content Trends, and Technological Adoption 1991–1999

Dr. Laura Johansson, Kevin Zhang, Dr. Marcus Webb

This paper applies computational text analysis and computer vision techniques to a corpus of 892,000 archived HTML documents from the first decade of the World Wide Web. We identify macro-level trends in web design aesthetics including color palette evolution, typography choices, layout patterns, and multimedia adoption rates. Our NLP analysis reveals shifts in content tone, keyword frequency, and topical interests across the 1990s. The findings provide a quantitative foundation for understanding the formative period of web culture and inform current debates about platform design homogenization.

πŸ“… April 2023 πŸ“– Digital Humanities Quarterly, Vol. 17 πŸ‘οΈ 7,330 views ⬇️ 2,560 downloads πŸ“ 38 citations
Preservation Technology Published

Recovering Deprecated Web Technologies: A Framework for Executing Live Applets, MIDI Soundtracks, and Legacy JavaScript on Modern Infrastructure

Dr. Yuki Tanaka, James Chen, Dr. Hiroshi Watanabe

We present WebLegacy, a comprehensive preservation framework that enables the execution and rendering of deprecated web technologies including Java applets, ActiveX controls, MIDI audio streams, Shockwave animations, and JavaScript 1.0–1.2 code. Our system employs containerized legacy runtime environments with automated dependency resolution, allowing researchers to interact with preserved pages exactly as they originally functioned. Evaluation across 340,000 archived pages demonstrates 87% successful execution of originally interactive elements.

πŸ“… February 2023 πŸ“– IEEE Internet Computing, Vol. 27 πŸ‘οΈ 4,120 views ⬇️ 1,430 downloads πŸ“ 22 citations
Archival Science Published

The Web Ring Phenomenon: Networked Communities and Early Distributed Content Curation on the World Wide Web

Dr. Ingrid Larsson, Amara Osei, Dr. David Park

This paper provides the first comprehensive analysis of the Web Ring network system, which connected over 30,000 themed websites in a decentralized navigation structure during the mid-to-late 1990s. Using our recovered archive data, we map the topology of 4,200+ active Web Rings, quantify their cross-linking patterns, and analyze how they fostered early forms of community-driven content curation. We argue that Web Rings represent an important precursor to modern social graph architectures and decentralized content discovery systems.

πŸ“… December 2022 πŸ“– Journal of Internet Archaeology, Vol. 3 πŸ‘οΈ 3,870 views ⬇️ 1,120 downloads πŸ“ 19 citations
Early Web Studies Under Review

From Gopher to HTTP: A Comparative Analysis of Pre-Web and Early-Web Information Retrieval Systems

Dr. TomΓ‘Ε‘ NovotnΓ½, Dr. Raj Patel

This historical-comparative study examines the transition from Gopher and WAIS directory-based information systems to the link-based architecture of the early World Wide Web. Using recovered Gopher server dumps and early HTTP server logs, we analyze the technical and cultural factors that led to HTTP's dominance, documenting the lost protocols, search methodologies, and information architectures that preceded the modern web. The study provides critical context for contemporary discussions about decentralized information systems and protocol sovereignty.

πŸ“… August 2024 πŸ“– ACM Transactions on the Web (under review) πŸ‘οΈ 1,240 views ⬇️ 340 downloads πŸ“ Under review
Web Preservation Published

Policy and Practice: Building International Standards for Early Web Archival Through Multilateral Collaboration

Dr. Elena Vasquez, Dr. Chen Wei, Dr. Laura Johansson, Sofia Kim

This policy paper outlines the institutional frameworks required for international collaboration in early web preservation. Drawing on our experience coordinating with archival institutions across 38 countries, we propose a standardized metadata schema for early web content, a mutual legal assistance protocol for cross-border archival access, and a governance model for a distributed early web preservation consortium. We present case studies from our work with the Internet Archive, the British Library, and the National Diet Library of Japan.

πŸ“… May 2023 πŸ“– IFLA Journal, Vol. 49 πŸ‘οΈ 2,650 views ⬇️ 780 downloads πŸ“ 15 citations

πŸ“¬ Stay Updated

Subscribe to receive notifications when new research papers and publications are added to our archive.

"}