Our R&D team develops novel algorithms for web archaeology, era-accurate rendering engines, and large-scale preservation pipelines that keep digital heritage accessible for generations.
Our R&D efforts focus on four foundational areas that advance digital preservation science.
Algorithmic reconstruction of lost URLs, dead links, and fragmented domain histories using cross-referenced wayback snapshots and DNS logs.
Custom rendering pipelines that emulate Netscape Navigator 3.0, IE4, and early WebKit engines to display pages exactly as they appeared.
Protocol-level emulation of 56k dial-up constraints, early HTTP/1.0 behaviors, and deprecated MIME types for authentic archival playback.
ML-driven trend analysis tracking the evolution of web design, technology adoption, and cultural shifts across 1990–1999 web content.
Track our open research projects, internal tools, and collaboration frameworks.
Next-generation distributed crawler with parallel snapshot resolution and automated dead-link stitching.
Automated parser that reconstructs broken frameset layouts from fragmented HTML4 archives.
Lossless archival pipeline for embedded audio/video assets using FFmpeg legacy codecs and container normalization.
Built on modern foundations, engineered for historical fidelity.
Peer-reviewed findings and open-source contributions from our lab.
J. Chen, A. Vance, L. Okoro // Digital Heritage Conference 2024
M. Rossi, S. Patel // ACM Web Archive Symposium 2023
K. Tanaka, R. Dubois // IEEE Digital Preservation Journal Vol.12
Open to academic partnerships, institutional grants, and developer contributions. Help us preserve the foundational web.
Request Research Access →