Transparent architecture, open APIs, and production-ready SDKs for researchers, historians, and developers building on preserved web data.
A distributed, immutable stack engineered for scale, authenticity, and long-term digital preservation.
Custom multi-threaded spiders with protocol fallbacks for HTTP/1.0, FTP, and Gopher. Respects robots.txt historically while supporting aggressive archival modes.
Append-only object storage compliant with WARC/WACZ standards. Triple-replicated across global regions with cryptographic integrity verification.
Headless browser farm configured for legacy user-agent strings. Captures full-page screenshots, PDFs, and interactive DOM states.
Full-text and vector search across 4.2M pages. Supports temporal filtering, technology tagging, and semantic content matching.
Query, download, and reconstruct archived content programmatically. Full REST & GraphQL endpoints available.
Production-ready libraries for the most popular development environments. All SDKs share a unified type system and handle pagination, rate limits, and authentication automatically.
Everything needed to integrate, monitor, and contribute to the preservation ecosystem.
| Resource | Description | Status | Endpoint / Link |
|---|---|---|---|
| CLI Tool | Terminal interface for bulk crawling, downloading, and verifying WARC files | Stable | npm i -g @1990web/cli |
| Webhooks | Real-time notifications for crawl completion, integrity alerts, and dataset releases | Stable | /webhooks |
| GraphQL API | Flexible querying for relationships, metadata graphs, and cross-temporal analysis | Beta | api.1990archive.dev/graphql |
| WARC Validator | d>Standards compliance checker for IIPC/WARC metadata and stream integrity | Dev | github.com/1990web/validator |
We believe digital heritage should be transparent, auditable, and collaboratively improved.
Read and propose changes to our archiving protocols, metadata schemas, and API versioning policies.
Browse RFCs →Our core crawler, WARC parser, and legacy rendering modules are publicly available under MIT/Apache 2.0.
View Repositories →Request dataset access for research. We provide compute credits and dedicated archival queues for universities.
Apply for Access →Join 12k+ developers and historians. Get support, share datasets, and track infrastructure updates.
Join Community →Generate an API key, clone a repository, or join our developer program. The first decade of the web is waiting.