Bit-Level Fidelity, Contextual Integrity
Digital preservation isn't about taking screenshots. It's about capturing the raw bytes, the server responses, the deprecated MIME types, and the client-side rendering quirks that defined an era.
Our methodology follows the OAIS reference model, adapted specifically for stateless, hyperlink-driven content. Every captured artifact is stored with its complete request/response chain, metadata lineage, and cryptographic verification hashes.
When a webpage vanishes, it's not just text lost. It's typography, layout logic, audio cues, and cultural context erased. We preserve the complete user experience.
How We Capture & Maintain
Our automated ingestion system runs continuously across decentralized nodes, following a strict five-stage pipeline designed for maximum fidelity and long-term viability.
Discovery & Crawl
Seed URLs are expanded using era-appropriate link parsers, following frames, tables, and deprecated navigation structures.
Artifact Capture
Full HTTP streams are recorded, including headers, cookies, and binary payloads. Assets are de-duplicated and tagged.
Normalization
Raw data is converted into WARC/Cdx standards. Metadata is extracted, dated, and cryptographically signed.
Storage & Replication
Archives are distributed across geographically dispersed cold storage with erasure coding and periodic bit-rot checks.
Emulation & Render
Historical browser engines and OS environments are maintained to ensure accurate visual and interactive playback.
Formats We Support & Preserve
📄 Document & Markup
- HTML 2.0 / 3.2 / 4.01
- SGML & Early XML
- Frame sets & Applets
- CGI & Server-side scripts
🖼️ Media & Assets
- GIF89a & Animated GIF
- PCX, TGA, RAS, BMP
- MIDI, MOD, S3M trackers
- Early Flash & Shockwave
⚙️ Protocols & Archives
- WARC / ARC / CDX Indexes
- HTTP/1.0 & FTP Streams
- Gopher & WAIS Archives
- Memento TimeGate Links
Preservation at Scale
Help Us Keep the Web Alive
Whether you're a researcher, a former webmaster, or just someone who remembers the old internet, you can contribute to our preservation efforts. Submit URLs, donate storage, or apply for API access.