The Early Web: A Historical Record
Origins of the World Wide Web
On August 6, 1991, Sir Tim Berners-Lee posted a brief description of the World Wide Web project on the alt.hypertext newsgroup. That first message, sent from a NeXT computer at CERN, launched a revolution that would reshape human civilization.
The first website, hosted at http://info.cern.ch/hypertext/WWW/, explained what the World Wide Web was, how users could use it, and how to create their own web pages. It contained no images, no animations, no JavaScript — just text links on a white background. And it was breathtakingly revolutionary.
The 1990 Web Archive preserves the digital artifacts of this formative period — every surviving GeoCities page, every Angelfire personal homepage, every university server page, every early commercial website. Our collection spans the years 1990 through 1999, documenting the web's transformation from an academic tool to a global cultural phenomenon.
"If people had understood how to develop the web, as we had hoped rather than what it became, we could have gotten out of the dark ages into the information age much, much quicker."
— Tim Berners-Lee, 1999Archive Collections
Our collections are organized by era, technology platform, and cultural movement. Each collection contains fully preserved pages with original HTML, CSS, and embedded media where recoverable.
[ Best viewed in 800x600 ]
[ Netscape Now! ]
Academic & Research Pages
Early university, CERN, and NSFNET documentation pages from the pre-commercial web.
Browse collection →[ Under Construction ]
[ WebRing ▼ ]
[ Sign Guestbook ]
GeoCities Archive
182,400+ personal homepages from the GeoCities platform, preserved in their original glory.
Browse collection →║ Angelfire ║
║ Free Pages ║
╚══════════╝
Angelfire Collection
Personal and fan sites from the Angelfire free hosting platform, a major alternative to GeoCities.
Browse collection →║ Geocities ║
║ Paris ║
╚══════════════╝
Geocities Paris
French-language GeoCities pages (geocities-av.yahoo.net) with regional cultural content.
Browse collection →║ Slashdot ║
║ .org ║
║ est. 1997 ║
╚═══════════════╝
Early Commercial Sites
Dot-com era websites including Netscape, Yahoo!, Amazon, and the first search engines.
Browse collection →║ Alt.Hypertext ║
║ Berners-Lee ║
║ First WWW ║
╚═══════════════╝
The First Website
Full preservation of the original info.cern.ch HTTPd server content from Tim Berners-Lee's NeXT computer.
View preservation →Chronology of the Early Web
The table below traces the major milestones in the history of the World Wide Web as documented in our archive. Each event is tied to preserved web content.
| Date | Event | Archive ID | Tags |
|---|---|---|---|
| Aug 1991 | First website goes live at info.cern.ch | ARCH-001 | PRE-COMMERCIAL |
| Dec 1991 | Alt.hypertext newsgroup established | ARCH-014 | PRE-COMMERCIAL |
| Apr 1993 | Mosaic browser released by NCSA | ARCH-042 | BROWSER WARS |
| Apr 1994 | GeoCities founded by David Bohnett | ARCH-089 | PLATFORM |
| Sep 1994 | Yahoo! launched by Jerry Yang & David Filo | ARCH-102 | DIRECTORY |
| Jan 1995 | Amazon.com founded as "Cadabra" | ARCH-115 | E-COMMERCE |
| Sep 1995 | Angelfire launches free hosting | ARCH-134 | PLATFORM |
| Mar 1996 | Google founded by Larry & Sergey | ARCH-156 | SEARCH |
| 1997–99 | Dot-com boom peaks; GeoCities at 37M users | ARCH-201 | DOT-COM |
| Mar 2001 | GeoCities shut down by Yahoo! | ARCH-340 | LOST MEDIA |
Technologies Preserved
Our archive captures not just content but the full stack of early web technologies. Each preserved page is catalogued by its constituent technologies, allowing researchers to study the evolution of web standards. Our crawl engines can reconstruct pages with era-accurate rendering engines.
⚡ Rendering Engine Note
Pages in our archive are rendered using era-accurate browser emulation (Netscape Navigator 3.0–4.x, Internet Explorer 3.0–5.x). This ensures that table-based layouts, CSS quirks, and JavaScript behaviors are reproduced exactly as users experienced them in the 1990s.
Research & Academic Use
The 1990 Web Archive serves researchers in digital humanities, media studies, computer science, and cultural history. Our datasets are freely available under a CC BY 4.0 license for academic and non-commercial use.
Key research areas supported by our archive include:
- Evolution of web design patterns and aesthetics
- Cultural documentation of early internet communities
- Tracing the spread of web technologies across institutions
- Analysis of early e-commerce business models
- Study of lost media and digital preservation challenges
- Linguistic analysis of early internet vernacular
We partner with the Internet Archive, the British Library, and over 40 universities worldwide. Researchers can request access to our full dataset through our Academic Access Program. API access is available for automated research workflows.
Accessing the Archive
Our archive is accessible through multiple interfaces. Browse the collections below, use our search engine to find specific pages by URL, title, or era, or download our datasets for offline research.
📡 Full Dataset Download
The complete archive dataset (4.2TB compressed) is available for download via HTTP FTP at archive.1990web.archive.org/dataset/v3.2/. Torrent distribution is also available. Dataset includes raw HTML, recovered images, CSS, and metadata for every archived page.