The Early Web: A Historical Record

Origins of the World Wide Web

On August 6, 1991, Sir Tim Berners-Lee posted a brief description of the World Wide Web project on the alt.hypertext newsgroup. That first message, sent from a NeXT computer at CERN, launched a revolution that would reshape human civilization.

The first website, hosted at http://info.cern.ch/hypertext/WWW/, explained what the World Wide Web was, how users could use it, and how to create their own web pages. It contained no images, no animations, no JavaScript — just text links on a white background. And it was breathtakingly revolutionary.

The 1990 Web Archive preserves the digital artifacts of this formative period — every surviving GeoCities page, every Angelfire personal homepage, every university server page, every early commercial website. Our collection spans the years 1990 through 1999, documenting the web's transformation from an academic tool to a global cultural phenomenon.

"If people had understood how to develop the web, as we had hoped rather than what it became, we could have gotten out of the dark ages into the information age much, much quicker."

— Tim Berners-Lee, 1999

Archive Collections

Our collections are organized by era, technology platform, and cultural movement. Each collection contains fully preserved pages with original HTML, CSS, and embedded media where recoverable.

[ Welcome to My Page ]
[ Best viewed in 800x600 ]
[ Netscape Now! ]
1990–1993

Academic & Research Pages

Early university, CERN, and NSFNET documentation pages from the pre-commercial web.

Browse collection →
★ GeoCities ★
[ Under Construction ]
[ WebRing ▼ ]
[ Sign Guestbook ]
1994–1999

GeoCities Archive

182,400+ personal homepages from the GeoCities platform, preserved in their original glory.

Browse collection →
╔══════════╗
║ Angelfire ║
║ Free Pages ║
╚══════════╝
1995–2001

Angelfire Collection

Personal and fan sites from the Angelfire free hosting platform, a major alternative to GeoCities.

Browse collection →
╔══════════════╗
║ Geocities ║
║ Paris ║
╚══════════════╝
1996–2001

Geocities Paris

French-language GeoCities pages (geocities-av.yahoo.net) with regional cultural content.

Browse collection →
╔═══════════════╗
║ Slashdot ║
║ .org ║
║ est. 1997 ║
╚═══════════════╝
1996–2000

Early Commercial Sites

Dot-com era websites including Netscape, Yahoo!, Amazon, and the first search engines.

Browse collection →
╔═══════════════╗
║ Alt.Hypertext ║
║ Berners-Lee ║
║ First WWW ║
╚═══════════════╝
1991

The First Website

Full preservation of the original info.cern.ch HTTPd server content from Tim Berners-Lee's NeXT computer.

View preservation →

Chronology of the Early Web

The table below traces the major milestones in the history of the World Wide Web as documented in our archive. Each event is tied to preserved web content.

Date Event Archive ID Tags
Aug 1991 First website goes live at info.cern.ch ARCH-001 PRE-COMMERCIAL
Dec 1991 Alt.hypertext newsgroup established ARCH-014 PRE-COMMERCIAL
Apr 1993 Mosaic browser released by NCSA ARCH-042 BROWSER WARS
Apr 1994 GeoCities founded by David Bohnett ARCH-089 PLATFORM
Sep 1994 Yahoo! launched by Jerry Yang & David Filo ARCH-102 DIRECTORY
Jan 1995 Amazon.com founded as "Cadabra" ARCH-115 E-COMMERCE
Sep 1995 Angelfire launches free hosting ARCH-134 PLATFORM
Mar 1996 Google founded by Larry & Sergey ARCH-156 SEARCH
1997–99 Dot-com boom peaks; GeoCities at 37M users ARCH-201 DOT-COM
Mar 2001 GeoCities shut down by Yahoo! ARCH-340 LOST MEDIA

Technologies Preserved

Our archive captures not just content but the full stack of early web technologies. Each preserved page is catalogued by its constituent technologies, allowing researchers to study the evolution of web standards. Our crawl engines can reconstruct pages with era-accurate rendering engines.

archive-tech-query — 1990webarchive
$ archive list --technology "1990s" --all

[RESULTS] Technologies detected in archive:
├─ HTML 1.0–4.01        4,217,893 pages
├─ CSS 1.0–2.0          892,104 pages
├─ JavaScript 1.0–1.5    156,892 pages
├─ CGI/Bash scripts       89,441 pages
├─ Flash 1.0–3.0         234,567 pages
├─ MIDI/SMF music        421,003 pages
├─ Animated GIF          1,892,334 pages
├─ Java Applets           78,234 pages
├─ DHTML/frames          345,892 pages
├─ VBScript               12,445 pages
└─ PHP/ASP                267,891 pages

[OK] Query complete. 4,217,893 total entries.
[WARN] 14.2% of pages contain broken image references
[WARN] 23.7% reference dead CGI scripts

$ echo "Preservation in progress..."
Preservation in progress...
$

⚡ Rendering Engine Note

Pages in our archive are rendered using era-accurate browser emulation (Netscape Navigator 3.0–4.x, Internet Explorer 3.0–5.x). This ensures that table-based layouts, CSS quirks, and JavaScript behaviors are reproduced exactly as users experienced them in the 1990s.

Research & Academic Use

The 1990 Web Archive serves researchers in digital humanities, media studies, computer science, and cultural history. Our datasets are freely available under a CC BY 4.0 license for academic and non-commercial use.

Key research areas supported by our archive include:

  • Evolution of web design patterns and aesthetics
  • Cultural documentation of early internet communities
  • Tracing the spread of web technologies across institutions
  • Analysis of early e-commerce business models
  • Study of lost media and digital preservation challenges
  • Linguistic analysis of early internet vernacular

We partner with the Internet Archive, the British Library, and over 40 universities worldwide. Researchers can request access to our full dataset through our Academic Access Program. API access is available for automated research workflows.


Accessing the Archive

Our archive is accessible through multiple interfaces. Browse the collections below, use our search engine to find specific pages by URL, title, or era, or download our datasets for offline research.

4.2M
Pages Archived
182K
GeoCities Pages
37K
Angelfire Pages
1990
Earliest Source

📡 Full Dataset Download

The complete archive dataset (4.2TB compressed) is available for download via HTTP FTP at archive.1990web.archive.org/dataset/v3.2/. Torrent distribution is also available. Dataset includes raw HTML, recovered images, CSS, and metadata for every archived page.