Official public portalBDA-WEB-PUBLIC

COL-000002

Early Web Infrastructure

A curated collection examining foundational systems through which the early World Wide Web was published, organized, discovered, and searched.

Collection ID
COL-000002
Status
Current
Published
2026-09-07
Digital Objects
3
Publications
0

Curatorial Record

Overview

The early World Wide Web required more than the ability to publish interconnected documents. As the number of available websites expanded, users also required mechanisms for locating, organizing, and retrieving information distributed across the network.

This collection brings together three Digital Objects representing different stages in the development of that environment: the first website at CERN, Yahoo! Directory, and WebCrawler.

Together they document the emergence of web publishing and two influential approaches to the subsequent problem of information discovery.

From Publication to Discovery

The first website at CERN represents the environment in which the World Wide Web itself was developed. It demonstrated a system of interconnected hypertext resources accessible through common network protocols and provided information concerning the developing Web project.

As adoption of the Web increased, the expanding number of available resources created a new problem: discovery.

Yahoo! Directory addressed this problem through human organization. Websites were selected and arranged within a hierarchy of subjects and subcategories through which users could browse.

WebCrawler represented a different approach. Automated collection and full-text indexing allowed users to locate webpages by searching for words contained within their content.

These approaches — hierarchical browsing and automated textual retrieval — represent important stages in the development of mechanisms through which users navigated the increasingly complex information environment of the World Wide Web.

Curatorial Note

The Digital Objects within this collection are not presented as a complete history of early web infrastructure or information retrieval.

Instead, they have been selected to illustrate a historical progression from the establishment of the Web as a publishing environment to the development of systems intended to make its rapidly expanding information space navigable.

The relationship between Yahoo! Directory and WebCrawler is particularly significant. Developed during the same period, they demonstrate contrasting responses to the same underlying problem: how to find useful information on a Web growing beyond the scale at which individual resources could be discovered through direct knowledge and manually maintained links alone.

The collection therefore interprets these Digital Objects not merely as surviving early Internet services, but as evidence of the evolving information architecture of the public Web.

Origins of the Digital Objects

First-publication dates recorded in FWX, from earliest to latest. Date precision follows each record.

  1. 1990
    The First Website — CERN

    FWX-000008 · First published

  2. 1994
    Yahoo! Directory

    FWX-000006 · First published

  3. 1994-04-20
    WebCrawler

    FWX-000009 · First published

No supporting publications are currently available for this collection.

Tags

  • Early Web
  • World Wide Web
  • Information Retrieval
  • Web Discovery
  • Internet Infrastructure