Official public portalBDA-WEB-PUBLIC

FWX-000009

WebCrawler

WebCrawler was a full-text World Wide Web search and indexing service developed by Brian Pinkerton at the University of Washington and made publicly available in 1994. Unlike earlier web-discovery systems that indexed more limited information, WebCrawler enabled users to search for words occurring throughout indexed webpages and helped establish full-text retrieval as a fundamental model for web search.

Record Number
FWX-000009
Release
Public
Preservation Status
Archival Record
Authority
Forgotten Web Exchange

Object Registration

Digital Object ID
FWX-000009
Object Type
Web Search Service
Classification
Public
Operational Status
Superseded
Preservation Status
Archival Record
Creator
Brian Pinkerton
First Published
1994-04-20
Last Verified
2026-09-07
Historical Address
http://www.webcrawler.com/

Historical Overview

WebCrawler was a full-text World Wide Web search and indexing service developed by Brian Pinkerton at the University of Washington and made publicly available in 1994. Unlike earlier web-discovery systems that indexed more limited information, WebCrawler enabled users to search for words occurring throughout indexed webpages and helped establish full-text retrieval as a fundamental model for web search.

Current Preservation Status

The Digital Object remains Historical States Accessible Through Web Archives at a currently recorded network location.

Periodic Archival Review conducted by the Records & Provenance Division has identified the current integrity status as Original Search Infrastructure No Longer Operational.

Preservation status remains Archival Record.

Monitoring Programme
Periodic Archival Review
Integrity Status
Original Search Infrastructure No Longer Operational
Accessibility
Historical States Accessible Through Web Archives

Appears in Collections

Narrative Record

Overview

WebCrawler was developed by Brian Pinkerton at the University of Washington during 1994. The project began as a desktop application for locating information on the World Wide Web before being adapted into a publicly accessible web search service.

WebCrawler became publicly available on 20 April 1994 with an index containing material drawn from several thousand websites.

Its principal historical significance lies in full-text web search. WebCrawler enabled users to search for words occurring throughout the content of indexed webpages rather than relying solely on manually organized categories or limited document metadata.

This approach anticipated a fundamental characteristic of subsequent web search engines: automated collection and indexing of webpage content followed by user retrieval through textual queries.

WebCrawler subsequently passed through several owners, including America Online and Excite. By 2001, WebCrawler no longer operated its original independent search index and the name continued through later search implementations.

Preservation Note

A website using the WebCrawler name remains publicly accessible. The Bureau does not consider continuity of the WebCrawler brand and domain sufficient evidence that the original 1994 search system remains operational.

For catalogue purposes, this record represents the original WebCrawler full-text search and indexing service and its subsequent development as an independent search system.

Later services operating under the WebCrawler name are treated as successor implementations rather than uninterrupted instances of the Digital Object represented by this record.

Surviving historical interfaces, documentation, archived webpages, and technical records provide evidence of earlier states of the service. The Bureau therefore records the original WebCrawler system as an archival preservation case.

Publication Details

Prepared by
Office of Digital Preservation
Published through
Forgotten Web Exchange
Issuing Authority
Bureau of Digital Antiquities
Document Revision
1.0
Record Identifier
FWX-000009
Last Verification
2026-09-07

Tags

  • Search Engine
  • Web Search
  • Information Retrieval
  • Web Crawling
  • Full-Text Search
  • Early Web
  • University of Washington
  • Internet History