06145003 is referenced by 134 patents and cites 14 patents.

A computer-based system and method of retrieving information pertaining to Web documents on a computer network is disclosed. The method includes maintaining an address map that associates primary addresses with secondary addresses. A primary address includes a network retrieval protocol and a network address. The secondary address may include a different retrieval protocol or a different network address from the primary document address. A Web crawler retrieves a Web document using the primary document address, and determines whether the address map contains a secondary document address prefix corresponding to the primary document address prefix. If a secondary document address prefix exists, the Web crawler creates a secondary address, retrieves additional information pertaining to the Web document, and combines the additional information with the data retrieved from the Web document. The combined data may be stored in an index, and subsequently used to perform a document search.

Title
Method of web crawling utilizing address mapping
Application Number
8/992329
Publication Number
6145003
Application Date
December 17, 1997
Publication Date
November 7, 2000
Inventor
Dmitriy Meyerzon
Bellevue
WA, US
Sankrant Sanu
Redmond
WA, US
Agent
Christenson O Connor Johnson Kindness PLLC
Assignee
Microsoft Corporation
WA, US
IPC
G06F 15/173
View Original Source