Learning the process of World Wide Web data retrieval
Name
62277965-MIT.pdf
Description
Full printable version
Size
4.98 MB
Format
Adobe PDF
Checksum (MD5)
13a9c96fe2ee38e7c49bac4b593ff3ab
Author(s)
Manuel, Ryan A
Advisor(s)
David R. Karger.
Alternative Title
Learning the process of WWW data retrieval
Date Issued
2005
Publisher
Massachusetts Institute of Technology
Abstract
We develop a method for extracting and internalizing web site form submissions which we refer to as web operations. To begin the process, a user performs a sample submission of the form. From that submission, our system determines all of the necessary information to store the web operation. Through a simple user interface the user can view and modify the web operation to the extent that he wants or needs to. With the operation now stored, the user can invoke the operation without browsing to the web site on which the operation was originally contained. By utilizing the web site information extraction techniques contained in the Haystack information management system, we give the user the option to extract information off of web operation results pages. Thus, when using our system to the fullest extent, a user can invoke web operations and view and make use of the results without viewing any web pages.
Description
Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 2005.
Includes bibliographical references (leaf 65).
Subjects
Electrical Engineering and Computer Science.
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission.
Persistent DSpace Link