Back to results

Brigham Young University - Provo

Query Rewriting for Extracting Data behind HTML Forms

Abstract

dc:description.abstract

Much of the information on the Web is stored in specialized searchable databases and can only be accessed by interacting with a form or a series of forms. As a result, enabling automated agents and Web crawlers to interact with form-based interfaces designed primarily for humans is of great value. This thesis describes a system that can fill out Web forms automatically according to a given user query against a global schema for an application domain and, to the extent possible, extract just the relevant data behind these Web forms. Experimental results on two application domains show that the approach is reasonable for HTML forms.

Degree

thesis:*
Name thesis:degree_name
MS
Grantor dc:publisher
Brigham Young University - Provo

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Chen, Xueqi

Subjects

dc:subject × 4

Rights

Language dc:language
English

Identifiers

dc:identifier.*
Repository record dc:identifier
https://scholarsarchive.byu.edu/etd/25
OAI identifier oai:identifier
oai:scholarsarchive.byu.edu:etd-1024

Chain of custody

source
Harvested from
Brigham Young University
Base URL
scholarsarchive.byu.edu/do/oai/
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
citation

Chen, Xueqi. Query Rewriting for Extracting Data behind HTML Forms. Brigham Young University - Provo, https://scholarsarchive.byu.edu/etd/25