Back to search

Virginia Tech

Table Understanding for Information Retrieval

Abstract

dc:description.abstract

This thesis proposes a novel approach for finding tables in text files containing a mixture of unstructured and structured text. Tables may be arbitrarily complex because the data in the tables may themselves be tables and because the grouping of data elements displayed in a table may be very complex. Although investigators have proposed competence models to explain the structure of tables, there are no computationally feasible performance models for detecting and parsing general structures in real data. Our emphasis is placed on the investigation of a new statistical procedure for detecting basic tables in plain text documents. The main task here is defining and testing this theory in the context of the Odessa Digital Library.

Degree

thesis:*
Name thesis:degree_name
Master of Science
Level thesis:degree_level
masters
Discipline thesis:degree_discipline
Computer Science
Department dc:contributor.department
Computer Science
Grantor dc:publisher
Virginia Tech
Year dc:date.issued
2002

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Pande, Ashwini K.
Chair dc:contributor.committeechair
  • Ehrich, Roger W.
Committee members dc:contributor.committeemember
  • Fox, Edward A.
  • North, Christopher L.

Subjects

dc:subject × 5

Rights

dc:rights
Statement dc:rights
  • In Copyright

Identifiers

dc:identifier.*
Dc Identifier Other
etd-08282002-151909
OAI identifier oai:identifier
oai:vtechworks.lib.vt.edu:10919/34820

Chain of custody

source
Harvested from
Virginia Tech
Base URL
vtechworks.lib.vt.edu/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
citation

Pande, Ashwini K.. Table Understanding for Information Retrieval. masters thesis, Virginia Tech, 2002. http://hdl.handle.net/10919/34820