Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 6432 for “"documents"”.

  1. XML documents schema design

    … for defining the syntax and structure of XML documents. To enable efficient usage of XML documents in any application in large scale electronic environment, it is necessary to avoid data redundancies and update anomalies. Redundancy and anomalies in XML documents can lead not only to higher …

    hull Repository record for XML documents schema design (opens in a new tab)

  2. Semantic Processing of Digital Documents

    … applications like verifying the consistency of documents. Creating such models for common documents is currently an expensive and error-prone process. In this thesis we present a novel approach to modelling and processing digital documents that uses semantic technologies. In contrast to other …

    passau-thes Repository record for Semantic Processing of Digital Documents (opens in a new tab)

  3. Efficient processing of XML documents

    In this thesis, we advocate storing XML documents in a relational DBMS, and address the related challenges. In particular, we set out to address the issues of mapping, indexing and updating XML documents.

    nus Repository record for Efficient processing of XML documents (opens in a new tab)

  4. Annotation persistence over dynamic documents

    … in the paper world to augment the usefulness of documents. By annotation, we include a large variety of creative manipulations by which the otherwise passive reader becomes actively involved in a document. Annotations in digital form possess many benefits paper annotations do not enjoy, such as …

    mit Repository record for Annotation persistence over dynamic documents (opens in a new tab)

  5. Machine learning on Web documents

    … news, and automatically sorting and compiling documents. We adapt and create machine learning algorithms for use with the Web's distinctive structures: large-scale, noisy, varied data with potentially rich, human-oriented features. We adapt two standard classification algorithms, the slow but …

    mit Repository record for Machine learning on Web documents (opens in a new tab)

  6. Automatically Extract Information from Web Documents

    The Internet could be considered to be a reservoir of useful information in textual form — product catalogs, airline schedules, stock market quotations, weather forecast etc. There has been much interest in building systems that gather such information on a user's behalf. But because these …

    wku-diss Repository record for Automatically Extract Information from Web Documents (opens in a new tab)

  7. Managing the consistency of distributed documents

    Many businesses produce documents as part of their daily activities: software engineers produce requirements specifications, design models, source code, build scripts and more; business analysts produce glossaries, use cases, organisation charts, and domain ontology models; service providers and …

    ucl Repository record for Managing the consistency of distributed documents (opens in a new tab)

  8. Complying shipping documents under UCP 600

    … against the backdrop of the question: ‘what documents must a beneficiary, acting as seller under an international sale of goods carried by sea, present to a bank, and how must he present them, in order for the presentation to be considered compliant?’. It interprets the rules through the …

    soton Repository record for Complying shipping documents under UCP 600 (opens in a new tab)

  9. Geometric correction of historical Arabic documents

    Geometric deformations in historical documents significantly influence the success of both Optical Character Recognition (OCR) techniques and human readability. They may have been introduced at any time during the life cycle of a document, from when it was first printed to the time it was digitised …

    salford Repository record for Geometric correction of historical Arabic documents (opens in a new tab)

  10. Safe Template Processing of XML Documents

    Templates sind eine etablierte Technik zur Zusammenführung verschiedener Datenquellen, welche zum Zwecke der Separation of Concerns getrennt gehalten werden sollen. Bei der Benutzung bestehender Ansätze tritt das Problem auf, dass zur Zeit der Erstellung eines Templates keine Aussagen über die …

    qucosa-diss

  11. Topic specific spider for taxonomic documents

    … to develop a subsystem that collects taxonomic documents available on the World Wide Web using a combination of spidering and document classification techniques. To increase the number of documents collected, two query expansion techniques have been considered and evaluated. We found that …

    ku Repository record for Topic specific spider for taxonomic documents (opens in a new tab)

  12. Model-based identification of Oriental documents

    … capability of identifying languages printed in documents can support many potential applications including document classification for character recognition, translation, and language understanding. Language identification is normally done manually. However, the high volume and variety of …

    concordia Repository record for Model-based identification of Oriental documents (opens in a new tab)

  13. Automatic classification of multi-lingual documents

    … (LC) refers to the categorization of text documents into different natural language groups, whereas language identification (LI) determines the language used in a document. LC and LI play important roles in document processing systems, because they can perform initial classifications to …

    concordia Repository record for Automatic classification of multi-lingual documents (opens in a new tab)

  14. Natural language search of structured documents

    … around the problem of searching through XML documents, each of which describes the play-by-play events of a baseball game. These events are collected from Major League Baseball games between 2004 and 2008, containing information detailing the outcome of every pitch thrown. My techniques are …

    mit Repository record for Natural language search of structured documents (opens in a new tab)

  15. Content-based indexing of low resolution documents

    … the methods used for the purpose of identifying documents that are captured using image capturing devices. In addition, the thesis also concerns with a technique that can be used to retrieve images from an indexed image database. Both concerns above apply digital image processing technique. To …

    uthm Repository record for Content-based indexing of low resolution documents (opens in a new tab)

  16. A Common Representation Format for Multimedia Documents

    Multimedia documents are composed of multiple file format combinations, such as image and text, image and sound, or image, text and sound. The type of multimedia document determines the form of analysis for knowledge architecture design and retrieval methods. Over the last few decades, theories of …

    unt Repository record for A Common Representation Format for Multimedia Documents (opens in a new tab)

  17. Taintx: A System for Protecting Sensitive Documents

    … and have access to important company documents. There have been several studies suggesting that employees are taking critical information after learning they will be laid off. This becomes an issue and a threat to a corporation's security. Corporations are then placed in a position to …

    uno Repository record for Taintx: A System for Protecting Sensitive Documents (opens in a new tab)

  18. Effect of OCR errors on short documents

    … is a study of the effect of OCR errors on short documents. OCR recognizes and translates text image into ASCII format. When this data is retrieved in response to a query, the retrieval performance depends on the efficiency of the OCR device used. Measures like recall, precision and ranking were …

    unlv Repository record for Effect of OCR errors on short documents (opens in a new tab)

  19. Enhancing retrieval and discovery of desktop documents

    … Most of this information is in the form of documents in files organized in the hierarchical folder structures provided by the operating system. Operating system-provided access to these data is mainly through structure-guided navigation, and more recently through keyword …

    soton Repository record for Enhancing retrieval and discovery of desktop documents (opens in a new tab)

Page 1 of 322