Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 9 of 9 for “"document similarity"”.
-
Representing semantic relatedness
… question we must address is how to represent documents. The way a document is organised reflects certain explicit and implicit semantic and syntactical coupling relationships which are embedded in its contents. The effective capturing of such content couplings is thereby crucial for a genuine …
-
Interaction harvesting for document retrieval
… to provide meaningful search terms for non-text documents. Unfortunately, such systems usually require the author to enter the keywords manually, a task that is commonly neglected, or is executed poorly. This thesis proposes an approach to document categorization called Interaction Harvesting, …
-
An automated framework for problem report triage in large-scale open source problem repositories
… fully automated framework that utilizes multiple document similarity measures, summary statistics describing each report and user behavior attributes to determine if the problem at hand is new or duplicate. The framework relies on making as few assumptions as possible on the data in order to …
-
Beyond pre-training: continual learning and hallucinations in transformer-based language models.
… norm for various language modelling tasks from document similarity analysis and text classification to natural language generation. Despite the impressive performance on benchmark datasets, adopting pretrained models for real-world applications often requires additional stages of model …
-
Scalability of Stepping Stones and Pathways
… it, and returns the result as a ranked list of documents. However, this approach is not the most effective method to support the task of finding document associations (relationships between concepts or queries) both for direct or indirect relationships. The Stepping Stones and Pathways (SSP) …
-
The Cluster Hypothesis: A Visual/Statistical Analysis
… judgments based on a small number of exemplar documents to be applied to a larger number of unexamined documents, clustered presentation of search results represents an intuitively attractive possibility for reducing the cognitive resource demands on human users of information retrieval …
-
State Political Parties in American Politics: Innovation and Integration in the Party System
… why state platforms vary in their degree of similarity to the national platform. I analyze an extensive platform dataset, using cluster analysis and document similarity measures to compare platform content across the 1952 to 2014 period. The analysis shows that, as a group, Democratic and …
-
Role of semantic indexing for text classification.
… semantic relatedness between terms means that document similarity is not properly captured in the VSM. To address this problem, semantic indexing approaches have been proposed for modelling the semantic relatedness between terms in document representations. Accordingly, in this thesis, we …
-
Educational System for Recommending Study Activities
… je využíváno metod Term Frequency - Inverse Document Frequency a vnoření slov.Pro používání modulu a jeho komunikaci modulu s rozhraním OU Analyse je implementováno RESTful API.