Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 397 for “"Corpora"”.

  1. Extracting paraphrases from aligned corpora

    Thesis (M.Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 2002.

    mit Repository record for Extracting paraphrases from aligned corpora (opens in a new tab)

  2. Phonetic Transcriptions of Large Speech Corpora

    Contains fulltext : 27415.pdf (Publisher’s version ) (Open Access) Contains fulltext : 41404.pdf (Publisher’s version ) (Open Access)

    radboud Repository record for Phonetic Transcriptions of Large Speech Corpora (opens in a new tab)

  3. Discourse models for collaboratively edited corpora

    … discourse models for collaboratively edited corpora. Due to the exponential growth rate and significant stylistic and content variations of collaboratively edited corpora, models based on professionally edited texts are incapable of processing the new data effectively. For these methods to …

    mit Repository record for Discourse models for collaboratively edited corpora (opens in a new tab)

  4. Using corpora to aid in learning collocations

    The potential of corpora, language databases comprised of authentic language materials from a variety of sources, has gradually trickled down to ESL and EFL classrooms (McCarthy, O'Keeffe, & Walsh, 2010) and has been associated with data-driven learning (DDL) where learners observe language …

    uiuc Repository record for Using corpora to aid in learning collocations (opens in a new tab)

  5. Annotation-free location mention mining from text corpora

    Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2025-02-04 without embargo terms

    uiuc Repository record for Annotation-free location mention mining from text corpora (opens in a new tab)

  6. Annotation-free knowledge mining from massive text corpora

    Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2022-11-11 without embargo terms

    uiuc Repository record for Annotation-free knowledge mining from massive text corpora (opens in a new tab)

  7. Autoentity: automated entity detection from massive text corpora

    Entity detection is one of the fundamental tasks in Natural Language Processing and Information Retrieval. Most existing methods rely on human annotated data and hand-crafted linguistic features, which makes it hard to apply the model to an emerging domain. In this paper, we propose a novel …

    uiuc Repository record for Autoentity: automated entity detection from massive text corpora (opens in a new tab)

  8. Efficient near duplicate document detection for specialized corpora

    Knowledge of near duplicate documents can be adventagous to search engines, even those that only cover a small enterprise or specialized corpus. In this thesis, we investigate improvements to simhash, a signature-based method which can be used to efficiently detect near duplicate documents. We …

    mit Repository record for Efficient near duplicate document detection for specialized corpora (opens in a new tab)

  9. The automatic extraction of linguistic information from text corpora

    This is a study exploring the feasibility of a fully automated analysis of linguistic data. It identifies a requirement for large-scale investigations, which cannot be done manually by a human researcher. Instead, methods from natural language processing are suggested as a way to analyse large …

    birmingham Repository record for The automatic extraction of linguistic information from text corpora (opens in a new tab)

  10. Rule based learning of word pronunciations from training corpora

    This paper describes a text-to-pronunciation system using transformation-based error-driven learning for speech-recognition purposes. Efforts have been made to make the system language independent, automatic, robust and able to generate multiple pronunciations. The learner proposes initial …

    mit Repository record for Rule based learning of word pronunciations from training corpora (opens in a new tab)

  11. When more is less : identifying biases in large Icelandic corpora

    … devices, in their own language. Linguistic corpora are used for this purpose, to train language models which can predict and model human languages. Various corpora have been assembled for Icelandic and language models have been trained on them. These models predict words and sentences, and …

    reykjavik Repository record for When more is less : identifying biases in large Icelandic corpora (opens in a new tab)

  12. Using Corpora with Taiwanese college students in a Student-centred

    Previous studies show that corpora are helpful to translation teaching and learning in numerous ways; however, the students’ use of and attitudes towards corpus-assisted translation are seldom discussed. This research addresses the following two issues regarding the implementation of a …

    qu-belfast Repository record for Using Corpora with Taiwanese college students in a Student-centred (opens in a new tab)

  13. Enhancing Competitive Intelligence Using Geospatial Analysis and Social-Media Corpora

    … analysis, which is an essential element of corporate strategy, is a technique used to assess competitor (current and future) threats. Firms cannot ignore their competitors. The challenges of keeping their competitive advantage and increasing growth can be a difficult task for businesses when …

    claremont Repository record for Enhancing Competitive Intelligence Using Geospatial Analysis and Social-Media Corpora (opens in a new tab)

  14. TextDive: construction, summarization and exploration of multi-dimensional text corpora

    … techniques. In our view, documents in text corpora contain informative explicit meta-attributes (e.g., category, date, author, etc.) and implicit attributes (e.g., sentiment), forming one or a set of highly-structured multi-dimensional spaces. Much knowledge can be derived if we develop …

    uiuc Repository record for TextDive: construction, summarization and exploration of multi-dimensional text corpora (opens in a new tab)

  15. A Comparative Study of Mechanisms of Maintenance of Corpora Lutea

    Made available in DSpace on 2014-12-05T17:47:39Z (GMT). No. of bitstreams: 1 6000233.pdf: 3523564 bytes, checksum: 2ffc4f50f02184e990a891c1496f36c7 (MD5) Previous issue date: 1959

    uiuc Repository record for A Comparative Study of Mechanisms of Maintenance of Corpora Lutea (opens in a new tab)

  16. The Formation and Maintenance of Corpora Lutea in Cycling Gilts

    Made available in DSpace on 2014-12-05T17:47:51Z (GMT). No. of bitstreams: 1 6305075.pdf: 2817449 bytes, checksum: 7b5b0420ba1b93949c093401468a937a (MD5) Previous issue date: 1963

    uiuc Repository record for The Formation and Maintenance of Corpora Lutea in Cycling Gilts (opens in a new tab)

  17. Patent semantics : analysis, search and visualization of large text corpora

    Patent Semantics is system for processing text documents by extracting features capturing their semantic content, and searching, clustering, and relating them by those same features. It is set apart from existing methodologies by combining a visualization scheme that integrates retrieval and …

    mit Repository record for Patent semantics : analysis, search and visualization of large text corpora (opens in a new tab)

  18. A framework for automated landmark recognition in community contributed image corpora

    Any large library of information requires efficient ways to organise it and methods that allow people to access information efficiently and collections of digital images are no exception. Automatically creating high-level semantic tags based on image content is difficult, if not impossible to …

    dcu Repository record for A framework for automated landmark recognition in community contributed image corpora (opens in a new tab)

Page 1 of 20