Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 36 for “"human evaluation"”.

  1. Recycling texts: human evaluation of example-based machine translation subtitles for DVD

    … this prospective profiling phase, we also elicit human judgements (through a combined questionnaire and interview) on the quality of the corpus data and on the reusability in new contexts of the TL subtitles. The intelligibility and acceptability of EBMT-produced subtitles (dependent variables) …

    dcu Repository record for Recycling texts: human evaluation of example-based machine translation subtitles for DVD (opens in a new tab)

  2. Response Generation in Longitudinal Dialogues

    … the most challenging type of conversations for human-machine dialogue systems. LDs include the recollections of events, personal thoughts, and emotions specific to each individual in a sparse sequence of dialogue sessions. Dialogue systems designed for LDs should uniquely interact with the users …

    trento Repository record for Response Generation in Longitudinal Dialogues (opens in a new tab)

  3. Component Study of Co -Word Analysis

    The results of human evaluation and statistical analyses show that both the strength index component and input term component affect the output from co-word analysis in significant ways. The resulting maps from the inclusion index test are much more general than those from the equivalence index. …

    uiuc Repository record for Component Study of Co -Word Analysis (opens in a new tab)

  4. The integration of machine translation and translation memory

    … effort. We perform both automatic and human evaluation on these models. When measured against the consensus of human judgement, the recommendation model obtains 0.91 precision at 0.93 recall, and the reranking model obtains 0.86 precision at 0.59 recall. The high precision of these …

    dcu Repository record for The integration of machine translation and translation memory (opens in a new tab)

  5. Lactose Hydrolysis by Fungal and Yeast Lactase: Influence on Freezing Point and Dipping Characteristics of Ice Cream

    … methods could be used as an alternative to human testing of dippability. In the first experiment, ice cream mixes were treated with lactase (EC 3.2.1.23) to cause 0 to 83% lactose hydrolysis. Lactose hydrolysis decreased the freezing point from -1.63oC in the control (0% hydrolysis) to …

    vt Repository record for Lactose Hydrolysis by Fungal and Yeast Lactase: Influence on Freezing Point and Dipping Characteristics of Ice Cream (opens in a new tab)

  6. Language style transfer

    … and word order recovery. In both automatic and human evaluation our method achieves strong performance.

    mit Repository record for Language style transfer (opens in a new tab)

  7. Evaluating style modification in text

    … transfer is bottlenecked by a lack of standard evaluation practices. We define three key aspects of interest (style transfer intensity, content preservation, and naturalness) and show how to obtain more reliable measures of them from human evaluation than in previous work. We also demonstrate …

    mit Repository record for Evaluating style modification in text (opens in a new tab)

  8. Evaluating Differences in GPT-4 Treatment by Gender in Healthcare Applications

    … explicit gender cues. We conduct a large-scale human evaluation of GPT-4 responses to medical questions, including counterfactual gender pairs for each question. Our findings reveal differential treatment based on the original patient gender. Specifically, responses for women more often …

    mit Repository record for Evaluating Differences in GPT-4 Treatment by Gender in Healthcare Applications (opens in a new tab)

  9. Definition modelling for English and Portuguese: a comparison between models and settings

    … provide clarity, and pervade all areas of human activity. Hence, having access to definitions is essential for many professions, but it is crucial for translation and interpreting. Definition Modelling (DM) is a task concerned with automatically generating definitions from embeddings. While …

    wlv Repository record for Definition modelling for English and Portuguese: a comparison between models and settings (opens in a new tab)

  10. An Investigation into Automatic Translation of Prepositions in IT Technical Documentation from English to Chinese

    … texts to better suit the RBMT system. Overall evaluation results (both human evaluation and automatic evaluation) show the potential of our new approaches in improving the translation of prepositions. In addition, the current study also reveals a new function of automatic metrics in assisting …

    dcu Repository record for An Investigation into Automatic Translation of Prepositions in IT Technical Documentation from English to Chinese (opens in a new tab)

  11. AI-driven optimization framework with domain-specific Large Language Models for renewable energy and hydrogen deployment

    … is validated through LLM-as-Judge and human evaluation methodologies, with RE-LLaMA significantly outperforming the base model in renewable and hydrogen energy tasks, while MG-OPT's sophisticated algorithms enable precise resource allocation in grid-connected operation, ultimately …

    uoit Repository record for AI-driven optimization framework with domain-specific Large Language Models for renewable energy and hydrogen deployment (opens in a new tab)

  12. Multi-theme sentiment analysis with sentiment shifting

    … word feature representations of reviews, and the human evaluation shows its successful discovery of multi-theme sentiment words and automatic effect quantification of contextual valence shifters.

    uiuc Repository record for Multi-theme sentiment analysis with sentiment shifting (opens in a new tab)

  13. Investigating Natural Language Interactions in Communities

    … pretrained language models. We perform extensive human evaluation and discuss potential shortcomings of our system’s generations. Second, we describe work on investigating how NLP models degrade over time due to the dynamic nature of most communities. We find evidence that static models, like …

    washington Repository record for Investigating Natural Language Interactions in Communities (opens in a new tab)

  14. Generative Chatbot Framework for Cybergrooming Prevention

    … in social media environments to interact with human users (i.e., youth) and observe the conversations that the youth users respond to strangers or acquaintances when they are asked for private or sensitive information by the perpetrator. We evaluated the quality of conversations generated by …

    vt Repository record for Generative Chatbot Framework for Cybergrooming Prevention (opens in a new tab)

  15. People-search : searching for people sharing similar interests from the web

    … to the main proposed algorithm. Two kinds of evaluations were conducted. In the automatic evaluation, precision, recall, F and Kruskal-Goodman F measures were used to compare these algorithms. In the human evaluation, the effectiveness of the main proposed algorithm and two other important …

    njit Repository record for People-search : searching for people sharing similar interests from the web (opens in a new tab)

  16. -ing words in RBMT: multilingual evaluation and exploration of pre- and post-processing solutions

    … using a customised RBMT system. A feature-based human evaluation is then performed in order to obtain information about the specific feature under study. The results showed that 73% of the -ing words were correctly translated in terms of grammaticality and accuracy for German, Japanese and …

    dcu Repository record for -ing words in RBMT: multilingual evaluation and exploration of pre- and post-processing solutions (opens in a new tab)

  17. On development and perception of theomorphic robots: the uncanny valley effect to Artificial Intelligence (AI)

    This thesis presents the development and evaluation of a theomorphic robot, the Korean Pensive Buddha Robot, a life-sized, articulated robotic statue modelled after the Korean National Treasure No. 83, the Pensive Bodhisattva. The project integrates three-dimensional (3D) printing, servomotor …

    uoit Repository record for On development and perception of theomorphic robots: the uncanny valley effect to Artificial Intelligence (AI) (opens in a new tab)

  18. A tree-to-tree model for statistical machine translation

    … model to solve this learning problem. A human evaluation of the AEP-based translation approach in a German-to-English task shows significant improvements in the grammaticality of translations. This thesis also presents a statistical parser for Spanish that could be used as part of a …

    mit Repository record for A tree-to-tree model for statistical machine translation (opens in a new tab)

  19. Corpus-based machine translation evaluation via automated error detection in output texts

    … MT has increased dramatically. Consequently, the evaluation of MT systems is crucial for all stakeholders. However, the human evaluation of MT output is expensive and time-consuming, often relying on subjective quality judgements and requiring human `reference translations' against which the …

    whiterose Repository record for Corpus-based machine translation evaluation via automated error detection in output texts (opens in a new tab)

  20. Empathetic Text Style Transfer Layer for Conversational Agents a modular architecture and new resource for the italian language

    … a central challenge in the advancement of Human-Computer Interaction (HCI), particularly in domains such as healthcare where communication requires both factual accuracy and appropriate emotional attenuation. Despite advances in natural language generation, current conversational agents …

    trento Repository record for Empathetic Text Style Transfer Layer for Conversational Agents a modular architecture and new resource for the italian language (opens in a new tab)

Page 1 of 2