Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 20 of 118 for “"Evaluation metrics"”.
-
Safe shared autonomy and evaluation metrics
Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2024-05-01
-
A Test Suite for Saliency Method Evaluation Metrics
… processes by providing a comprehensive evaluation platform for saliency methods.
-
Multi-Dimensional Evaluation Metrics for Chest X-Ray Reports
… accuracy is the most important factor. Current evaluation metrics evaluate reports in one dimension. This work proposes the use of multiple dimensions (factual correctness, comprehensiveness, style, and overall quality) to better capture evaluation preferences of a clinical text generating model …
-
Wingsail design methodology and performance evaluation metrics for autonomous sailing
This dissertation explores an innovative approach to the design of aerodynamically-actuated wingsails, along with advancements in marine vehicle autonomy focused on vessel tracking and collision avoidance. The first segment of this research introduces a deterministic wingsail design optimization …
-
Transformation Tolerance of Facial Recognition Technology and Informative Evaluation Metrics
… methodological inspiration for more informative evaluation techniques, including the characterization of recognition performance as a function of a quantifiable input transformation. This work performs such an analysis. The performance scores of five state-of-the-art FR models are characterized …
-
Wingsail Design Methodology and Performance Evaluation Metrics for Autonomous Sailing
This dissertation explores an innovative approach to the design of aerodynamically-actuated wingsails, along with advancements in marine vehicle autonomy focused on vessel tracking and collision avoidance. The first segment of this research introduces a deterministic wingsail design optimization …
-
Fault Attacks on Cryptosystems: Novel Threat Models, Countermeasures and Evaluation Metrics
… secure systems against fault attacks • Building evaluation metrics to measure the security of designs In this regard, my thesis has covered the first bullet, by proposing the Differential Fault Intensity Analysis (DFIA) based on the biased fault model. The biased fault model in this attack means …
-
Evaluation of Automatic Text Summarization Using Synthetic Facts
… However, the unreliability of the existing evaluation metrics hinders its practical usage and slows down its progress. To address this issue, we propose an automatic reference-less text summarization evaluation system with dynamically generated synthetic facts. We hypothesize that if a …
-
Evaluating websites using a practical quality model.
Many of the existing website evaluation methods and criteria for evaluating website quality are not able to sufficiently assess the performance and quality of a website, and most of them focus on usability and accessibility. This thesis aims at proposing the website quality metrics and methods to …
-
A novel dependency-based evaluation metric for machine translation
Automatic evaluation measures such as BLEU (Papineni et al. (2002)) and NIST (Doddington (2002)) are indispensable in the development of Machine Translation (MT) systems, because they allow MT developers to conduct frequent, fast, and cost-effective evaluations of their evolving translation models. …
-
A Systematic Evaluation of Climate Services and Decision Support Tools for Climate Change Adaptation
… support tools based on previously proposed evaluation framework. This evaluation framework originally consists of four design elements which are divided into nine evaluation metrics for this study. These evaluation metrics are: identification of decision making context, discussion of the …
-
Code Summarization and Program Synthesis with Large Language Models
… against appropriate baselines using suitable evaluation metrics.
-
A comparative analysis of machine learning models for forecasting JSE Stock Returns
… variables comprise nine firm-specific financial metrics, motivated by prior research. The sample is divided into a training period (2005–2016) and a testing period (2016–2021), further split into 1-year, 3-year, and 5-year testing intervals. The results show that the LSTM model performsbest …
-
Automated Self-Admitted Technical Debt Tracking, Classification, and Repayment for Sustainable Software Development
… for Java) and introducing three new, effective evaluation metrics: BLEU-diff, CrystalBLEU-diff, and LEMOD. Using our new benchmarks and evaluation metrics, we evaluate two types of automated SATD repayment methods: fine-tuning smaller models, and prompt engineering with five large-scale models. …
-
Multimodal machine translation
… tested on multiple datasets using the automatic evaluation metrics like METEOR and BLEU. The experiments show that the proposed models outperform the text-only baseline model.
-
Automated Discovery of Big Data Workload Types
… techniques are compared in terms of different evaluation metrics, and the ones with the highest performance are introduced. The DBSCAN algorithm has shown the best performance and adequacy with 71% and 80% for the Purity, and Windows Type Accuracy (Awt), respectively. Ultimately, the …
-
TechCommix: A Tool and Foundation for Rethinking and Restructuring Technical Documentation
… and step-oriented, in terms of understanding and evaluation metrics, including those related to user experience. Results indicate that comics as a documentation style can offer enhanced, more positive user experiences, albeit not being overall better than the other styles.
-
Evaluating, Understanding, and Mitigating Unfairness in Recommender Systems
… In particular, we study appropriate unfairness evaluation metrics, examine the relation between bias in recommender models and inequality in the underlying population, as well as propose effective unfairness mitigation approaches. We start with exploring the implication of fairness in …
-
Enhancing Telecom Churn Prediction: Adaboost with Oversampling and Recursive Feature Elimination Approach
… It achieves the highest rank in all three evaluation metrics: recall (0.841), f1-score (0.655), and roc_auc (0.793), further indicating that the proposed approach effectively predicts churn and provides valuable insights into customer behavior.</p>
-
Visual questioning agents
… visual dialog. We propose novel approaches and evaluation metrics for these tasks. For visual question generation, we combined language models with variational autoencoders to enhance diversity in text generations. We also suggest diversity metrics to quantify these improvements. For visual …
Page 1 of 6