Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 20 of 2886 for “"benchmark"”.
-
Benchmark skyshine exposure rates
Digitized by Kansas Correctional Industries
-
Evaluate and Benchmark Aris
In this paper, we present evaluation and benchmark of Aris (Analogical Reasoning for reuse of Implementation & Specification). Aris aims to increase the number of verified programs by promotes the advantages of code reuse and the possibility of transferring specifications between similar …
-
Effect of Principal Leadership Strategies on Teachers' Use of Data in Benchmark and Non-Benchmark Middle Schools
… a greater impact on teachers' use of data in benchmark than non-benchmark schools. The purpose of the study was to determine the extent to which principal leadership strategies (independent variable) influenced teachers' use of data in benchmark and non-benchmark schools (dependent variable). …
-
SWE-Bench+: Enhanced Coding Benchmark for LLMs
… we propose SWE-Bench+, a refined version of the benchmark using two LLM-based tools: SoluLeakDetector to identify solution-leak issues and TestEnhancer to reduce weak test cases. SWE-Bench+ identifies solution-leak issues with 86% accuracy and reduces suspicious patches by 19%. To reduce the risk …
-
Benchmark indices, alpha creation and performance persistence
… empirical studies which investigate the role of benchmark indices, alpha creation and performance persistence. In the first essay, we re-visit the performance of 887 active UK equity mutual funds due to the fact that recent academic literature documents that standard benchmark models, such as FF3 …
-
Reactivity feedback analysis for EBR-II benchmark
Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2022-04-06 without embargo terms
-
Long-range Genomics Benchmark Technology and More
… mechanism, the creation of a genomics long range benchmark (GLRB), and the evaluation of various transformer and other non-transformer architectures. These efforts collectively develop the groundwork supporting the development of a robust genomics foundation model, opening new possibilities for …
-
Large Language Model Routing with Benchmark Datasets
… of open-source Large Language Models (LLMs) and benchmark datasets to compare them. While some models dominate these benchmarks, no single model typically achieves the best accuracy in all tasks and use cases. With a new dataset, it can be difficult to determine which LLM is best suited to the …
-
Synthesizing a Hybrid Benchmark Suite with BenchPrime
This paper presents BenchPrime, an automated benchmark analysis toolset that is systematic and extensible to analyze the similarity and diversity of benchmark suites. BenchPrime takes multiple benchmark suites and their evaluation metrics as inputs and generates a hybrid benchmark suite comprising …
-
A benchmark for impact assessment of affordable housing
… in the built environment for the significance of benchmarking. It is recognized as a key driver for measuring success criteria in the built environment sector. In spite of the huge application of this technique to the sector and other sectors, very little is known of it in affordable housing …
-
Robustness of the achievable benchmark of care method
<p>The Achievable Benchmark of Care Method is a process of care performance improvement measurement approach for identifying top performing healthcare providers. The purpose of this study was to investigate the robustness of the method. This was achieved by comparing the robustness of the standard …
-
DOCKGROUND Protein Docking Benchmark Sets and Assessment Resource
… of docking methods, high quality datasets and benchmarking tools are needed. This work provides such datasets, and a tool for benchmarking protein docking methods. The datasets are the protein-protein bound dataset, and the protein-RNA bound dataset. Both sets automatically update regularly on …
-
STREETS: a benchmark dataset for suburban traffic forecasting
In this work, we introduce and benchmark STREETS, a novel traffic flow dataset from publicly available web cameras in the suburbs of Chicago, IL. STREETS addresses multiple limitations of existing vehicular traffic datasets. Many current datasets lack a coherent traffic network graph to describe …
-
VisText: A Benchmark for Semantically Rich Chart Captioning
Captions that describe or explain charts help improve recall and comprehension of the depicted data and provide a more accessible medium for people with visual disabilities. However, current approaches for automatically generating such captions struggle to articulate the perceptual or cognitive …
-
Finite element comparison for a geologically motivated benchmark
Geologic deformation in three dimensions can be modeled using finite element analysis. In choosing the elements used to solve a model it is important to consider the accuracy of the solution and the computational intensity. The results for models using six element types and six element side lengths …
-
Minimum Variance Benchmark and Performance Assessment for PID Controllers
… An explicit explanation for the minimum variance benchmark is proposed to be applicable in industrial cases as the current online performance monitoring methods for PID controllers were inadequate. Control performance assessment (CPA) compares the actual output variance with the minimum variance …
-
Machine learning for selecting parallel I/O benchmark applications
… applications. Accurate I/O performance benchmarking, which can help us better understand the causes of these bottlenecks and to guide the performance optimization of poor performing applications, is therefore an important problem. We investigate the use of submodular function …
-
Grounded SCAN Human: A Benchmark for Zero-Shot Generalizations
In this work, we collect a new human annotated dataset called Grounded SCAN Human (gSCAN Human) as an extension of the original Grounded SCAN (gSCAN) dataset. The original gSCAN dataset was created to test various compositional generalizations by holding out certain examples during train time. …
-
The Extreme Benchmark Suite : measuring high-performance embedded systems
The Extreme Benchmark Suite (XBS) is designed to support performance measurement of highly parallel "extreme" processors, many of which are designed to replace custom hardware implementations. XBS is designed to avoid many of the problems that occur when using existing benchmark suites with …
-
A High Performance C++ Generic Benchmark for Computational Epidemiology
… resource intensive. In this work, we design a benchmark consisting of several kernels which capture the essential compute, communication, and data access patterns for such applications. For each kernel, the benchmark provides different evaluation strategies. The goal is to (a) derive …
Page 1 of 145