Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 60 for “"hardware accelerators"”.

  1. Exocompilation for Productive Programming of Hardware Accelerators

    … kernel libraries are critical to exploiting accelerators and specialized instructions in many applications. Because compilers are difficult to extend to support diverse and rapidly-evolving hardware targets, and automatic optimization is often insufficient to guarantee state-of-the-art …

    mit Repository record for Exocompilation for Productive Programming of Hardware Accelerators (opens in a new tab)

  2. Designing Hardware Accelerators for Solving Sparse Linear Systems

    … optimizing linear solvers for high performance hardware. However, despite their efforts, existing hardware has let them down. State-of-the-art linear solvers often utilize < 1% of available compute throughput on existing architectures such as CPUs and GPUs. There are many different algorithms …

    mit Repository record for Designing Hardware Accelerators for Solving Sparse Linear Systems (opens in a new tab)

  3. Evaluating the Feasibility of Transaction Scheduling via Hardware Accelerators

    … of accelerating transaction scheduling via hardware, leveraging FPGAs to ofoad scheduling logic from the CPU. We revisit Puppetmaster, a hardware transaction scheduler, and present a redesigned architecture emphasizing deployability, modularity, and evaluation. We implement both an optimized …

    mit Repository record for Evaluating the Feasibility of Transaction Scheduling via Hardware Accelerators (opens in a new tab)

  4. Energy Efficient Hardware Accelerators for Packet Classification and String Matching

    … algorithms and energy efficient high throughput hardware accelerators that implement packet classification and fixed string matching. These computationally heavy and memory intensive tasks are used by networking equipment to inspect all packets at wire speed. The constant growth in Internet usage …

    dcu Repository record for Energy Efficient Hardware Accelerators for Packet Classification and String Matching (opens in a new tab)

  5. Optimized ray tracing for real time use without hardware accelerators

    … images. However, in recent years graphics hardware has begun including specialized cores to accelerate ray tracing, an alternative rendering solution to rasterization and a type of light transport simulation. Ray tracing is generally the more desirable choice due to its ability to present …

    vt Repository record for Optimized ray tracing for real time use without hardware accelerators (opens in a new tab)

  6. Performance and Energy Efficiency Insights in LLM Inference Across Hardware Accelerators

    … LLM serving often depends on specialised hardware acceleration. While GPUs continue to dominate, domain-specific accelerators like TPUs and dataflow architectures are becoming increasingly compelling alternatives. In this thesis, we provide a comprehensive empirical performance study of …

    uic

  7. New techniques for the Reliability Evaluation of AI-oriented Hardware Accelerators

    L'abstract è presente nell'allegato / the abstract is in the attachment

    poli-torino Repository record for New techniques for the Reliability Evaluation of AI-oriented Hardware Accelerators (opens in a new tab)

  8. Designing Highly-Efficient Hardware Accelerators for Robust and Automatic Deep Learning Technologies

    … which poses a daunting challenge to traditional hardware platforms, such as CPUs/GPUs. The second issue lying in the conventional DNN applications is the laboring-intensive design period. The actual architecture design of a DNN model demands significant amount of efforts and cycles from machine …

    houston Repository record for Designing Highly-Efficient Hardware Accelerators for Robust and Automatic Deep Learning Technologies (opens in a new tab)

  9. Hardware Accelerator Generation Framework for Cryptographic Primitives

    … for secure communication; thus, dedicated hardware accelerators are required in resource and latency-constrained environments. High-Level Synthesis (HLS) generates hardware from high-level implementations in languages like C, enabling the rapid prototyping and evaluation of designs, leading …

    gatech Repository record for Hardware Accelerator Generation Framework for Cryptographic Primitives (opens in a new tab)

  10. GePSeA: A General-Purpose Software Acceleration Framework for Lightweight Task Offloading

    Hardware-acceleration techniques continue to be used to boost the performance of scientific codes. To do so, software developers identify portions of these codes that are amenable for offloading and map them to hardware accelerators. However, offloading such tasks to specialized hardware

    vt Repository record for GePSeA: A General-Purpose Software Acceleration Framework for Lightweight Task Offloading (opens in a new tab)

  11. FPGA-Roofline: An Insightful Model for FGPA-based Hardware Acceleration in Modern Embedded Systems

    … applications. This trend has led to emergence of hardware accelerators in embedded systems. While the processing power of dedicated hardware modules seems appealing, they require significant effort of development and integration to gain performance benefit. Thus, it is prudent to investigate and …

    vt Repository record for FPGA-Roofline: An Insightful Model for FGPA-based Hardware Acceleration in Modern Embedded Systems (opens in a new tab)

  12. An Energy and Area Estimation Plugin for Accelerator Architecture Simulation

    Development of domain-specific hardware accelerators has been an important focus for high performance computing research in recent years, enabling significant gains in a variety of practical applications. Of particular interest is accelerator design for applications involving sparse data. Such …

    mit Repository record for An Energy and Area Estimation Plugin for Accelerator Architecture Simulation (opens in a new tab)

  13. Parallel algorithms for placement and routing in VLSI design

    … of heuristic algorithms, special purpose hardware accelerators, or parallel algorithms for the numerous design tasks to decrease the time required for solution. In this thesis, we propose two new parallel algorithms for two VLSl synthesis tasks, standard cell placement and global routing.

    uiuc Repository record for Parallel algorithms for placement and routing in VLSI design (opens in a new tab)

  14. Heterogeneous multiprocessor pipeline design for H.264 video encoder

    … (ASIPs) with local and shared memories, and hardware accelerators (in the form of custom instructions). Our platform can be configured to use a particular number of ASIPs (slices/Group of Macroblocks per video frame) for a specific video resolution at design-time. The MPSoC architecture is …

    unsw Repository record for Heterogeneous multiprocessor pipeline design for H.264 video encoder (opens in a new tab)

  15. A Reconfigurable FPGA Overlay Architecture for Matrix-Matrix Multiplication

    … language has inspired many attempts to develop hardware accelerators for matrix-matrix multiplication. Both application-specific integrated circuits (ASICs), and field-programmable arrays (FPGAs) are used for this purpose. However, a trade-off between the two platforms is that ASICs provide …

    wustl Repository record for A Reconfigurable FPGA Overlay Architecture for Matrix-Matrix Multiplication (opens in a new tab)

  16. Branch Prediction For Network Processors

    … computation intensive tasks to dedicated hardware logic or through increased parallelism. While parallelism retains flexibility, challenges such as load-balancing limit its scope. On the other hand, hardware offloading allows complex algorithms to be implemented at high speed but sacrifice …

    dcu Repository record for Branch Prediction For Network Processors (opens in a new tab)

Page 1 of 3