Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 24 for “"Multi-GPU"”.

  1. Fluid Dynamics Simulations on Multi-GPU Systems

    … design, implementation and testing of the multi-GPU version of two fluid flow simulation models, focusing on the cellular automaton MAGFLOW lava flow simulator and the GPU-SPH model for Navier-Stokes. In both cases, a spatial subdivision of the domain is performed, with a minimal overlap to …

    catania Repository record for Fluid Dynamics Simulations on Multi-GPU Systems (opens in a new tab)

  2. Multi-GPU Load Balancing for Simulation and Rendering

    GPU computing can significantly improve performance by taking advantage of massive parallelism of GPUs for data parallel applications. Computation in visualization applications is suitable for parallelization on the GPU, which can improve performance and interactivity in these applications. If used …

    vt Repository record for Multi-GPU Load Balancing for Simulation and Rendering (opens in a new tab)

  3. Multi-GPU Based Lattice Boltzmann Flow Simulations in Porous Media

    … to simulate fluid flows in porous media at multiple length scales. The LB simulation is typically resource intensive due to its computational complexity and hence faces great numerical challenges in extremely large-scale computation. In this dissertation, I propose a multi-GPU solution to …

    houston Repository record for Multi-GPU Based Lattice Boltzmann Flow Simulations in Porous Media (opens in a new tab)

  4. A Multi-GPU Compute Solution for Optimized Genomic Selection Analysis

    … optimization presented in this thesis utilizes GPU computing to exploit the data-level parallelism within each of these iterations. In addition, it allows for the efficient management of memory, the pipelining of CUDA kernels, and the use of multiple GPUs. The optimizations presented show …

    calpoly Repository record for A Multi-GPU Compute Solution for Optimized Genomic Selection Analysis (opens in a new tab)

  5. Multi-GPU Accelerated High-Fidelity Simulations of Beam-Beam Effects in Particle Colliders

    … electron-ion colliders using a cluster of NVIDIA GPUs. The parallel implementation is optimized to minimize the communication overhead and the performance scales near linearly with number of GPUs. Further, the new code enables tracking and collision of the beams for millions of turns, thereby …

    odu Repository record for Multi-GPU Accelerated High-Fidelity Simulations of Beam-Beam Effects in Particle Colliders (opens in a new tab)

  6. CHAOS: A multi-GPU PIC-DSMC solver for modeling gas and plasma flows

    … simulations of these systems that evolve over multiple length- and time-scales is computationally expensive. Until recently, approximations were used to keep computational costs tenable, which in turn, increased the uncertainty in predictions and offered limited insights into the micro-scale …

    uiuc Repository record for CHAOS: A multi-GPU PIC-DSMC solver for modeling gas and plasma flows (opens in a new tab)

  7. OpenMP-CUDA implementation of the moment method and multilevel fast multipole algorithm on multi-GPU computing systems

    … this thesis, the method of moments (MoM) and the multilevel fast multipole algorithm (MLFMA) are implemented for GPU computation based on the hybrid OpenMP-CUDA parallel programming model. The resultant algorithms are called the OpenMP-CUDA-MoM and the OpenMP-CUDA-MLFMA, respectively. Both of the …

    uiuc Repository record for OpenMP-CUDA implementation of the moment method and multilevel fast multipole algorithm on multi-GPU computing systems (opens in a new tab)

  8. Multi-level Parallelism with MPI and OpenACC for CFD Applications

    … abstracts the details of implementation on the GPU. Although OpenACC generally limits the performance of the GPU, this model significantly reduces the work required to port an existing code to any accelerator platform, including GPUs. The purpose of this research is twofold: to investigate the …

    vt Repository record for Multi-level Parallelism with MPI and OpenACC for CFD Applications (opens in a new tab)

  9. SemCache: Semantics-Aware Caching for Efficient GPU Offloading

    <p>Graphical Processing Units (GPUs) offer massive, highly-efficient parallelism, making them an attractive target for computation-intensive applications. However, GPUs have a separate memory space which introduces the complexity of manually handling explicit data movements between GPU and CPU …

    purdue-thes Repository record for SemCache: Semantics-Aware Caching for Efficient GPU Offloading (opens in a new tab)

  10. Declarative Analytics on Heterogeneous HPC Systems

    … computing (HPC) powered by extensive use of GPUs. GPGPU's popularity in HPC, due to performance gains and power efficiency, demands redesigning traditional algorithms to exploit GPU parallelism. However, declarative languages, like Datalog, can directly leverage these advancements due to …

    uic

  11. Decentralized Baseband Processing for Massive MU-MIMO Systems

    … high spectral efficiency in realistic massive multi-user (MU) multiple-input multiple-output (MIMO) wireless systems requires computationally-complex algorithms for data detection in the uplink (users transmit to base-station) and beamforming in the downlink (base-station transmits to users). …

    rice Repository record for Decentralized Baseband Processing for Massive MU-MIMO Systems (opens in a new tab)

  12. Optimizations to a massively parallel database and support of a shared scan architecture

    … MapD, a database server which uses a hybrid of multi-CPU/multi-GPU architecture for query execution and analysis. We tackle the challenge of partitioning the data across multiple nodes with many CPUs and GPUs by means of an indexing framework. We implement a QuadTree spatial partitioning scheme …

    mit Repository record for Optimizations to a massively parallel database and support of a shared scan architecture (opens in a new tab)

  13. Implementation and performance evaluation of a GPU particle-in-cell code

    … (PIC) code on a graphical processing unit (GPU) using NVIDA's Compute Unified Architecture (CUDA). The massively parallel nature of computing on a GPU nessecitated the development of new methods for various steps of the PIC method. I investigated different algorithms and data structures used …

    mit Repository record for Implementation and performance evaluation of a GPU particle-in-cell code (opens in a new tab)

  14. Performance and Energy Efficiency Insights in LLM Inference Across Hardware Accelerators

    … on specialised hardware acceleration. While GPUs continue to dominate, domain-specific accelerators like TPUs and dataflow architectures are becoming increasingly compelling alternatives. In this thesis, we provide a comprehensive empirical performance study of six datacenter-grade GPUs from …

    uic

  15. Acceleration of Eulerian Multi-Material Methods on Highly Parallel Compute Architectures

    … compute architectures, with a focus on multiple graphics processing units (GPUs), as well as to develop algorithms for the numerical simulation of multiple interacting materials on modern, massively parallel, computer hardware. The Ripple framework is applicable to a wide range of HPC …

    cambridge Repository record for Acceleration of Eulerian Multi-Material Methods on Highly Parallel Compute Architectures (opens in a new tab)

  16. Acceleration Techniques for Industrial Large Eddy Simulation with High-Order Methods on CPU-GPU Clusters

    … acceleration techniques are investigated: the p-multigrid algorithm and Mach number preconditioning. The Weiss and Smith low Mach number preconditioner is used together with the p-multigrid method, and the third order explicit Runge-Kutta (RK3) scheme is considered as the smoother to reduce …

    ku Repository record for Acceleration Techniques for Industrial Large Eddy Simulation with High-Order Methods on CPU-GPU Clusters (opens in a new tab)

  17. Argon bubble transport and capture in continuous casting with an external magnetic field using GPU-based large eddy simulations

    … of the molten steel flow. A general purpose multi-GPU Navier-Stokes solver, CUFLOW, is developed. A Coherent-Structure Smagorinsky LES model is implemented to model the turbulent flow. A two-way coupled Lagrangian particle tracking model is added to track the motion of argon bubbles. A …

    uiuc Repository record for Argon bubble transport and capture in continuous casting with an external magnetic field using GPU-based large eddy simulations (opens in a new tab)

  18. HIGH PERFORMANCE AGENT-BASED MODELS WITH REAL-TIME IN SITU VISUALIZATION OF INFLAMMATORY AND HEALING RESPONSES IN INJURED VOCAL FOLDS

    The introduction of clusters of multi-core and many-core processors has played a major role in recent advances in tackling a wide range of new challenging applications and in enabling new frontiers in BigData. However, as the computing power increases, the programming complexity to take optimal …

    maryland Repository record for HIGH PERFORMANCE AGENT-BASED MODELS WITH REAL-TIME IN SITU VISUALIZATION OF INFLAMMATORY AND HEALING RESPONSES IN INJURED VOCAL FOLDS (opens in a new tab)

Page 1 of 2