Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 5 of 5 for “"TensorRT"”.
-
TensorRT inference performance study in MLModelScope
… and deploying the well-trained models, NVIDIA TensorRT is the leading framework that is exclusively developed for inference. It allows the developer to optimize the model to facilitate high-performance inference. While it has been shown extensively that TensorRT can significantly boost the …
-
Deep learning methods for real-time corneal and needle segmentation in volumetric OCT scans
… in real-time when deployed using the NVIDIA TensorRT framework.
-
Time-Optimal Re-planning of Quadrotor Trajectories
… a compiled real time inference library (NVIDIA TensorRT). The optimized planner is demonstrated to provide a 14.84 times increase in throughput and over 95% reduction in latency. The increase in throughput can be translated to better efficiency, and the reduction in latency is critical for …
-
Co-Designing Efficient Systems and Algorithms for Sparse and Quantized Deep Learning Computing
… quantization, enhancing the throughput of NVIDIA TensorRT-LLM by 1.2-2.4× on A100 GPUs. Finally, we introduce HART, an efficient autoregressive image generation method that achieves 4.5-7.7× higher throughput compared to diffusion models while maintaining visual quality. HART achieves this …
-
Design and optimization of an embedded machine learning image processing system with a Linux SOC for large-scale livestock tallying
… and YOLOv11 and their size variants. NVIDIA’s TensorRT acceleration framework is used to optimise these custom-trained architectures for FP32, FP16, and INT8 quantization formats, enabling further performance assessment on the selected NVIDIA Jetson Orin Nano platform. The YOLOv9’s large model …