Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 1 of 1 for “"quantization error measurement (QEM)"”.
-
Comparing the Performance of Small Word-Size Floating-Point Numerics to Fixed-Point Numerics in Neural Networks
… it suffers from limited dynamic range and quantization inflexibility. This thesis introduces an alternative approach—Adaptive Precision Training (APT)—which leverages reduced-precision floating-point formats (FP8, FP12, FP16) for dynamic, layer-wise quantization during training. APT …